[ Web Proxy ]
URL:
Viewing: https://commandcode.ai/docs/plans/max [Back]  [Original]

Max Plans | Command Code - Command Code Docs

Max Plans

The Max plans are the high-volume Command Code tiers for power users running agents all day. Both unlock every model - open-source and premium, including Claude Opus, Claude Fable, and Fugu Ultra - with no restrictions.

See every plan side by side on Pricing & Limits, or pick a plan from Studio > Billing.

PlanPrice/moCredits/moIncluded LLM UsageModels
Max 10$100$150~230K requestsEvery model, no restrictions
Max 20$200$300~370K requestsEvery model, no restrictions

Both include more credits for the same cost as raw API pricing. Credits reset monthly, top up at model cost, roll over, and never expire.

Max 10 gives you a $150 standard model usage limit and a $100 premium model usage limit (Claude Opus and Fable among them). Max 20 gives you $300 standard and $200 premium. See How credits work.

$100/month $150 of credits ~230K requests. For power users running Command Code throughout the workday.

  • Every model, including Claude Opus, Fable, and Fugu Ultra. No model restrictions.
  • $150 standard usage limit, $100 premium usage limit.
  • ~230K requests/month on a balanced mix (~43% standard / ~57% premium).
  • Best for engineers who keep an agent running all day across multiple projects.

$200/month $300 of credits ~370K requests. The top tier - maxed-out daily usage at the highest request volume.

  • Every model, no restrictions, including Claude Opus, Fable, and Fugu Ultra.
  • $300 standard usage limit, $200 premium usage limit.
  • ~370K requests/month on a balanced mix (~43% standard / ~57% premium).
  • Best for the heaviest users and small teams who want the most usage on one account.

Max 10 and Max 20 include all 63 models - the full open-source + premium catalog with no restrictions, including Claude Opus, Claude Fable, and Fugu Ultra. Switch any time with /model. Browse them all on Available Models.

Max 10 and Max 20 include the following usage limits:

Max 10Max 20
5-hour limit$45 of usage$90 of usage
Weekly limit$90 of usage$180 of usage
Monthly limit$150 of usage$300 of usage

Max 10 gives you a $150 standard model usage limit and a $100 premium model usage limit. Max 20 gives you $300 standard and $200 premium. See How credits work.

How far each window goes depends on the model - cheaper models allow more requests, pricier models fewer. We have a usage estimation calculator. Here's what each window buys per model on Max 10. Max 20 doubles every number below:

ModelRequests / 5 hoursRequests / weekRequests / month
Laguna S 2.1FreeFreeFree
MiniMax M3FreeFreeFree
MiniMax M2.7FreeFreeFree
Tencent Hy4 Preview13,80027,50045,900
Tencent Hy322,80045,50075,800
Kimi K32,2104,4107,350
Kimi K2.7 Code4,0708,14013,600
Kimi K2.7 Code HighSpeed2,0304,0706,780
Kimi K2.64,7109,41015,700
Kimi K2.57,40014,80024,700
GLM-5.3 Flash26,50053,10088,500
GLM-5.33,0406,09010,100
GLM-5.23,0406,09010,100
GLM-5.2 Fast1,5603,1105,180
GLM-5.13,0406,09010,100
GLM-53,9907,98013,300
MiniMax M313,30026,50044,200
MiniMax M2.713,30026,50044,200
MiniMax M2.523,80047,60079,400
DeepSeek V4 Pro (latest)22,20044,50074,100
DeepSeek V4 Flash (latest)68,400137,000228,000
DeepSeek V4 Flash Vision (exp)68,400137,000228,000
Qwen 3.8 Max2,9405,8809,800
Qwen 3.8 27B15,40030,80051,400
Qwen 3.6 Max Preview2,8805,7709,620
Qwen 3.6 Plus7,50015,00025,000
Qwen 3.7 Max1,5803,1605,260
Qwen 3.7 Plus9,70019,40032,300
Qwen 3.8 Flash44,00088,100147,000
Qwen 3.7 Flash129,000257,000429,000
Step 3.7 Flash18,80037,70062,800
Step 3.5 Flash39,50078,900132,000
MiMo V2.5 Pro64,100128,000214,000
MiMo V2.5146,000292,000487,000
Nemotron 3 Ultra6,47012,90021,600
Claude Fable 54488961,490
Claude Opus 58961,7902,990
Claude Opus 4.88961,7902,990
Claude Opus 4.78961,7902,990
Claude Opus 4.68961,7902,990
Claude Sonnet 52,2404,4807,460
Claude Sonnet 4.61,4902,9904,980
Claude Haiku 4.54,4808,96014,900
GPT-5.6 Sol1,3302,6604,440
GPT-5.6 Terra2,2204,4407,400
GPT-5.6 Luna33,30066,600111,000
GPT-5.58881,7802,960
GPT-5.41,7803,5505,920
GPT-5.4 Mini5,92011,80019,700
GPT-5.3 Codex1,0802,1503,590
Gemini 3.7 Flash8,82017,60029,400
Gemini 3.6 Flash2,9405,8809,800
Gemini 3.5 Flash2,8605,7109,520
Gemini 3.5 Flash Lite13,40026,80044,600
Gemini 3.1 Flash Lite15,00030,00050,000
Fugu Ultra8571,7102,860
Muse Spark 1.24,8109,63016,000
Muse Spark 1.2 Contributor205,000409,000682,000
Muse Spark 1.13,2106,42010,700
Grok 4.61,6203,2405,400
Grok 4.51,6203,2405,400
Inkling4,4508,90014,800
Inkling Small7,98016,00026,600

The estimates assume a typical agent request - ~7001K fresh input tokens, ~42K56K cache-read tokens, and ~125200 output tokens depending on the model family. Long conversations cost more, since every turn re-reads the full context - use /clear or a new session for unrelated tasks. The Usage page in Studio always reflects the actual price charged per request.

Every model on Max is listed below, with its monthly credit limit on each plan:

ModelInputOutputCache ReadCache WriteMax 10 creditsMax 20 credits
Laguna S 2.1$0.00$0.00$0.00-FreeFree
MiniMax M3$0.00$0.00$0.00-FreeFree
MiniMax M2.7$0.00$0.00$0.00-FreeFree
Tencent Hy4 Preview$0.834$2.501$0.042-$150$300
Tencent Hy3$0.14$0.58$0.035-$150$300
Kimi K3$3.00$15.00$0.30-$150$300
Kimi K2.7 Code$0.95$4.00$0.19-$150$300
Kimi K2.7 Code HighSpeed$1.90$8.00$0.38-$150$300
Kimi K2.6$0.95$4.00$0.16-$150$300
Kimi K2.5$0.60$3.00$0.10-$150$300
GLM-5.3 Flash$0.15$0.50$0.03-$150$300
GLM-5.3$1.40$4.40$0.26-$150$300
GLM-5.2$1.40$4.40$0.26-$150$300
GLM-5.2 Fast$3.00$10.25$0.50-$150$300
GLM-5.1$1.40$4.40$0.26-$150$300
GLM-5$1.00$3.20$0.20-$150$300
MiniMax M3$0.30$1.20$0.06-$150$300
MiniMax M2.7$0.30$1.20$0.06-$150$300
MiniMax M2.5$0.30$1.20$0.03-$150$300
DeepSeek V4 Pro (latest)$0.66$1.98$0.022-$150$300
DeepSeek V4 Flash (latest)$0.22$0.66$0.007-$150$300
DeepSeek V4 Flash Vision (exp)$0.22$0.66$0.007-$150$300
Qwen 3.8 Max$2.00$6.00$0.25$2.50$150$300
Qwen 3.8 27B$0.40$3.00$0.04-$150$300
Qwen 3.6 Max Preview$1.30$7.80$0.26$1.63$150$300
Qwen 3.6 Plus$0.50$3.00$0.10-$150$300
Qwen 3.7 Max$2.50$7.50$0.50$3.13$150$300
Qwen 3.7 Plus$0.40$1.60$0.08$0.50$150$300
Qwen 3.8 Flash$0.16$0.47$0.016-$150$300
Qwen 3.7 Flash$0.03$0.13$0.006$0.038$150$300
Step 3.7 Flash$0.20$1.15$0.04-$150$300
Step 3.5 Flash$0.10$0.30$0.02-$150$300
MiMo V2.5 Pro$0.435$0.87$0.0036-$150$300
MiMo V2.5$0.14$0.28$0.0028-$150$300
Nemotron 3 Ultra$0.60$2.40$0.12-$150$300
Claude Fable 5$10.00$50.00$1.00$12.50$100$200
Claude Opus 5$5.00$25.00$0.50$6.25$100$200
Claude Opus 4.8$5.00$25.00$0.50$6.25$100$200
Claude Opus 4.7$5.00$25.00$0.50$6.25$100$200
Claude Opus 4.6$5.00$25.00$0.50$6.25$100$200
Claude Sonnet 5$2.00$10.00$0.20$2.50$100$200
Claude Sonnet 4.6$3.00$15.00$0.30$3.75$100$200
Claude Haiku 4.5$1.00$5.00$0.10$1.25$100$200
GPT-5.6 Sol$5.00$30.00$0.50$6.25$150$300
GPT-5.6 Terra$2.00$12.00$0.20$2.50$100$200
GPT-5.6 Luna$0.20$1.20$0.02$0.25$150$300
GPT-5.5$5.00$30.00$0.50-$100$200
GPT-5.4$2.50$15.00$0.25-$100$200
GPT-5.4 Mini$0.75$4.50$0.075-$100$200
GPT-5.3 Codex$2.00$8.00$0.50-$100$200
Gemini 3.7 Flash$0.75$3.75$0.075$0.04167$150$300
Gemini 3.6 Flash$1.50$7.50$0.15-$100$200
Gemini 3.5 Flash$1.50$9.00$0.15-$100$200
Gemini 3.5 Flash Lite$0.30$2.50$0.03-$100$200
Gemini 3.1 Flash Lite$0.25$1.50$0.03-$100$200
Fugu Ultra$5.00$30.00$0.50-$100$200
Muse Spark 1.2$1.25$4.25$0.15-$150$300
Muse Spark 1.2 Contributor$0.10$0.20$0.002-$150$300
Muse Spark 1.1$1.25$4.25$0.15-$100$200
Grok 4.6$2.00$6.00$0.50-$150$300
Grok 4.5$2.00$6.00$0.50-$150$300
Inkling$1.00$4.05$0.17-$150$300
Inkling Small$0.50$1.20$0.10-$150$300

For the full detail on any model's pricing - including peak/off-peak rates where they apply - see Available Models.

On Max 10, you have a $150 standard model usage limit and a $100 premium model usage limit. On Max 20, that's $300 standard and $200 premium.

Billing category decides which limit applies, not branding. GPT-5.6 Sol and Luna, Grok 4.5 and 4.6, Gemini 3.7 Flash, and Muse Spark 1.2 are marketed as premium but bill as standard, so they use the standard limit like other standard models.

A few other things worth knowing:

  • Extra pay-as-you-go credits aren't capped and work on every model.
  • Credits reset at the start of each billing cycle.
  • Top up at model cost any time. Extra credits roll over and never expire.
  • Deals apply to every request, including extra credits, and stack on top of your plan credits - a model at 50% off makes them go twice as far. See Deals for what's running now.

  • Max 10 - for power users running Command Code all day.
  • Max 20 - for maxed-out daily usage at the highest request volume.

Need more than your plan's monthly limit? Buy extra pay-as-you-go credits any time from Studio > Billing, at model cost. They roll over and never expire, or turn on auto top-up to add them automatically when your balance runs low. For raw pay-as-you-go API access with no markup, see the Provider API. Teams and organizations: contact support@commandcode.ai about Team & Enterprise plans.

  1. Install: npm i -g command-code
  2. Sign in and start coding with cmd
  3. Pick Max 10 or Max 20 from Studio > Billing or the pricing page

Full comparison and FAQs live on Pricing & Limits.

The Max plan has API access. You call the Provider API endpoints with your Command Code API key, and usage is metered against your Max credits and the models included above.

1

Subscribe

Pick the Max plan from Studio > Billing or the pricing page. Every plan except the Go plan has API access.

2

Create an API key

The same key authenticates the CLI and the API. Create one from your API keys page in Studio.

3

Call the API

Point any OpenAI or Anthropic compatible client at https://api.commandcode.ai/provider/v1Copy and send your first request. Premium models take the Anthropic Messages shape on /v1/messages, open models take the OpenAI shape on /v1/chat/completions.

Provider API Example Request

curl https://api.commandcode.ai/provider/v1/messages \ -H "Authorization: Bearer <CMD_API_KEY>" \ -H "Content-Type: application/json" \ -d '{ "model": "claude-opus-5", "max_tokens": 1024, "messages": [{"role": "user", "content": "Write a haiku about race conditions."}] }'
Copy

Read the Provider API quickstart for more details.

We still recommend the Command Code CLI: harness engineering makes the same credits go further - the Read tool keeps ~25 billion junk tokens a month out of your context, tool call repairs are free on every plan, and ~98% cache hit rates mean most of every request bills as cheap cache reads.

What models are available on the Max plans?

Every model, with no restrictions - the full open and premium catalog, including Claude Opus, Claude Fable, and Fugu Ultra. Switch any time with /model in interactive mode. Browse them all on Available Models, with per-token rates on Pricing & Limits.

Should I pick Max 10 or Max 20?

Max 10 is $100/month for $150 of credits (~230K requests) - built for engineers who keep an agent running all day. Max 20 is $200/month for $300 of credits (~370K requests) - the top tier for the heaviest users. Both unlock every model; the only difference is how much you get.

What are the usage limits on the Max plans?

On Max 10, $45 of usage in any 5 hours and $90 in any 7 days; on Max 20, $90 and $180. Windows roll from your first request and reset on their own - see Usage Limits. Extra pay-as-you-go credits have no usage limits.

How much can I spend on premium vs. standard models?

Max 10 gives you a $150 standard model usage limit and a $100 premium model usage limit. Max 20 gives you $300 standard and $200 premium. Claude Opus and Fable use the premium limit. Billing category decides which limit applies, not branding: GPT-5.6 Sol and Luna, Grok, Gemini 3.7 Flash, and Muse Spark 1.2 are marketed as premium but bill as standard, so they use the standard limit. The usage calculator reflects this - pick a premium model on a Max plan and it divides the premium limit, not the plan total. Extra pay-as-you-go credits aren't capped and work on every model.

Can I buy extra credits?

Yes, any time. Buy pay-as-you-go credits at model cost - they unlock every model, roll over, never expire, and skip the usage limits entirely.

Can I use the Max plans via API?

Yes. Max has API access - you use the same Provider API the Provider plan uses (OpenAI Chat Completions and Anthropic Messages endpoints, one key for the CLI and the API), so you can integrate Max with any other agent. We still recommend the Command Code CLI, because our harness engineering makes the same credits go further: our Read tool keeps ~25 billion junk tokens a month out of your context, tool call repairs are free on every plan, and we run ~98% cache hit rates so most of every request bills as cheap cache reads.

Are the Max plans available worldwide?

Yes. Open-source models run on infrastructure in the US, EU, and Singapore for reliable access worldwide.

Can I enforce zero data retention?

Yes. Most models are ZDR by default - agreements are renewed monthly and can take time for brand-new models. You can also enforce ZDR on every request (which can change model prices based on the provider, as explained in Pricing & Limits). Run the CLI with CMD_ZDR=1 (for example, CMD_ZDR=1 cmd) to enforce zero data retention and no prompt training on every request.

Can I upgrade, downgrade, or cancel any time?

Yes. Change or cancel your plan from Studio > Billing at any time. Upgrades take effect immediately and start both usage windows fresh; downgrades take effect at your next billing cycle. Extra credits you have bought stay with your account - they roll over and never expire.

Table of ContentsOn This Page

Join our Discord

Chat with us and become part of the Command Code community.

Join Discord

Need help?

Open an issue on GitHub, ask on Discord, or reach out on .

Get help

Web Proxy Viewer  |  New URL  |  Original Page