GOAT Plan

The Command Code GOAT plan is the best low-cost coding plan on the market today: unlimited coding on 30+ top open and closed models for $10/month. Your $10 buys $70 of credits - a 7x multiplier, the highest of any $10 coding plan. With deals, that stretches beyond $100: more than 10x what you pay.

See every plan side by side on Pricing & Limits, or pick a plan from Studio > Billing.

At Command Code, we're incessantly curious about 1) open models, 2) how to get the best value out of them, and 3) how to make that value accessible to everyone. The GOAT plan is our answer to all three.

While almost every other coding agent was built to serve closed models, we built Command Code to be the best coding agent harness for open models.

We started with the $1 Go plan. The sheer audacity of that plan was a hit, but it was also a bit of a tease: a great way to get started, not enough usage to last the month.

The GOAT plan fixes that - it's the best value in the coding market today, with enough credits to do something meaningful with open models.

Coding with open models comes with hard problems - cost, reliability, safety, and the sheer complexity of running them well. We set out to solve all of it, so you can pick an open model and just code.

From getting DeepSeek to beat Opus, to building infra with leading ~98% cache hit rates, to fixing AI slop with the built-in /design skill - Command Code has made its way to the top of the open model coding world, everywhere. The GOAT plan is the next step in that journey: open models, accessible to everyone.

Open models are already beating frontier closed models at coding. How rich is this moment if you think of it.

With the v1 release, which is a full rewrite of this 6-year-old codebase, it's now arguably one of the best, most mature coding agent harnesses. Check out the new Mods API - you can build anything you can imagine.

We can't wait to see what you build with Command Code.

We're also open sourcing Command Code later this month.

Let's connect on @CommandCodeAI and in our Discord community.

Every model below is included on the GOAT plan - switch between them any time with /model in interactive mode. Rates are per 1M tokens; the full catalog, including the premium models that need Pro or Max, lives on commandcode.ai/models.

Caps
DeepSeek V4 Flash Vision (exp)Off-peak shown (17h/day) · peak $0.44 / $1.32 01–04 & 06–10 UTC
1Mnot yet scored$0.22$0.66$0.01
1Mnot yet scoredFreeFreeFree
1M59.580$1.40$4.40$0.26
262K52.0$0.40$3.00$0.04
DeepSeek V4 Pro (latest)Off-peak shown (17h/day) · peak $1.32 / $3.96 01–04 & 06–10 UTC
1M53.275$0.66$1.98$0.022
Gemini 3.7 Flash-50%Ends December 31, 2026
1M56.0339$1.50$0.75$7.50$3.75$0.15$0.075$0.08334$0.04167
500K60.966$2.00$6.00$0.50
1M56.8$1.25$4.25$0.15
1M56.8$0.10$0.20$0.002
1M58.147$2.00$6.00$0.25$2.50
DeepSeek V4 Flash (latest)Off-peak shown (17h/day) · peak $0.44 / $1.32 01–04 & 06–10 UTC
1M52115$0.22$0.66$0.007
1M41.261$0.50$1.20$0.10
1Mnot yet scored$0.03$0.13$0.006$0.038
256Knot yet scoredFreeFreeFree
256K42.370$1.00$4.05$0.17
1M59.739$3.00$15.00$0.30
1.1M52.3165$0.20$1.20$0.02$0.25
1.1M60.969$5.00$30.00$0.50$6.25
500K55.851$2.00$6.00$0.50
262K42.275$0.14$0.58$0.035
1Mnot yet scored$3.00$10.25$0.50
1M52.669$1.40$4.40$0.26
262Knot yet scored$1.90$8.00$0.38
256K43.038$0.95$4.00$0.19
1M38.3136$0.60$2.40$0.12
1M45.4104$0.60$0.30$2.40$1.20$0.12$0.06
1M39.456$0.40$1.60$0.08$0.50
256K30.9120$0.20$1.15$0.04
1M38.062$0.80$0.14$4.00$0.28$0.16$0.0028
1M42.958$2.00$0.435$6.00$0.87$0.40$0.0036
1M46.7$2.50$7.50$0.50$3.13
1M26.5$0.10$0.30$0.02
200K41.028$1.40$4.40$0.26
200K38.9$0.30$1.20$0.06
200K41.1$1.30$7.80$0.26$1.63
200K40.5$0.50$3.00$0.10
256K45.1$0.95$4.00$0.16
200K40.6$1.00$3.20$0.20
256K36.0$0.60$3.00$0.10
200K34.5$0.30$1.20$0.03
40/58 Models · Per 1M tokens · USD
Text Vision Reasoning

The GOAT plan includes the following usage limits:

  • 5-hour limit - $14 of usage
  • Weekly limit - $35 of usage
  • Monthly limit - $70 of usage

How far each window goes depends on the model - cheaper models allow more requests, pricier models fewer. We have a usage estimation calculator. Estimated request counts per limit window:

ModelRequests / 5 hoursRequests / weekRequests / month
GPT-5.6 Sol4141,0402,070
DeepSeek V4 Flash (latest)18,20045,60091,200
GLM-5.29472,3704,740
GPT-5.6 Luna2,9607,40014,800
Tencent Hy37,08017,70035,400
Qwen 3.8 27B4,79012,00024,000
Qwen 3.8 Max2616541,310
Qwen 3.7 Max2325791,160
Qwen 3.7 Plus1,4203,5607,110
Qwen 3.6 Plus1,1002,7505,500
MiniMax M32,7706,93013,900
Kimi K2.7 Code1,0802,7105,420
MiMo V2.519,50048,70097,400
DeepSeek V4 Pro (latest)1,9804,9409,880
MiMo V2.5 Pro5,70014,20028,500
DeepSeek V4 Flash Vision (exp)4,95012,40024,800
GLM-5.32716771,350
Muse Spark 1.24281,0702,140
Muse Spark 1.2 Contributor18,20045,50090,900
Kimi K3196490980
Kimi K2.7 Code HighSpeed181452904
Grok 4.5144360719
Grok 4.6144360719
Gemini 3.7 Flash1,5703,9207,840
GLM-5.2 Fast138346691
Inkling3969891,980
Inkling Small7091,7703,550
Step 3.7 Flash1,6704,1808,370
Step 3.5 Flash3,5108,77017,500
Nemotron 3 Ultra5751,4402,870

Model the math yourself with the calculator on Pricing & Limits - it knows every model's GOAT allowance.

The estimates assume a typical agent request - ~800 fresh input tokens, ~50,000 cache-read tokens, and ~125-200 output tokens depending on the model family - at the prices below per 1M tokens. Most of an agent's volume is cache reads, which is why cache hit rates matter so much. Browse every model on Available Models, with per-token rates on Pricing & Limits.

Every model on the GOAT plan is listed below:

ModelInputOutputCache ReadCache WriteMonthly credits
$5.00$30.00$0.50$6.25$70
GLM-5.2$1.40$4.40$0.26-$70
Tencent Hy3$0.14$0.58$0.035-$70
Qwen 3.8 27B$0.40$3.00$0.04-$70
$0.22$0.66$0.007-$60
Kimi K2.7 Code$0.95$4.00$0.19-$60
$0.30$1.20$0.06-$47
Qwen 3.7 Max$2.50$7.50$0.50$3.13$33
$0.40$1.60$0.08$0.50$33
$0.50$3.00$0.10-$33
MiMo V2.5$0.14$0.28$0.0028-$30
$0.20$1.20$0.02$0.25$20
Qwen 3.8 Max$2.00$6.00$0.25$2.50$20
$0.66$1.98$0.022-$20
MiMo V2.5 Pro$0.435$0.87$0.0036-$20

New models start at 2x credits i.e. $20 credits on the $10/mo GOAT plan until we negotiate better deals. That's still 2x your money at public API prices, and the allowance updates automatically the moment a deal starts - no codes, no toggles.

ModelInputOutputCache ReadCache WriteMonthly credits
$0.22$0.66$0.01-$20
GLM-5.3$1.40$4.40$0.26-$20
Muse Spark 1.2$1.25$4.25$0.15-$20
Muse Spark 1.2 Contributor$0.10$0.20$0.002-$20
Kimi K3$3.00$15.00$0.30-$20
Kimi K2.7 Code HighSpeed$1.90$8.00$0.38-$20
Grok 4.5$2.00$6.00$0.50-$20
$2.00$6.00$0.50-$20
Gemini 3.7 Flash$0.75$3.75$0.075$0.04167$40
GLM-5.2 Fast$3.00$10.25$0.50-$20
Inkling$1.00$4.05$0.17-$20
Inkling Small$0.50$1.20$0.10-$20
Step 3.7 Flash$0.20$1.15$0.04-$20
Step 3.5 Flash$0.10$0.30$0.02-$20
Nemotron 3 Ultra$0.60$2.40$0.12-$20

Speed variants are separate models with their own credits: GLM-5.2 Fast and Kimi K2.7 Code HighSpeed carry the standard 2x credits ($20 on the $10 GOAT plan), not their base model's boosted allowance.

We're constantly negotiating better terms with providers, so the allowances above can change at any time. When a deal starts, your allowance updates by itself.

You pay $10 and code with up to $70 of per-model credit allowances. We benchmarked the GOAT plan to be the best value on the market. This is possible through phenomenal harness and inference engineering. We work directly with model teams and providers to negotiate capacity for the models we recommend, and we run our infrastructure at leading ~95-98% cache hit rates.

Most of a coding agent's volume is cache reads, so serving the cache well makes every request dramatically cheaper - and those savings come back to you as bigger credits.

Deals with the best pricing apply automatically. Many providers and coding agents charge up to 400% over model list prices, while we regularly offer deals matching - and sometimes beating - the labs' own API prices. We run our own infrastructure and have deep experience running open models at scale, which lets us offer better pricing than most providers.

That's also why credits vary by model. Where we've negotiated capacity, credits are boosted - up to the full $70 on models like GPT-5.6 Sol, GLM-5.2 and Tencent Hy3, with boosted credits across the Qwen 3.6/3.7 family. New models, and models whose public pricing already discounted via deals, carry the 2x credits ($20 credits on the $10/mo GOAT plan) and move up as we land better terms. We regularly move capacity to the latest models that perform better.

Partner with us

If you're a model lab or provider and want to partner with Command Code, reach out to us to get your model in front of over 100K developers.

  • Every major open model: Switch any time with /model in interactive mode. Browse them all on Available Models, with per-token rates on Pricing & Limits.
  • Usage limits: $14 of usage in any 5 hours, $35 in any 7 days - see Usage Limits.
  • Extra credits any time: Buy pay-as-you-go credits at model cost - they unlock every model (including premium), roll over, and never expire.
  • Global availability: Open-source models run on infrastructure in the US, EU, and Singapore for reliable access worldwide.
  • Zero data retention available: Most models are ZDR by default - agreements are renewed monthly and can take time for brand-new models. You can also enforce ZDR on every request (which can change model prices based on the provider, as explained in Pricing & Limits).

Your budget refreshes at the start of every billing cycle, with every model discount already baked into the allowances - no codes, no toggles.

Buy extra credits at model cost any time: they roll over, never expire, and bill at the model's regular rate, since the bigger allowances apply to your monthly budget only.

Past a limit, requests fall back to those credits - and without them, paid models pause until the window or cycle resets while the free models keep working.

  1. Install: npm i -g command-code
  2. Sign in and start coding with cmd (macOS/Linux) or cmdc (Windows)
  3. Pick the GOAT plan from Studio > Billing or the pricing page

Need premium models? Compare the Pro and Max plans, or use the pay-as-you-go Provider API. Full comparison and FAQs live on Pricing & Limits.

The GOAT plan has API access. You call the Provider API endpoints with your Command Code API key, and usage is metered against your GOAT credits and the models included above.

1

Subscribe

Pick the GOAT plan from Studio > Billing or the pricing page. Every plan except the Go plan has API access.

2

Create an API key

The same key authenticates the CLI and the API. Create one from your API keys page in Studio.

3

Call the API

Point any OpenAI or Anthropic compatible client at https://api.commandcode.ai/provider/v1 and send your first request.

Provider API Example Request

curl https://api.commandcode.ai/provider/v1/chat/completions \ -H "Authorization: Bearer <CMD_API_KEY>" \ -H "Content-Type: application/json" \ -d '{ "model": "deepseek/deepseek-v4-flash", "messages": [{"role": "user", "content": "Write a haiku about race conditions."}] }'

Read the Provider API quickstart for more details.

We still recommend the Command Code CLI: harness engineering makes the same credits go further - the Read tool keeps ~25 billion junk tokens a month out of your context, tool call repairs are free on every plan, and ~98% cache hit rates mean most of every request bills as cheap cache reads.

What models are available in GOAT plan?

GOAT plan offers open and closed models. Switch any time with /model in interactive mode. Browse them all on Available Models, with per-token rates on Pricing & Limits.

What are the usage limits on the GOAT plan?

$14 of usage in any 5 hours, $35 in any 7 days, and $70 per month. Limits reset with each window and with your billing cycle - see Usage Limits.

Can I buy extra credits?

Yes, any time. Buy pay-as-you-go credits at model cost - they unlock every model (including premium), roll over, and never expire.

Can I use the GOAT plan via API?

Yes. Every plan except the Go plan has API access - you use the same Provider API the Provider plan uses (OpenAI Chat Completions and Anthropic Messages endpoints, one key for the CLI and the API), so you can integrate GOAT with any other agent. We still recommend the Command Code CLI, because phenomenal harness engineering makes the same credits go further: our Read tool keeps ~25 billion junk tokens a month out of your context, tool call repairs are free on every plan, and we run ~98% cache hit rates so most of every request bills as cheap cache reads.

Is the GOAT plan available worldwide?

Yes. Open-source models run on infrastructure in the US, EU, and Singapore for reliable access worldwide.

Can I enforce zero data retention?

Yes. Most models are ZDR by default - agreements are renewed monthly and can take time for brand-new models. You can also enforce ZDR on every request (which can change model prices based on the provider, as explained in Pricing & Limits). Run the CLI with CMD_ZDR=1 (for example, CMD_ZDR=1 cmd) to enforce zero data retention and no prompt training on every request.