Use /plan-review mode to comment on the plan, revise, and approve it when it’s ready.
CommandCode.ai
@commandcode.ai
Command Code with taste; the first coding agent that observes how you write code and adapts to your preferences over time with meta neuro-symbolic AI `taste-1`.
Tested FlappyBench with /design on Hy4 Preview, Kimi K3, and GLM 5.3. 🔹 Hy4 Preview ($0.0480): Most distinct UI and gameplay 🔹 Kimi K3 ($0.0740): Hardest gameplay and the most expensive 🔹 GLM 5.3 ($0.0184): Smooth gameplay and the cheapest of the three
Tencent Hy4 Preview is live in Command Code 🐐 770B mode, 49B active, 1M context. `cmd update` available in GOAT and above plans. Our early internal benchmarks show strong perf yet cheaper per task compared to GLM 5.3 Flash. Y'all share how it goes.
Type /resume and get back to your previous conversation • $ cmd --resume opens a picker for earlier sessions • $ cmd --continue resumes your latest session
Qwen 3.8 Flash is now available in Command Code with 2x usage limits on the GOAT plan. 🐐
GLM 5.3 Flash aka Ox Alpha is now live in Command Code 4x usage in GOAT plan $40 credits. ~24K reqs. ~1.2B tokens Available across all plans and API 🐐
Minimax M3 and M2.7 are now free in Command Code. next 10 days. all subs. all plans. 🐐
Paste images as your chat input. • Take a screenshot • Paste it into the chat • Describe what’s wrong • Watch Command Code fix the issue Let your screenshots be the context.
Alipay is now supported in Command Code. More ways to pay and be a GOAT.
Tested FlappyBench on Qwen 3.8 27B, DeepSeek V4 Pro 0813, and Gemini 3.7 Flash All scored 9/10, it comes down to cost, quality, and iterations 🔹 DSV4 ($0.0036): Low cost, solid, but movement need work 🔹 Qwen ($0.0068): One-shot, playable, with UI finish 🔹 Gemini ($0.0083): Needed 2–3 iterations
Text-only models now have vision in Command Code. • A vision model reads the attached image • Your main model gets the context in the same thread • /config lets you choose the vision model Stop describing screenshots. Just attach them.
Qwen 3.8 27B is now live in Command Code. Beats Opus 4.6 Max, GPT 5.6 Luna Max on Artificial Analysis Intelligence Index. Best for coding and long agent tasks. Multimodal with reasoning toggled on/off Available on all plans & API 🐐
GLM-5.3 is already live in Command Code. Same price as GLM-5.2, ~50% better on Zai's internal Code Bench. Still 5–10x cheaper than the frontier closed models (per 1M output tokens, list price today). Plans: $1 Go ($10 credits) · $10 GOAT ($20 credits) 🐐
Launching `a11y` mode in `/design` Run `/design a11y` to catch: • Tap targets that are too small • Focus management problems • Broken keyboard navigation • ARIA and accessible name gaps • Missing or incorrect form labels and errors Run `cmd update`, then try `/design a11y` in your project
GPT 5.6 Sol is now in Command Code GOAT 🐐 Best coding plan with best GPT model. $70 credits. $10/mo GOAT plan gets you: GPT Sol has $70 credits ~105M tokens ~2.1K reqs We've also achieved 99.43% cache hit ratio. Available on GOAT + above. Limited time deal. 🐐
Writing tests is now the easiest part of your workflow. Pick a file, and Command Code will: - Analyze the code - Generate the right test cases - Run the test suite - Resolve failures - Re-run until all tests pass Try it on a file that hasn’t been tested yet.
Tested FlappyBench with GLM 5.3, Fable 5 and GPT-5.6 Sol 3 models. Same prompt with /design command. Scored on features, UX/UI, and cost. 🔹 Fable 5 → 9.5/10 · $0.420 🔹 GLM 5.3 → 9/10 · $0.018 🔹 GPT-5.6 Sol → 9/10 · $0.150 Fable 5 wins on quality, but costs 23x more than GLM 5.3
GLM 5.3 is live in Command Code (2x usage) The best deal: 2x usage, permanently. 🐐 $10 GOAT plan has $20 worth of GLM 5.3 usage. That's ~70M tokens & ~1.3K requests. • Text-reasoning. 1M context • IN $1.4/M · OUT $4.4/M · Cache $0.26/M • Available on all plans & API
Gemini 3.7 Flash is now live in Command Code at 50% OFF. - It's fast with frontier level intelligence - $0.75 input/$3.75 output per 1M tokens - Available on GOAT, Max and Team plans Starts at $10/month $ npm i -g command-code
Pull requests are easier now. Just ask to open a PR, It will: • Read your changes • Draft the description and test plan • Commit and push your branch • Open a well-documented PR Try it with our $10/mo GOAT plan.
Tested FlappyBench with Grok 4.6, GPT-5.6 Sol, and Opus 5. 3 frontier models. Same prompt with /design command. Scored on gameplay features, UX/UI, and cost. 🔹 Grok 4.6 → 9.5/10 · $0.095 🔹 GPT-5.6 Sol → 9/10 · $0.150 🔹 Opus 5 → 9/10 · $0.253 Grok 4.6 gave best output with lowest cost.
Tested New DSV4 Pro 0813, Kimi K3, and GLM 5.2 with our FlappyBench 3 models, same prompt with the /design command. Reviewed gameplay features, UX/UI, and cost. 🔹 DSV4 Pro 0813 → 8/10 · $0.0005 🔹 Kimi K3 → 9.5/10 · $0.0740 🔹 GLM 5.2 → 9/10 · $0.0480
Code reviews should help you ship faster, not block you. Run /review [PR number] and get your code reviews directly in the terminal • scans the diff • flags risky changes • finds missing tests • posts a ready-to-ship PR review Try now with our $10/mo GOAT plan.
Browse through plans with /plans Browse, review, and annotate saved plans. Select plans from current session vs all. Search across all plans. Keep up with all your planning.
We benchmarked Muse Spark 1.2, Kimi K3, and GLM 5.2 with Flappy Bench 3 models, same prompt with the /design command. Reviewed gameplay features, UX/UI, and cost. 🔹 Kimi K3 → 9.5/10 · $0.0740 🔹 GLM 5.2 → 9/10 · $0.0480 🔹 Muse Spark 1.2 → 8/10 · $0.0187
Muse Spark 1.2 by Meta now live in Command Code 🔹 Muse Spark 1.2 • Available in plans: GOAT, Pro, Max • $1.25 input, $4.25 output, $0.15 cache 🔹Muse Spark 1.2 Contributor • Available in plans: Go, GOAT, Pro, Max • Contributor is ~95% cheaper than 1.2 • $0.10 input, $0.20 output, $0.002 cache
Plan mode removes the guesswork. When the change is too big for a quick fix: • Your code stays untouched • Plan it first, then review it • Apply exactly as planned No surprises, just clarity.