CommandCode Model Efficiency Report


Verdict

Keep GLM-5.3-Flash as the daily driver. It is already the most cost-efficient model on the CommandCode catalog — by a wide margin, not a marginal one. No switch recommended.


Efficiency Table

Sorted by Artificial Analysis' full Intelligence Index evaluation suite cost per Index point (their apples-to-apples efficiency measure). All prices are list prices per 1M tokens as tracked by Artificial Analysis (first-party serving, Sep 2026).

ModelII[^1]In $/1MOut $/1MBlend $/1M[^2]Suite cost per II pt (USD)tok/sParams (total / active)
GLM-5.3-Flash420.150.500.246.760320B / 18B
DeepSeek V4 Flash350.441.320.6613.5126284B / 13B
DeepSeek V4 Pro361.323.961.9831.2771.6T / 49B
GLM-5.3451.404.402.1555.671753B / 40B
Gemini 3.5 Flash331.509.003.3865.8216n/a
Qwen3.8 Max402.006.003.0079.340n/a
Kimi K3443.0015.006.0083.1422.8T / 104B
GPT-5.5395.0030.0011.25135.890n/a
GLM-5.2391.404.402.15n/a62n/a

[^1]: II = Artificial Analysis Intelligence Index v3.4.3-class score as published on each model page (higher is better; page-stated medians ranged 17–24 at fetch time). [^2]: Blend = 3:1 input:output token mix, a standard workload approximation.

GLM-5.3-Flash is roughly 8x cheaper per quality point than its own flagship (GLM-5.3) and ~20x cheaper per point than GPT-5.5.


Findings

1. GLM-5.3-Flash profile

2. Coding gap to the flagship is small; the price gap is ~9x

Ampere.sh (Aug 2026) side-by-side, both at reasoning effort:

BenchmarkGLM-5.3-FlashGLM-5.3
Terminal Bench84.388.2
DeepSWE v163.466.9
AutomationBench48.848.2
AA Intelligence Index5760

3. If you ever want more headroom

4. Speed alternative

DeepSeek V4 Flash is the throughput king of the catalog (126 tok/s, rank 7/112 on AA's speed chart) at 2.8x flash's cost per point but still cheap in absolute terms. Consider it only for long high-volume batch jobs where wall-clock time dominates token cost.


Caveats


Current configuration (verified working, no change needed)

SurfaceSetting
Hermes daily driverz-ai/glm-5.3-flash via commandcode provider
omp @smol rolecommandcode/z-ai/glm-5.3-flash:max (~/.omp/agent/config.yml → modelRoles.smol)
omp pluginpi-commandcode-provider@0.6.4 (~/.omp/plugins)
AuthCOMMAND_CODE_API_KEY in ~/.zshenv (omp) / COMMANDCODE_API_KEY in ~/.hermes/.env (Hermes)

References

Scraped metrics and computed efficiency ratios (blend pricing, suite-cost-per-II-point) are in the session scratch files under /tmp/aa-.json; the raw Artificial Analysis page text is preserved in /tmp/aa-raw.json for audit.*