Skip to content
New research:Dots vs. Muse for Developers

Coding Models Index

Compare the models behind coding agents by release date and API price, with links to the providers’ documentation.

Latest releases

Recent additions, newest first. API prices in USD per million tokens; each card names its tier and checked date. Context tiers, caching, and tool fees can change the bill. Agent subscriptions are priced separately.

· OpenAI

GPT-6.1 Sol

$2 / $10

Input / output per 1M tokens · standard up to 272K input; cache reads $0.10; above 272K: $4 input / $15 output

Pricing checked October 2, 2026

The updated Sol model for coding and computer use, with a 1.05M-token context window. Use the Responses API for tool calling.

· Anthropic

Claude Sonnet 5.5

$2 / $10

Input / output per 1M tokens · standard API; cache reads $0.20

Pricing checked October 2, 2026

A faster, lower-cost alternative to Opus 5.5, with a 1M-token context window.

· Anthropic

Claude Opus 5.5

$4 / $20

Input / output per 1M tokens · standard API; cache reads $0.20

Pricing checked October 2, 2026

The newest Opus release. Standard input and output tokens cost 20% less than Opus 5.

· OpenAI

GPT-6 Sol

$2 / $10

Input / output per 1M tokens · standard; input up to 272K

Pricing checked October 2, 2026

The original GPT-6 Sol release. GPT-6.1 Sol is the newer version at the same standard input and output rates, with a lower cache-read price.

· OpenAI

GPT-6 Luna

$0.1 / $0.5

Input / output per 1M tokens · standard; input up to 272K

Pricing checked October 2, 2026

The lower-cost GPT-6 option for routine work and high-volume agent tasks.

· xAI

Grok 4.7

$2 / $6

Input / output per 1M tokens · standard; prompts below 200K

Pricing checked October 2, 2026

A 500K-context model with configurable reasoning. Its faster variant is available through Cursor and Grok Build.

· DeepSeek

DeepSeek V4.1 Flash

$0.3 / $1.2

Input / output per 1M tokens · peak, uncached; off-peak $0.15 input / $0.60 output; cache hits $0.006 peak / $0.003 off-peak

Pricing checked October 2, 2026

An open-weights model with native image input and 1M context. The deepseek-flash API alias serves V4.1; legacy V4 Flash IDs also route here.

· OpenAI

GPT-6 Astra

$10 / $50

Input / output per 1M tokens · standard; input up to 272K

Pricing checked October 2, 2026

OpenAI’s most capable broadly available model, combining coding, computer use, and research in long-running tasks.

· Google

Gemini 3.8 Flash

$0.75 / $3.75

Input / output per 1M tokens · standard promo through Dec 31, 2026; $1.50 input / $7.50 output from Jan 1, 2027

Pricing checked October 2, 2026

Google’s latest Flash model for software engineering and agent workflows. The separate Cyber variant has restricted access.

· Anthropic

Claude Fable 5.1

$10 / $50

Input / output per 1M tokens · standard API

Pricing checked October 2, 2026

Generally available for coding and knowledge work. The related Mythos 5.1 model is limited to trusted-access programs.

Compare coding models

Newest releases appear first. Every price has a checked or snapshot date. Turn on historical columns to see the August 26, 2026 benchmark snapshot and dated usage figures. An unscored model has no comparable result in this snapshot; it is not a zero score.

Looking for typed choices, scores, or classification? See the Decision Models Index.

30 of 30 models · sort by release date or output price
Coding models, release dates, and dated prices. Historical scores and usage are optional columns.
ModelWeightsAgent examples
GPT-6.1 SolNew releaseOpenAIThe updated Sol model for coding and computer use, with a 1.05M-token context window. Use the Responses API for tool calling.Model and pricingRelease log$10 output$2 input / 1M tokensChecked October 2, 2026standard up to 272K input; cache reads $0.10; above 272K: $4 input / $15 outputSep 29, 2026ClosedCodex
Claude Sonnet 5.5New releaseAnthropicA faster, lower-cost alternative to Opus 5.5, with a 1M-token context window.Model, release date, and pricing$10 output$2 input / 1M tokensChecked October 2, 2026standard API; cache reads $0.20Sep 28, 2026ClosedClaude Code
Claude Opus 5.5New releaseAnthropicThe newest Opus release. Standard input and output tokens cost 20% less than Opus 5.Release and pricing$20 output$4 input / 1M tokensChecked October 2, 2026standard API; cache reads $0.20Sep 22, 2026ClosedClaude Code
GPT-6 SolNew releaseOpenAIThe original GPT-6 Sol release. GPT-6.1 Sol is the newer version at the same standard input and output rates, with a lower cache-read price.Model and pricingRelease log$10 output$2 input / 1M tokensChecked October 2, 2026standard; input up to 272KSep 22, 2026ClosedCodex
GPT-6 LunaNew releaseOpenAIThe lower-cost GPT-6 option for routine work and high-volume agent tasks.Model and pricingRelease log$0.5 output$0.1 input / 1M tokensChecked October 2, 2026standard; input up to 272KSep 22, 2026ClosedCodex
Grok 4.7New releasexAIA 500K-context model with configurable reasoning. Its faster variant is available through Cursor and Grok Build.Release and pricingModel details$6 output$2 input / 1M tokensChecked October 2, 2026standard; prompts below 200KSep 21, 2026ClosedGrok Build / Cursor
DeepSeek V4.1 FlashNew releaseDeepSeekAn open-weights model with native image input and 1M context. The deepseek-flash API alias serves V4.1; legacy V4 Flash IDs also route here.Release logAPI pricing and aliasesWeights and MIT license$1.2 output$0.3 input / 1M tokensChecked October 2, 2026peak, uncached; off-peak $0.15 input / $0.60 output; cache hits $0.006 peak / $0.003 off-peakSep 10, 2026OpenDeepSeek Harness / OpenCode
GPT-6 AstraNew releaseOpenAIOpenAI’s most capable broadly available model, combining coding, computer use, and research in long-running tasks.Model and pricingRelease log$50 output$10 input / 1M tokensChecked October 2, 2026standard; input up to 272KSep 3, 2026ClosedCodex
Gemini 3.8 FlashNew releaseGoogleGoogle’s latest Flash model for software engineering and agent workflows. The separate Cyber variant has restricted access.ReleaseAPI pricing$3.75 output$0.75 input / 1M tokensChecked October 2, 2026standard promo through Dec 31, 2026; $1.50 input / $7.50 output from Jan 1, 2027Sep 2, 2026ClosedGemini CLI / Antigravity
Claude Fable 5.1New releaseAnthropicGenerally available for coding and knowledge work. The related Mythos 5.1 model is limited to trusted-access programs.Release and pricingRelease date$50 output$10 input / 1M tokensChecked October 2, 2026standard APISep 1, 2026ClosedClaude Code
GLM 5.3 Flash (Ox Alpha)Z.ai (revealed Aug 26)Released as Ox Alpha before being identified as GLM 5.3 Flash. Volume and availability notes refer to the August snapshot.Description from August snapshot$0.25 output$0.075 input / 1M tokensSnapshot: August 26, 2026free during previewAug 20, 2026PartialOpenCode (launch partner)
GLM 5.3Z.aiTop-10 coding at a fraction of frontier pricing; 5.2 weights are open, 5.3 Flash weights promised.Description from August snapshot$4.4 output$1.4 input / 1M tokensSnapshot: August 26, 2026Aug 14, 2026PartialZCode
Gemini 3.7 FlashGoogleAugust's fastest riser: near-frontier coding at Flash prices.Description from August snapshot$1.875 output$0.375 input / 1M tokensSnapshot: August 26, 202675% promoAug 13, 2026ClosedGemini CLI / Antigravity
Grok 4.6xAICheapest model in the 76+ coding tier; AA flags it for cost efficiency.Description from August snapshot$6 output$2 input / 1M tokensSnapshot: August 26, 2026Aug 12, 2026Closed—
DeepSeek V4 ProDeepSeekDeepSeek's reasoning flagship, refreshed mid-August.Description from August snapshotDeepSeek continues to serve V4 Pro after its previously announced September 14 retirement date.Current API pricing and availability$3.96 output$1.32 input / 1M tokensChecked October 2, 2026peak, uncached; off-peak $0.66 input / $1.98 outputAug 12, 2026Partial—
Muse Spark 1.2MetaMeta's return to the coding table, ahead of Sonnet 5 on the Coding Index.Description from August snapshot$4.25 output$1.25 input / 1M tokensSnapshot: August 26, 2026Aug 5, 2026ClosedMuse Code
Qwen 3.8 MaxAlibabaAlibaba's frontier bid; the 2.4T-parameter open-weights sibling shipped Aug 12.Description from August snapshot$6 output$2 input / 1M tokensSnapshot: August 26, 2026Aug 3, 2026PartialQwen Code
DeepSeek V4 FlashDeepSeekThe volume king of coding: usage spiked ~570% after the July 31 agent-tuned retrain.Description from August snapshotRetired on the DeepSeek API. Its legacy model ID now serves V4.1 Flash at V4.1 prices; the figures below are historical.Retirement and alias routing$0.1 output$0.05 input / 1M tokensSnapshot: August 26, 2026approx; tieredJul 31, 2026Partial—
Claude Opus 5Anthropic#1 on AA's Intelligence Index; Claude Code + Opus 5 also tops their Coding Agent Index.Description from August snapshot$25 output$5 input / 1M tokensSnapshot: August 26, 2026Jul 24, 2026ClosedClaude Code
Kimi K3Moonshot AIThe open-weights frontier for coding, within ~2 points of Opus 5.Description from August snapshot$15 output$3 input / 1M tokensSnapshot: August 26, 2026Jul 16, 2026OpenKimi Code CLI
GPT-5.6 SolOpenAIHighest Coding Index score in the August 26 snapshot.Description from August snapshotAPI pricing$20 output$4 input / 1M tokensChecked October 2, 2026standard promo at least through Nov 21, 2026; input up to 272KJul 9, 2026ClosedCodex
GPT-5.6 TerraOpenAICoding-skewed 5.6 sibling: top-5 coding despite mid-pack general intelligence.Description from August snapshot$12 output$2 input / 1M tokensSnapshot: August 26, 2026Jul 9, 2026ClosedCodex
GPT-5.6 LunaOpenAIThe volume play: near-Sonnet coding scores at commodity pricing.Description from August snapshot$1.2 output$0.2 input / 1M tokensSnapshot: August 26, 2026Jul 9, 2026ClosedCodex
Claude Sonnet 5AnthropicThe workhorse mid-tier for agentic coding at a fifth of Fable pricing.Description from August snapshot$10 output$2 input / 1M tokensSnapshot: August 26, 2026Jun 30, 2026ClosedClaude Code
Claude Fable 5AnthropicThe hard-task tier: #2 on intelligence, priciest mainstream model per token.Description from August snapshot$50 output$10 input / 1M tokensSnapshot: August 26, 2026Jun 9, 2026ClosedClaude Code
MiniMax M3MiniMaxOpen-weights long-horizon agent model with a 1M context at commodity cost.Description from August snapshotAPI pricing$1.2 output$0.3 input / 1M tokensChecked October 2, 2026standard up to 512K input; above 512K: $0.60 input / $2.40 output; priority 1.5×May 31, 2026Open—
Composer 2.5CursorCursor’s own coding model. Standard and Fast variants have different token prices.Description from August snapshotCursor pricing and variants$2.5 output$0.5 input / 1M tokensChecked October 2, 2026Cursor standard; default Fast variant is $3 input / $15 outputMay 18, 2026ClosedCursor
GPT-5.5OpenAIThe previous OpenAI flagship, still top-10 while the 5.6 family displaces it.Description from August snapshot$30 output$5 input / 1M tokensSnapshot: August 26, 2026Apr 24, 2026ClosedCodex
Gemini 3.1 ProGoogleGoogle's Pro tier, now behind its own 3.7 Flash on both AA indexes.Description from August snapshot$12 output$2 input / 1M tokensSnapshot: August 26, 2026Feb 19, 2026ClosedGemini CLI / Antigravity
Claude Haiku 4.5AnthropicCheap, fast Claude tier for high-volume subagent and edit loops; aging against the 2026 field.Description from August snapshot$5 output$1 input / 1M tokensSnapshot: August 26, 2026Oct 15, 2025ClosedClaude Code
Historical usage and spending snapshots

These observations describe July and August 2026. They do not establish current market share or model quality.

OpenRouter volume · August 26, 2026

High-volume models in the August 26 OpenRouter snapshot. Routed token volume is not a measure of coding quality or total adoption.

MiMo-V2.5Xiaomi9.92T/wk#3 on OpenRouter; Pro variant scores 60.2 on the Coding Index (secondhand)
Hy3Tencent7.16T/wkopen weights; Coding Index 58.8 (secondhand)
Nemotron 3 Ultra (free)NVIDIA5.4T/wkfree tier absorbing enormous routed volume
GLM 5.2Z.ai3.22T/wkopen weights (MIT); Coding Index 68.8; superseded by 5.3 but still huge
Solar Pro 4Upstage573B/wkAug 12 release; a sleeper climbing the rankings

API spending on Ramp · July 2026

Share of model API spend observed among US businesses on Ramp in July 2026. This measures spending within Ramp’s coverage, not the whole market or coding-agent usage.

Claude Opus 4.8
28%-2.9 pt
GPT-5.6 Sol
14.9%+14.9 pt
Claude Sonnet 4.6
8.3%-4.4 pt
Claude Fable 5
8%+6.2 pt
GPT-5.5
7.1%-4.7 pt
Claude Opus 4.6
6.9%-6.4 pt
Claude Sonnet 5
3.6%+3.6 pt
Claude Opus 5
3.5%+3.5 pt
GPT-5.4
2.4%-1.3 pt
GPT-5.6 Terra
2%+2 pt

Emerald = Anthropic, grey = OpenAI. Delta vs the prior month. Source: Ramp AI Index. Limited to spending visible in this source. Missing providers and models should not be read as having zero spend.

Sources and methodology

The Amplifying Coding Score blends two independent signals: 60% Artificial Analysis Coding Index (Terminal-Bench v2.1 + SciCode, a capability benchmark) and 40% arena coding Elo (OpenLM's Arena+ coding column, Aug 22 snapshot, an LLM-judged preference signal). Each is min-max normalized within this cohort, so 100 is the cohort frontier and 0 the cohort floor, not absolute quality. A model missing one component is scored on the other alone and marked °; missing both means no score and no rank.

Coding Index is a third-party benchmark, quoted at the effort setting noted per row. These are not Amplifying product benchmarks. Scores marked † were recorded from public mirrors of the AA board. They have not been reverified against the current primary board. Releases after the snapshot and models without captured benchmark evidence carry no score. This does not mean that public results are unavailable today.

Tokens per week is OpenRouter's routed volume in the dated snapshot. It covers requests routed through OpenRouter, not the whole market. First-party usage (Claude Code on Anthropic's API, Codex on OpenAI's, Gemini CLI on Google's) never appears there, so frontier-lab volumes are heavily understated. Figures marked ‡ are from a July 22 community snapshot because the model has since left the visible top lists.

Biz spend is the model's share of API spend among US businesses on Ramp's card data (July 2026, per the Ramp AI Index). A dash means no figure was captured, not zero spend. Spending and routed token volume do not feed into the coding score.

Tokenizers differ per provider. Price per token alone does not tell you the cost of a completed task: reasoning, retries, caching, and tool calls all affect the bill. Test models with the same agent, task, and settings before drawing conclusions about cost or quality.

Curated, not exhaustive. Release details and benchmark snapshots are updated separately. Official sources are linked on the new releases; historical benchmark and adoption sources include: Artificial Analysis, OpenRouter rankings, OpenRouter models API, and recent release coverage.

Models are half the story

The same model can behave differently with different tools, instructions, and permissions. The Coding Agents Index tracks publicly attributed agent activity: pull requests, commits, merge outcomes, and review patterns.

Coding Agents Index