Cost-performance analysis
Which Model Is Best for Coding?
A cost-performance leaderboard of the models coding agents run on - benchmark scores, real token-mix pricing, latency, and decoding speed.
Models used by coding agents
Sort by:
- GPT-5.6 Sol OpenAI#1DF Score 87.4Avg task cost $0.486Latency —Context window 1.05M
- GPT-5.5 OpenAI#2DF Score 85.0Avg task cost $0.490Latency 118.46sContext window 1.05M
- GPT-5.6 Terra OpenAI#3DF Score 83.5Avg task cost $0.293Latency —Context window 1.05M
- Claude Opus 5 Anthropic#4DF Score 81.3Avg task cost $0.565Latency 56.78sContext window 1M
- Gemini 3.5 Flash Google#5DF Score 81.3Avg task cost $0.620Latency 20.59sContext window 1M
- Claude Fable 5 Anthropic#6DF Score 78.5Avg task cost $2.36Latency 239.01sContext window 1M
- GPT-5.4 OpenAI#7DF Score 75.9Avg task cost $0.254Latency 113.80sContext window 1M
- Gemini 3.1 Pro (Preview) Google#8DF Score 74.5Avg task cost $0.783Latency 25.74sContext window 1M
- GPT-5.6 Luna OpenAI#9DF Score 70.9Avg task cost $0.151Latency —Context window 1.05M
- Claude Opus 4.7 Anthropic#10DF Score 69.4Avg task cost $1.58Latency —Context window 1M
- Claude Sonnet 5 Anthropic#11DF Score 69.2Avg task cost $0.969Latency 144.37sContext window 1M
- Claude Opus 4.8 Anthropic#12DF Score 67.8Avg task cost $1.16Latency 30.05sContext window 1M
- GPT-5.3 Codex OpenAI#13DF Score 61.9Avg task cost —Latency 79.89sContext window 400k
- GPT-5.4 mini OpenAI#14DF Score 59.9Avg task cost $0.183Latency 11.87sContext window 400k
- Gemini 3 Flash (Preview) Google#15DF Score 57.0Avg task cost —Latency 7.60sContext window 1M
- GPT-5.4 nano OpenAI#16DF Score 53.6Avg task cost —Latency 4.74sContext window 400k
- Claude Sonnet 4.6 Anthropic#17DF Score 50.0Avg task cost —Latency 1.16sContext window 1M
- Kimi-K2.7-Code Moonshot AI#18DF Score 48.8Avg task cost —Latency 2.92sContext window 256k
- Claude Opus 4.6 Anthropic#19DF Score 45.4Avg task cost $0.581Latency —Context window 1M
- Claude Opus 4.5 Anthropic#20DF Score 37.2Avg task cost —Latency 15.04sContext window 200k
- Claude Sonnet 4.5 Anthropic#21DF Score 33.8Avg task cost —Latency 1.36sContext window 1M
- Claude Haiku 4.5 Anthropic#22DF Score 28.0Avg task cost $0.159Latency 0.81sContext window 200k
- GPT-5 mini OpenAI#23DF Score 26.8Avg task cost —Latency 85.30sContext window 400k
- GPT-OSS 120B OpenAI#24DF Score 15.9Avg task cost —Latency 0.94sContext window 131k
- Gemini 2.5 Pro Google#25DF Score 12.1Avg task cost —Latency 22.54sContext window 1M
- GPT-5.3 Codex Spark OpenAI#26DF Score —Avg task cost —Latency —Context window
- Claude Opus 4.8 Fast Anthropic#27DF Score —Avg task cost —Latency —Context window 1M
- MAI-Code-1-Flash Microsoft#28DF Score —Avg task cost —Latency —Context window
- Raptor mini GitHub#29DF Score —Avg task cost —Latency —Context window 400k
Help us to detect updates in plans values
The prices and token mixes above are driven by real coding usage. Share yours via letmecode to allow us discover updates in plans values.
$ npx letmecode@latest
FAQ
Something unclear or missing?
If any of the numbers or terms above don't add up, or you spotted something that looks off, tell us - we'll clarify and keep the data sharper for everyone.
Report inconsistency