Cost-performance analysis
Which Model Is Best for Coding?
A cost-performance leaderboard of the models coding agents run on - benchmark scores, real token-mix pricing, latency, and decoding speed.
Models used by coding agents
Sort by:
- Claude Opus 5.5 Anthropic#1DF Score 99.6Avg task cost —Latency —Context window 1M
- Claude Fable 5.1 Anthropic#2DF Score 90.4Avg task cost $2.73Latency 210.79sContext window 1M
- GPT-6 Astra OpenAI#3DF Score 88.3Avg task cost $1.15Latency —Context window 1M
- GPT-6 Sol OpenAI#4DF Score 81.4Avg task cost —Latency 136.12sContext window 872k
- Claude Fable 5 Anthropic#5DF Score 80.9Avg task cost $2.36Latency 239.01sContext window 1M
- Claude Opus 5 Anthropic#6DF Score 79.7Avg task cost $0.821Latency 56.78sContext window 1M
- GPT-5.6 Sol OpenAI#7DF Score 79.3Avg task cost $0.350Latency —Context window 1.05M
- GPT-5.6 Terra OpenAI#8DF Score 73.7Avg task cost $0.234Latency —Context window 1.05M
- Gemini 3.8 Flash Google#9DF Score 67.9Avg task cost $0.135Latency 14.66sContext window 1M
- GPT-5.4 OpenAI#10DF Score 63.3Avg task cost $0.254Latency 113.80sContext window 1M
- GPT-5.5 OpenAI#11DF Score 62.8Avg task cost $0.490Latency 118.46sContext window 1.05M
- Claude Opus 4.8 Anthropic#12DF Score 61.3Avg task cost $1.16Latency 30.05sContext window 1M
- Claude Opus 4.7 Anthropic#13DF Score 60.7Avg task cost $1.58Latency —Context window 1M
- GPT-5.6 Luna OpenAI#14DF Score 60.4Avg task cost $0.0301Latency —Context window 1.05M
- Claude Sonnet 5 Anthropic#15DF Score 57.1Avg task cost $0.969Latency 144.37sContext window 1M
- Gemini 3.1 Pro (Preview) Google#16DF Score 52.5Avg task cost $0.783Latency 25.74sContext window 1M
- Claude Opus 4.6 Anthropic#17DF Score 51.6Avg task cost $0.581Latency —Context window 1M
- Gemini 3.5 Flash Google#18DF Score 48.3Avg task cost $0.620Latency 20.59sContext window 1M
- Claude Opus 4.5 Anthropic#19DF Score 48.1Avg task cost —Latency 15.04sContext window 200k
- Gemini 3 Flash (Preview) Google#20DF Score 46.3Avg task cost —Latency 7.60sContext window 1M
- Claude Sonnet 4.6 Anthropic#21DF Score 41.4Avg task cost —Latency 1.16sContext window 1M
- GPT-5.4 mini OpenAI#22DF Score 39.6Avg task cost $0.183Latency 11.87sContext window 400k
- Kimi-K2.7-Code Moonshot AI#23DF Score 38.3Avg task cost —Latency 2.92sContext window 256k
- GPT-5.4 nano OpenAI#24DF Score 28.6Avg task cost —Latency 4.74sContext window 400k
- Claude Sonnet 4.5 Anthropic#25DF Score 27.2Avg task cost —Latency 1.36sContext window 1M
- Claude Haiku 4.5 Anthropic#26DF Score 21.6Avg task cost $0.159Latency 0.81sContext window 200k
- Gemini 2.5 Pro Google#27DF Score 20.2Avg task cost —Latency 22.54sContext window 1M
- GPT-5 mini OpenAI#28DF Score 19.3Avg task cost —Latency 85.30sContext window 400k
- GPT-OSS 120B OpenAI#29DF Score 0.0Avg task cost —Latency 0.94sContext window 131k
- Raptor mini GitHub#30DF Score —Avg task cost —Latency —Context window 400k
- MAI-Code-1-Flash Microsoft#31DF Score —Avg task cost —Latency —Context window
- Claude Opus 4.8 Fast Anthropic#32DF Score —Avg task cost —Latency —Context window 1M
Help us to detect updates in plans values
The prices and token mixes above are driven by real coding usage. Share yours via letmecode to allow us discover updates in plans values.
$ npx letmecode@latest
FAQ
Something unclear or missing?
If any of the numbers or terms above don't add up, or you spotted something that looks off, tell us - we'll clarify and keep the data sharper for everyone.
Report inconsistency