// ranked · September 2026
Best LLM for Coding
Ranked by LiveBench’s coding category: code generation and completion on recent competitive-programming and LeetCode-style problems.
Claude Opus 5.5 leads on LiveBench coding at 89.3, ahead of Claude Fable 5.1 (86.4) and Claude Fable 5 (86.0).
Claude Opus 5.5 vs Claude Fable 5.1, head to headBest value in the top 10
Claude Opus 4.7
Lowest measured cost per point among the leaders — $0.0204 per point for a score of 82.1.
Best under $1.00 / 1M
GPT-5.6 Luna
Highest score at a blended list price of $1.00 per 1M tokens or less — 82.9 at $0.450.
Top 10
| # | Model | Coding score | Blended / 1M | Context | Capabilities |
|---|---|---|---|---|---|
| 1 | Claude Opus 5.5 anthropic/claude-opus-5.5 | 89.3 | $8.00 | 1M | reasoningtoolsvision |
| 2 | Claude Fable 5.1 anthropic/claude-fable-5.1 | 86.4 | $20.00 | 1M | reasoningtoolsvision |
| 3 | Claude Fable 5 anthropic/claude-fable-5 | 86.0 | $20.00 | 1M | reasoningtoolsvision |
| 4 | GPT-5.6 Sol openai/gpt-5.6-sol | 83.9 | $4.00 | 1.1M | reasoningtoolsvision |
| 5 | GPT-5.2-Codex openai/gpt-5.2-codex | 83.6 | $4.81 | 400K | reasoningtoolsvision |
| 6 | GPT-5.6 Luna openai/gpt-5.6-luna | 82.9 | $0.450 | 1.1M | reasoningtoolsvision |
| 7 | GPT-5.5 openai/gpt-5.5 | 82.1 | $11.25 | 1.1M | reasoningtoolsvision |
| 8 | Claude Opus 4.7 anthropic/claude-opus-4.7 | 82.1 | $10.00 | 1M | reasoningtoolsvision |
| 9 | Claude Opus 4.8 anthropic/claude-opus-4.8 | 81.8 | $10.00 | 1M | reasoningtoolsvision |
| 10 | GPT-6 Sol openai/gpt-6-sol | 81.8 | $4.00 | 1.1M | reasoningtoolsvision |
45 more models qualify — see the full leaderboard.
FAQ
What is the best LLM for Coding in September 2026?
Claude Opus 5.5 leads on LiveBench coding at 89.3, ahead of Claude Fable 5.1 (86.4) and Claude Fable 5 (86.0).
What is the best value in the top 10 for coding?
Claude Opus 4.7. Lowest measured cost per point among the leaders — $0.0204 per point for a score of 82.1.
What is the best under $1.00 / 1m for coding?
GPT-5.6 Luna. Highest score at a blended list price of $1.00 per 1M tokens or less — 82.9 at $0.450.
How is this ranking produced?
Ranked by LiveBench’s coding category: code generation and completion on recent competitive-programming and LeetCode-style problems.
Other rankings
How this ranking is produced
- One metric, stated above — nothing here is weighted or scored by us. Scores come from LiveBench release 2026-06-25; models without a published run don't appear in score-based rankings.
- Price, context and capabilities — live from OpenRouter, refreshed every 15 minutes.
- Ties — scores less than a point apart are called a tie; effort settings alone move a LiveBench score by more than that.
A public benchmark is someone else's workload. Before committing, see LLM & agent evaluation.