Sonnet 5.5 Coding Benchmark
How Sonnet 5.5 ranks on the LLM Coding Leaderboard, shown next to the three models directly above and below it, with the top three models included for reference. Sonnet 5.5 was tested at 2 effort levels, each ranked separately.
- Last evaluated: 7 hours ago (September 29, 2026), with Claude Code.
- Last evaluated: 7 hours ago (September 29, 2026), with Claude Code.
| # | Model | Total points (max 60) |
Avg cost per prompt |
Avg time per prompt |
Video | Tested with | Points per project | |||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|
|
Laravel Code Quality
(max 20) |
React-TS Code Quality
(max 20) |
CSV Import (PHP) | Offline Sync (PHP) | Bank Feed (Dart / Flutter) | Shipping Quotes (Go) | |||||||
| 1 | Opus 5.5 (High) | 57.83 | $0.79 | 03:10 | Claude Code | 18.5 | 19.33 | 5 | 5 | 5 | 5 | |
| 2 | Opus 5.5 (Medium) | 57.37 | $0.56 | 02:04 | Claude Code | 18.95 | 19.67 | 5 | 4.75 | 5 | 4 | |
| 3 | GPT-6-Astra (High) | 56.77 | $0.96 | 04:34 | Codex CLI | 18.6 | 18.17 | 5 | 5 | 5 | 5 | |
| 4 | Sonnet 5.5 (High) | 56.18 | $0.34 | 01:53 | Claude Code | 18.68 | 18.5 | 5 | 4 | 5 | 5 | |
| 5 | GPT-6-Astra (Medium) | 56.03 | $0.69 | 03:08 | Codex CLI | 18.2 | 17.83 | 5 | 5 | 5 | 5 | |
| 6 | GPT-5.6-Sol (High) | 53.95 | $1.18 | 10:15 | Codex CLI | 17.45 | 17 | 5 | 4.5 | 5 | 5 | |
| 7 | Opus 5 (High) | 53.72 | $1.65 | 06:33 | Claude Code | 18.55 | 17.67 | 5 | 3.5 | 4 | 5 | |
| 8 | Opus 5.5 (Low) | 53.69 | $0.37 | 01:13 | Claude Code | 17.52 | 17.17 | 5 | 5 | 5 | 4 | |
| 9 | Fable 5.1 (Medium) | 53.63 | $1.53 | 03:06 | Claude Code | 18.55 | 16.83 | 4.5 | 3.75 | 5 | 5 | |
| 10 | Sonnet 5.5 (Medium) | 53.46 | $0.20 | 00:58 | Claude Code | 17.88 | 18.33 | 4.5 | 3.75 | 5 | 4 | |
| 11 | Space Bunny (Max) | 52.61 | N/A | 19:19 | OpenCode | 17.61 | 16.5 | 5 | 4.5 | 4 | 5 | |
| 12 | Opus 5 (Medium) | 52.57 | $1.10 | 03:48 | Claude Code | 17.65 | 17.17 | 5 | 3.25 | 4.5 | 5 | |
| 13 | GPT-6-Sol (High) | 52.42 | $0.31 | 05:18 | Codex CLI | 17.67 | 17.5 | 3.5 | 4.25 | 4.5 | 5 | |
Tutorials about Sonnet 5.5
Video
· Sep 29, 2026
I Tried NEW Sonnet 5.5 on 27 Coding Prompts
How to read this
- Each score measures a model-and-harness configuration, not the model in isolation.
- Prices are calculated with API costs. Models tested on a subscription plan show N/A.
- Full methodology and scoring formulas are explained in this article.
See every tested model side by side on the full LLM Coding Leaderboard.