Opus 5 Coding Benchmark
How Opus 5 ranks on the LLM Coding Leaderboard, shown next to the three models directly above and below it. Opus 5 was tested at 2 effort levels, each ranked separately.
- Medium effort — place #2, 17.75 points (max 20). Last evaluated on August 10, 2026, with Claude Code.
- High effort — place #4, 17.5 points (max 20). Last evaluated on August 12, 2026, with Claude Code.
| # | Model | Total points (max 20) |
Avg cost per prompt |
Avg time per prompt |
Video | Tested with | Points per project (max 5) | |||
|---|---|---|---|---|---|---|---|---|---|---|
| CSV Import (PHP) | Offline Sync (PHP) | Bank Feed (Dart/Flutter) | Shipping Quotes (Go) | |||||||
| 1 | GPT-5.6-Sol (Medium) | 18 | $1.01 | 05:14 | Codex CLI | 5 | 4 | 4 | 5 | |
| 2 | Opus 5 (Medium) | 17.75 | $1.10 | 03:48 | Claude Code | 5 | 3.25 | 4.5 | 5 | |
| 3 | GPT-5.6-Luna (Max) | 17.5 | $0.10 | 14:18 | Codex CLI | 5 | 4.5 | 3 | 5 | |
| 4 | Opus 5 (High) | 17.5 | $1.65 | 06:33 | Claude Code | 5 | 3.5 | 4 | 5 | |
| 5 | Tencent Hy3 (High) | 16.45 | $0.05 | 06:35 | OpenCode | 3.2 | 3.75 | 4.5 | 5 | |
| 6 | GPT-5.6-Terra (Medium) | 16.45 | $0.20 | 02:44 | Codex CLI | 4.2 | 3.75 | 4 | 4.5 | |
| 7 | GPT-5.6-Luna (Xhigh) | 15.75 | $0.06 | 09:07 | Codex CLI | 4 | 4.75 | 2 | 5 | |
Tutorials about Opus 5
Video
· Jul 25, 2026
I Tested NEW Opus 5 on 11 Coding Prompts
How to read this
- Each score measures a model-and-harness configuration, not the model in isolation.
- Prices are calculated with API costs. Models tested on a subscription plan show N/A.
- Full methodology and scoring formulas are explained in this article.
See every tested model side by side on the full LLM Coding Leaderboard.