GPT-6-Astra Coding Benchmark
How GPT-6-Astra ranks on the LLM Coding Leaderboard, shown next to the three models directly above and below it, with the top three models included for reference.
- Last evaluated: 8 hours ago (September 6, 2026), with Codex CLI.
| # | Model | Total points (max 20) |
Avg cost per prompt |
Avg time per prompt |
Video | Tested with | Points per project (max 5) | |||
|---|---|---|---|---|---|---|---|---|---|---|
| CSV Import (PHP) | Offline Sync (PHP) | Bank Feed (Dart/Flutter) | Shipping Quotes (Go) | |||||||
| 1 | GPT-6-Astra (Medium) | 35.22 | $0.69 | 03:08 | Codex CLI | 5 | 5 | 5 | 5 | |
| 2 | GPT-5.6-Sol (High) | 34.72 | $1.18 | 10:15 | Codex CLI | 5 | 4.5 | 5 | 5 | |
| 3 | Opus 5 (High) | 33 | $1.65 | 06:33 | Claude Code | 5 | 3.5 | 4 | 5 | |
| 4 | Opus 5 (Medium) | 32.97 | $1.10 | 03:48 | Claude Code | 5 | 3.25 | 4.5 | 5 | |
Tutorials about GPT-6-Astra
PREMIUM
Article
· Sep 6, 2026
GPT-6-Astra in ChatGPT App: Browser Use and Light Level
How to read this
- Each score measures a model-and-harness configuration, not the model in isolation.
- Prices are calculated with API costs. Models tested on a subscription plan show N/A.
- Full methodology and scoring formulas are explained in this article.
See every tested model side by side on the full LLM Coding Leaderboard.