OpenAI · model
GPT-5 mini benchmark scores
As of 2026-10-07, GPT-5 mini (OpenAI) has measured scores on 1 benchmark and estimated scores on 8 more.
| Benchmark | Score | Source |
|---|---|---|
| AA-SciCode | 47.8% | estimated ± 3.8 pp, high confidence |
| AA Coding Index | 46.4% | estimated ± 6.8 pp, medium confidence |
| SWE-bench Verified | 74.3% | estimated ± 4.4 pp, high confidence |
| FrontierCode 1.1 Main | 5.6% | estimated ± 4.3 pp, low confidence |
| Vals SWE-bench | 69.8% | estimated ± 6.3 pp, medium confidence |
| PostTrainBench v1.1 | 19.9% | estimated ± 10.5 pp, low confidence |
| Vibe Code Bench | 14.2% | measured |
| React Native Evals | 61.7% | estimated ± 3.0 pp, low confidence |
| SWE-Rebench | 7.3% | estimated ± 6.5 pp, low confidence |