benchgap
General

SEC-Bench Pro leaderboard

As of 2026-10-07, the highest measured score on SEC-Bench Pro is 85.4% by GPT-6 Astra. 6 more models have estimated scores, calibrated from the benchmarks they were measured on.

Measured scores: benchlm.ai.

#ModelScoreSource
1GPT-6 Astra85.4%measured
2GPT-6.1 Sol78.8%measured
3Claude Mythos 576.1%estimated ± 4.3 pp, medium confidence
4Claude Mythos Preview73.3%estimated ± 4.3 pp, medium confidence
5GPT-5.6 Sol71.2%measured
6GLM-5.367.7%estimated ± 4.3 pp, medium confidence
7GPT-5.6 Terra67.0%estimated ± 4.3 pp, medium confidence
8GPT-6 Sol66.3%measured
9MiMo-V2.6-Pro66.3%measured
10DeepSeek V4.1 Flash62.8%measured
11GPT-5.6 Luna54.9%estimated ± 4.3 pp, medium confidence
12Kimi K353.9%estimated ± 4.3 pp, medium confidence
13MiMo-V2.6-Flash47.5%measured
14GPT-5.545.8%measured
15GPT-6 Luna34.2%measured

All leaderboards

Agentic · terminal

Agentic · tools

Coding

Math

Knowledge & reasoning

Instruction following

Multilingual

Vision & documents

Long context

General