← Back to New
BENCHMARKS

GPT-5.6 Sol sweeps all four Agents' Last Exam splits

Sol ranks first across the four views among 20 models with scores, while Claude Fable 5 remains in the top five on each.

Event date Published

Sol leads every Agents' Last Exam split

Four Agents' Last Exam panels compare GPT-5.6 Sol and Claude Fable 5 across Overall, Near-term, Full-Spectrum, and Last-Exam. Sol ranks first in every panel among 20 models with scores.
PanelItemALE score (%)
OverallGPT-5.6 Sol (#1)53.62
OverallClaude Fable 5 (#4)48.7
Near-termGPT-5.6 Sol (#1)78.82
Near-termClaude Fable 5 (#5)71.13
Full-SpectrumGPT-5.6 Sol (#1)46.58
Full-SpectrumClaude Fable 5 (#3)42
Last-ExamGPT-5.6 Sol (#1)19.42
Last-ExamClaude Fable 5 (#2)16.5

Overall is the aggregate view; Near-term, Full-Spectrum, and Last-Exam are the source's published splits. Among 20 models with scores, Sol leads all four. Its selected settings are xhigh, max, xhigh, and medium in that order.