Best AI models for reasoning
Ranked by a 0–100 composite of reasoning and knowledge benchmarks (GPQA Diamond, Humanity's Last Exam, SimpleQA Verified, SimpleBench), scored as percentile rank against every model released in the past year.
Reasoning score is a 0–100 percentile composite across that area’s benchmarks. Open a model for the raw scores.
Frequently asked questions
What is the best AI model for reasoning?
Claude Fable 5 (Anthropic) currently ranks first for reasoning on Model Beat, followed by GPT-5.6 Sol and Gemini 3.1 Pro. Ranked by a 0–100 composite of reasoning and knowledge benchmarks (GPQA Diamond, Humanity's Last Exam, SimpleQA Verified, SimpleBench), scored as percentile rank against every model released in the past year.
Which is the most affordable strong reasoning model?
Among the top-ranked reasoning models, Muse Spark 1.1 is the cheapest at $1.25 per million input tokens.
How are these reasoning rankings calculated?
Ranked by a 0–100 composite of reasoning and knowledge benchmarks (GPQA Diamond, Humanity's Last Exam, SimpleQA Verified, SimpleBench), scored as percentile rank against every model released in the past year.
Benchmarks & model data from Epoch AI (CC BY); pricing & specs from OpenRouter. ECI = Epoch Capabilities Index.