Best AI Models by Task: Coding, Math, Reasoning, Long Context, Open Weights, and Value
The leaderboard answers “best overall”. These pages answer “best for what” — each with a transparent, task-specific ranking rule and the same per-number sources.
Best AI Models for Coding
Coding and agentic models ordered by their published coding pillar, combining the benchmark variants available in this snapshot.
Leader: Qwen3.6 Max Preview · SI 66.4
Best AI Models for Math
Mathematical-reasoning models ordered by the math pillar from the available published evaluations.
Leader: GPT-6.1 Sol · SI 73.0
Best AI Models for Reasoning
The strongest general-reasoning models, ranked by the reasoning pillar of the SI Score (GPQA Diamond, Humanity's Last Exam, ARC-AGI-2).
Leader: Claude Opus 5.5 · SI 78.4
Best AI Models for Long Context
Models ordered by advertised input context window, with their SI Score alongside so you can trade reach against quality.
Leader: Llama 4 Scout 17B Instruct · SI 51.0
Best Open-Weight AI Models
Open-weight models ranked by SI Score, with the exact license each weight release carries shown next to it.
Leader: Kimi K3 · SI 71.0
Best AI Models by Price and Performance
The most SI Score per dollar, using a blended price of input and output tokens at a 3:1 mix.
Leader: Claude Haiku 5.5 · SI 64.8