The world's leading AI models, ranked
Aggregated live from Artificial Analysis and OpenRouter. Track every frontier model: intelligence, coding, math, speed and cost, all in one place.
| Rank | Model | |||||||||
|---|---|---|---|---|---|---|---|---|---|---|
| 1 | Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) Anthropic | 63.4 | 53.4 | 81.6 | 70 | $20 | ||||
| 2 | Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback) Anthropic | 62.9 | 53.2 | 80.7 | 63 | $20 | ||||
| 3 | GPT-6 Astra (Max) OpenAI | 61.2 | 52.7 | 76.9 | 56 | $20 | ||||
| 4 | Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback) Anthropic | 61 | 51.2 | 79.1 | 56 | $20 | ||||
| 5 | GPT-6 Astra (Xhigh) OpenAI | 60.7 | 52.4 | 75.9 | 52 | $20 | ||||
| 6 | Claude Opus 5 (Adaptive Reasoning, Max Effort) Anthropic | 60.4 | 50.8 | 78 | 0 | $10 | ||||
| 7 | GPT-6 Astra OpenAI📄🖼T | 60.1 | 50.9 | 77.1 | 52 | $20 | 1.1M | |||
| 8 | Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) Anthropic | 59.3 | 49.7 | 77 | 0 | $10 | ||||
| 9 | Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) Anthropic | 59.1 | 49.6 | 76.5 | 0 | $20 | ||||
| 10 | Claude Opus 5 (Adaptive Reasoning, High Effort) Anthropic | 58.1 | 48.1 | 76.5 | 0 | $10 |
Frequently asked questions
+ What is the best AI model right now?
By the DataCore composite score (intelligence, coding, math), Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) currently leads our ranking of 647 AI models from 65 providers, updated continuously.
+ Which AI model is the cheapest?
Several models is among the lowest-cost frontier options. You can sort and filter the leaderboard by price.
+ What is the fastest LLM?
Celeris-1 reaches the highest output speed (~2357 tokens/sec). Speed and latency are measured by Artificial Analysis.
+ What is an AI model "time horizon"?
It's METR's metric: the length of task (in human time) a model can complete with a 50% success rate. A longer time horizon means the model can handle longer, more complex agentic tasks.
+ What is the best open-source LLM?
The board tracks 249 open-weight models such as DeepSeek, Qwen, Llama and GLM. Use the "Open weights" filter to compare them head-to-head.
+ Where does the data come from and how often is it updated?
Aggregated from Artificial Analysis (intelligence, coding, math, speed, price), METR (time horizon) and OpenRouter (catalog, live pricing). Data refreshes continuously (≈10-minute cache).
Data sources
- OpenRoutercached 1h ago
Model catalog, context window, modalities, live pricing
- Artificial Analysiscached 1h ago
Intelligence Index, coding & math scores, speed, latency
50% task-completion time horizon (how long a task a model can finish)
- Aider Polyglotcached 1h ago
Polyglot code-editing benchmark (pass rate %)