Reasoning Models
Models with reasoning capability or strong reasoning benchmark results.
Based on 10 AI models
| Model | Provider | Score | Price (in) | Context |
|---|---|---|---|---|
| Claude 3.5 Sonnet | Anthropic | 78.4 | $3.00 | 200,000 |
| Gemini 1.5 Pro | 76.9 | $1.25 | 2,000,000 | |
| Mistral Large 2 | Mistral | 73.6 | $2.00 | 128,000 |
| GPT-4o | OpenAI | 73.1 | $2.50 | 128,000 |
| Claude 3 Opus | Anthropic | 60.3 | $15.00 | 200,000 |
| Llama 3.1 405B | Meta | 58.1 | $3.00 | 128,000 |
| GPT-4o mini | OpenAI | 55.1 | $0.15 | 128,000 |
| DeepSeek R1 | DeepSeek | 55 | $0.55 | 65,536 |
| Claude 3 Haiku | Anthropic | 54.6 | $0.25 | 200,000 |
| Gemini 2.0 Flash | 47.5 | $0.10 | 1,048,576 |
Frequently Asked Questions
What is the best model for Reasoning Models?
Based on our data, Claude 3.5 Sonnet ranks first for Reasoning Models with a score of 78.
How were these models selected?
Models are ranked by scenario-specific data: benchmark scores (coding/reasoning/math/vision), capability support, context window, or price. All data comes from our transparent model catalog, updated daily.
How fresh is this data?
Model data, prices, and rankings are refreshed on a regular schedule; ranking scores are snapshotted daily. Each model page shows its last verified date.