Compare AI models
Claude Opus 4.8vsgpt-4.1
Which should you pick?
- gpt-4.1 · overall 69.8 vs 58.3
- gpt-4.1 cheaper output · $8.00 vs $25.00 /1M
- gpt-4.1 longer context · 1.0M tokens vs 1.0M tokens
Specifications
| Claude Opus 4.8 | gpt-4.1 | |
|---|---|---|
| Tier | A | B |
| Context window | 1.0M tokens | 1.0M tokens |
| Input price (per 1M tokens) | $5.00 | $2.00 |
| Output price (per 1M tokens) | $25.00 | $8.00 |
| Provenance | Runs in:USMade in:United States | Runs in:USMade in:United States |
Quality by category
Win rate per category: how often each model beats a field-average model on a prompt from that category. 50% is average. It is not a percentage of correct answers.
| Category | Claude Opus 4.8 | gpt-4.1 |
|---|---|---|
| coding | 46.4% | 74.8% |
| creative | 43.9% | 83.0% |
| factual | 70.4% | 66.4% |
| multilingual | 57.8% | 64.2% |
| reasoning | 74.0% | 73.3% |
Anthropic
Full review of Claude Opus 4.8 →
OpenAI
Full review of gpt-4.1 →