73Fit score
Still available; Anthropic recommends migrating toward Opus 4.7 for latest improvements.
Model details →Writing quality is harder to benchmark than coding or math. The picks here weight model size, instruction-following, and reputational signal.
Still available; Anthropic recommends migrating toward Opus 4.7 for latest improvements.
Model details →Anthropic’s most capable generally available model for complex reasoning and agentic coding (per Anthropic models overview).
Model details →Balanced speed and intelligence with extended thinking support (per Anthropic models overview).
Model details →| # | Model | Provider | Context | Input price / 1M | Tier |
|---|---|---|---|---|---|
| 1 | Claude Opus 4.6 | Anthropic | 1M Ultra context (1M+) | Custom | Variable |
| 2 | Claude Opus 4.7 | Anthropic | 1M Ultra context (1M+) | Custom | Variable |
| 3 | Claude Sonnet 4.6 | Anthropic | 1M Ultra context (1M+) | Custom | Variable |
| 4 | Gemini 1.5 Flash | 1M Ultra context (1M+) | Custom | Variable | |
| 5 | Gemini 1.5 Pro | 1M Ultra context (1M+) | Custom | Variable | |
| 6 | Claude 3 Opus | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 7 | Claude 3.5 Sonnet | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 8 | Claude 3.7 Sonnet | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 9 | Claude Haiku 4.5 | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 10 | Claude Opus 4 | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 11 | Claude Opus 4.1 | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 12 | Claude Opus 4.5 | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 13 | Claude Sonnet 4 | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 14 | Claude Sonnet 4.5 | Anthropic | 200K Long context (128K+) | Custom | Variable |
| 15 | Gemini 2.0 Flash | 1M Ultra context (1M+) | Custom | Variable | |
| 16 | Llama 3.3 70B Versatile | Groq | 128K Long context (128K+) | Custom | Variable |
| 17 | Llama 3.1 405B Instruct | Meta | 128K Long context (128K+) | Custom | Variable |
| 18 | Llama 3.1 70B Instruct | Meta | 128K Long context (128K+) | Custom | Variable |
| 19 | Qwen2.5 72B Instruct | Qwen | 128K Long context (128K+) | Custom | Variable |
Subjective, but Claude (Sonnet/Opus), GPT-5, Gemini Pro, and Grok 4 are commonly cited as the most natural-sounding. Try the same prompt on each via the comparison pages.