| Spec | Claude Sonnet 4 | Gemini 2.5 Pro |
|---|
| Provider | Anthropic | |
| Release Date | May 2025 | May 2025 |
| Knowledge Cutoff | Mar 2025 | 2025-01-31 |
| Parameters | unknown | Undisclosed |
| Context Window | 200K | 1.0M |
| Max Output | — | — |
| Open Source | No | No |
| License | proprietary | proprietary |
| Tokenizer | — | — |
| Modality | — | — |
| Reasoning Model | No | No |
| Moderated | No | No |
| Price (per 1M tokens) | Claude Sonnet 4 | Gemini 2.5 Pro |
|---|---|---|
| Input | $3.00 | $1.25 |
| Output | $15.00 | $10.00 |
| Feature | Claude Sonnet 4 | Gemini 2.5 Pro |
|---|---|---|
| Text Input | ||
| Image Input (Vision) | ||
| Audio Input | ||
| Video Input | ||
| File Input | ||
| Image Output | ||
| Audio Output | ||
| Tool Use | ||
| Structured Output (JSON) | ||
| Streaming | ||
| Reasoning Tokens | ||
| Web Search | ||
| Temperature Control | ||
| Top-P Sampling | ||
| Stop Sequences | ||
| Seed (Reproducibility) | ||
| Log Probabilities |
| Benchmark | Claude Sonnet 4 | Gemini 2.5 Pro |
|---|---|---|
| MMLU-Pro | 80.2 | — |
| IFEval | 90.1 | — |
| Chatbot Arena Elo | 1310 | — |
| SimpleQA | — | 50.8 |
| Benchmark | Claude Sonnet 4 | Gemini 2.5 Pro |
|---|---|---|
| HumanEval | 93.8 | — |
| SWE-bench Verified | 53.2 | 63.2 |
| Benchmark | Claude Sonnet 4 | Gemini 2.5 Pro |
|---|---|---|
| BigBench-Hard | 90.8 | — |
| GPQA Diamond | 70.5 | 83 |
| Humanity's Last Exam | — | 17.8 |
| ARC-AGI v2 | — | 4.9 |
| Benchmark | Claude Sonnet 4 | Gemini 2.5 Pro |
|---|---|---|
| MMMU | 72.5 | 79.6 |