| Spec | Llama 4 Maverick | o3 |
|---|---|---|
| Provider | Meta | OpenAI |
| Release Date | Apr 2025 | Apr 2025 |
| Knowledge Cutoff | — | 2024-05-31 |
| Parameters | 400000000000B | Undisclosed |
| Context Window | — | — |
| Max Output | — |
| — |
| Open Source | Yes | No |
| License | llama_4_community_license_agreement | proprietary |
| Tokenizer | — | — |
| Modality | — | — |
| Reasoning Model | No | No |
| Moderated | No | No |
| Price (per 1M tokens) | Llama 4 Maverick | o3 |
|---|
| Feature | Llama 4 Maverick | o3 |
|---|---|---|
| Text Input | ||
| Image Input (Vision) | ||
| Audio Input | ||
| Video Input | ||
| File Input | ||
| Image Output | ||
| Audio Output | ||
| Tool Use | ||
| Structured Output (JSON) | ||
| Streaming | ||
| Reasoning Tokens | ||
| Web Search | ||
| Temperature Control | ||
| Top-P Sampling | ||
| Stop Sequences | ||
| Seed (Reproducibility) | ||
| Log Probabilities |
| Benchmark | Llama 4 Maverick | o3 |
|---|---|---|
| GPQA Diamond | 69.8 | 83.3 |
| Humanity's Last Exam | — | 14.7 |
| ARC-AGI v2 | — | 6.5 |
| Benchmark | Llama 4 Maverick | o3 |
|---|---|---|
| MMMU | 73.4 | 82.9 |
| MMMU-Pro | 59.6 | 76.4 |
| CharXiv Reasoning | — | 78.6 |
| Benchmark | Llama 4 Maverick | o3 |
|---|---|---|
| SWE-bench Verified | — | 69.1 |
| Benchmark | Llama 4 Maverick | o3 |
|---|---|---|
| AIME 2025 | — | 86.4 |
| FrontierMath | — | 15.8 |
| Benchmark | Llama 4 Maverick | o3 |
|---|---|---|
| BrowseComp | — | 49.7 |