GPT-4o (Aug 2024)
ModelActiveby OpenAI · family “gpto-aug”
Epoch Capabilities Index
128.8 #160 of 268
90% CI 122.9 – 130.9
Released
6 Aug 2024
Context window
Unknown
Max output
Unknown
Input $ / 1M tokens
Unknown
Output $ / 1M tokens
Unknown
Cached input $ / 1M
Unknown
Each benchmark shown separately with its own source
Benchmark results (12)
| Benchmark | Domain | Score | vs best recorded | Setting | Run | Source |
|---|---|---|---|---|---|---|
| FrontierMath-Tiers-1-3-v2-Private | math | 0.4% ±0.4 | 0% | — | 27 Aug 2026 | Eval log ↗ |
| Chess Puzzles | games | 13.0% ±3.4 | 18% | — | 15 Jul 2026 | Eval log ↗ |
| SimpleQA Verified | knowledge | 26.0% ±1.4 | 34% | — | 31 Aug 2026 | Eval log ↗ |
| CadEval | coding | 26.0% | 35% | — | — | External ↗ |
| METR Time Horizons | agents | 33.8% | 40% | — | — | External ↗ |
| FrontierMath-2025-02-28-Privatesuperseded | math | 0.3% ±0.3 | 1% | — | 7 Mar 2025 | Epoch ↗ |
| Aider polyglot | coding | 23.1% | 26% | — | — | External ↗ |
| OTIS Mock AIME 2024-2025 | math | 6.4% ±2.6 | 6% | — | 25 Feb 2025 | Eval log ↗ |
| SimpleBench | reasoning | 17.8% | 22% | — | — | External ↗ |
| GPQA diamond | science | 49.2% ±2.6 | 51% | — | 27 Jan 2025 | Eval log ↗ |
| MATH level 5 | math | 53.3% ±1.1 | 54% | — | 27 Jan 2025 | Epoch ↗ |
| MMLU | knowledge | 84.3% | 96% | — | — | External ↗ |
Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.
API list price over time
Price history
$ per 1M tokens
This model has no public API price listing
0 recorded prices on — (OpenRouter listing); re-read on every data refresh, most recently 11h ago. Steps show when the price changed.
Events
- BenchmarkEpoch AI evaluates GPT-4o (Aug 2024)Epoch AI Benchmarking Hub
- BenchmarkEpoch AI evaluates GPT-4o (Aug 2024)Epoch AI Benchmarking Hub