Skip to content

o1 (medium)

ModelActive
by OpenAI · family “o1-medium”
Epoch Capabilities Index
Not scored by Epoch AI
Released
17 Dec 2024
Context window
Unknown
Max output
Unknown
Input $ / 1M tokens
Unknown
Output $ / 1M tokens
Unknown
Cached input $ / 1M
Unknown
Each benchmark shown separately with its own source

Benchmark results (13)

BenchmarkDomainScorevs best recordedSettingRunSource
FrontierMath-Tiers-1-3-v2-Privatemath10.2% ±1.8
11%
medium27 Aug 2026Eval log ↗
Chess Puzzlesgames12.0% ±3.3
17%
medium7 Aug 2026Eval log ↗
CadEvalcoding56.0%
76%
medium—External ↗
METR Time Horizonsagents51.1%
60%
medium—External ↗
GeoBenchmultimodal80.0%
91%
medium—External ↗
Fiction.LiveBenchlong-context83.3%
86%
medium—External ↗
Lech Mazur Writingother70.2%
82%
medium—External ↗
VPCTmultimodal37.0%
41%
medium—External ↗
OTIS Mock AIME 2024-2025math73.3% ±6.7
73%
medium27 Feb 2025Eval log ↗
SimpleBenchreasoning36.7%
45%
medium—External ↗
GPQA diamondscience75.8% ±3.1
79%
medium27 Jan 2025Eval log ↗
MATH level 5math94.4% ±0.6
96%
medium27 Jan 2025Epoch ↗
ARC-AGIreasoning30.7%
31%
medium—External ↗

Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.

API list price over time

Price history

$ per 1M tokens
This model has no public API price listing

0 recorded prices on — (OpenRouter listing); re-read on every data refresh, most recently 11h ago. Steps show when the price changed.

Events