Skip to content

Llama 3.1-405B

ModelOpen sourceActive
by Meta · family “llamab”
Epoch Capabilities Index
128.8 #161 of 268
90% CI 123.0 – 130.8
Released
23 Jul 2024
Context window
Unknown
Max output
Unknown
Input $ / 1M tokens
Unknown
Output $ / 1M tokens
Unknown
Cached input $ / 1M
Unknown
Each benchmark shown separately with its own source

Benchmark results (15)

BenchmarkDomainScorevs best recordedSettingRunSource
DTBenchreasoning61.4%
62%
——External ↗
WeirdMLcoding21.4% ±0.0
23%
——External ↗
OTIS Mock AIME 2024-2025math9.7% ±3.2
10%
—25 Feb 2025Eval log ↗
The Agent Companyagents7.4%
17%
——External ↗
SimpleBenchreasoning23.0%
28%
——External ↗
Cybenchcoding7.5%
8%
——External ↗
GPQA diamondscience50.9% ±2.6
53%
—27 Jan 2025Eval log ↗
BBHreasoning82.9%
93%
——External ↗
MATH level 5math49.8% ±1.2
51%
—27 Jan 2025Epoch ↗
MMLUknowledge84.5%
96%
——External ↗
PIQAother85.9%
97%
——External ↗
Winograndeother89.2%
100%
——External ↗
HellaSwagreasoning89.2%
94%
——External ↗
ARC AI2other95.3%
100%
——External ↗
TriviaQAother82.7%
94%
——External ↗

Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.

API list price over time

Price history

$ per 1M tokens
This model has no public API price listing

0 recorded prices on — (OpenRouter listing); re-read on every data refresh, most recently 9h ago. Steps show when the price changed.

Events

No events recorded yet.