Skip to content

GLM 4.7 Flash

ModelOpen sourceActive
by Z.ai (Zhipu AI) · family “glm-flash”

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,... (description from the OpenRouter listing)

in: textout: textReasoningTool useStructured output
Epoch Capabilities Index
Not scored by Epoch AI
Listing ↗
Released
19 Jan 2026
Context window
200K tokens
Max output
118K tokens
Input $ / 1M tokens
$0.060
Output $ / 1M tokens
$0.40
Cached input $ / 1M
Unknown
Each benchmark shown separately with its own source

Benchmark results (3)

BenchmarkDomainScorevs best recordedSettingRunSource
Chess Puzzlesgames0.0%
0%
none28 Aug 2026Eval log ↗
OTIS Mock AIME 2024-2025math58.3% ±6.6
58%
—30 Aug 2026Eval log ↗
GPQA diamondscience60.5% ±2.9
63%
—30 Aug 2026Eval log ↗

Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.

API list price over time

Price history

$ per 1M tokens

9 recorded prices on 1 Feb 2026 (OpenRouter listing + Internet Archive snapshots); re-read on every data refresh, most recently 1h ago. Steps show when the price changed.

Events