GLM 4.7 Flash
ModelOpen sourceActiveby Z.ai (Zhipu AI) · family “glm-flash”
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,... (description from the OpenRouter listing)
in: textout: textReasoningTool useStructured output
Epoch Capabilities Index
Not scored by Epoch AI
Released
19 Jan 2026
Context window
200K tokens
Max output
118K tokens
Input $ / 1M tokens
$0.060
Output $ / 1M tokens
$0.40
Cached input $ / 1M
Unknown
Each benchmark shown separately with its own source
Benchmark results (3)
| Benchmark | Domain | Score | vs best recorded | Setting | Run | Source |
|---|---|---|---|---|---|---|
| Chess Puzzles | games | 0.0% | 0% | none | 28 Aug 2026 | Eval log ↗ |
| OTIS Mock AIME 2024-2025 | math | 58.3% ±6.6 | 58% | — | 30 Aug 2026 | Eval log ↗ |
| GPQA diamond | science | 60.5% ±2.9 | 63% | — | 30 Aug 2026 | Eval log ↗ |
Source: Epoch AI Benchmarking Hub (CC BY 4.0). “External” rows are leaderboard results Epoch collects from third parties. Best reported setting per benchmark is shown.
API list price over time
Price history
$ per 1M tokens
9 recorded prices on 1 Feb 2026 (OpenRouter listing + Internet Archive snapshots); re-read on every data refresh, most recently 1h ago. Steps show when the price changed.
Events
- BenchmarkEpoch AI evaluates GLM 4.7 FlashEpoch AI Benchmarking Hub
- BenchmarkEpoch AI evaluates GLM 4.7 FlashEpoch AI Benchmarking Hub
- Open releaseZ.ai releases GLM 4.7 FlashOpenRouter models API