GLM 4.7 Flash
by Z.aitext->text3 endpoints
Neura Intelligence Index
66.4/ 100
Rank
#108—
Confidence
mediumBenchmarks
6(15% coverage)
Overview
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
Capabilities
Text generation
Tool use / Function calling
Structured output
Extended reasoning
Long context
Top-P sampling
Stop sequences
Deterministic seed
Log probabilities
Modalities
Input
Text
Output
Text
Technical Specifications
Context Window
200K tokens
Max Output
118.0K tokens
Tokenizer
Other
Uptime (24h)
9887.3%
Parameters
358000000000B
License
mitOpen Weights
Country
🇨🇳CN
Release Date
December 22, 20258mo
Supported Parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Benchmark Performance
coding
| Benchmark | Score | Source |
|---|---|---|
| SWE-bench Verified | 73.8 | verified |
| TerminalBench | 33.3 | verified |
math
| Benchmark | Score | Source |
|---|---|---|
| AIME 2025 | 95.7 | verified |
reasoning
| Benchmark | Score | Source |
|---|---|---|
| GPQA Diamond | 85.7 | verified |
| Humanity's Last Exam | 42.8 | verified |
agent
| Benchmark | Score | Source |
|---|---|---|
| BrowseComp | 52 | verified |
Performance Over Time
Pricing
Per 1M tokens
Input$0.06
Output$0.40
Added
January 19, 2026
Last synced 9/16/2026