Z.ai

GLM 4.7 Flash

by Z.aitext->text3 endpoints
Neura Intelligence Index
66.4/ 100
Rank
#108
Confidence
medium
Benchmarks
6(15% coverage)

Overview

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

Capabilities

Text generation
Tool use / Function calling
Structured output
Extended reasoning
Long context
Top-P sampling
Stop sequences
Deterministic seed
Log probabilities

Modalities

Input
Text
Output
Text

Technical Specifications

Context Window
200K tokens
Max Output
118.0K tokens
Tokenizer
Other
Uptime (24h)
9887.3%
Parameters
358000000000B
License
mitOpen Weights
Country
🇨🇳CN
Release Date
December 22, 20258mo

Supported Parameters

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Benchmark Performance

coding

BenchmarkScoreSource
SWE-bench Verified
73.8
verified
TerminalBench
33.3
verified

math

BenchmarkScoreSource
AIME 2025
95.7
verified

reasoning

BenchmarkScoreSource
GPQA Diamond
85.7
verified
Humanity's Last Exam
42.8
verified

agent

BenchmarkScoreSource
BrowseComp
52
verified

Performance Over Time

Pricing

Per 1M tokens
Input$0.06
Output$0.40
Added
January 19, 2026
Last synced 9/16/2026