GLM 4.6
by Z.aitext->text4 endpoints
Neura Intelligence Index
57.3/ 100
Rank
#152—
Confidence
mediumBenchmarks
6(15% coverage)
Overview
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
Capabilities
Text generation
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
Top-P sampling
Stop sequences
Deterministic seed
Modalities
Input
Text
Output
Text
Technical Specifications
Context Window
204.8K tokens
Max Output
16.4K tokens
Knowledge Cutoff
2025-03-31
Tokenizer
Other
Uptime (24h)
9975.4%
Parameters
357000000000B
License
mitOpen Weights
Country
🇨🇳CN
Release Date
September 30, 202511mo
Supported Parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Benchmark Performance
coding
| Benchmark | Score | Source |
|---|---|---|
| SWE-bench Verified | 68 | verified |
| TerminalBench | 40.5 | verified |
math
| Benchmark | Score | Source |
|---|---|---|
| AIME 2025 | 93.9 | verified |
reasoning
| Benchmark | Score | Source |
|---|---|---|
| GPQA Diamond | 81 | verified |
| Humanity's Last Exam | 17.2 | verified |
agent
| Benchmark | Score | Source |
|---|---|---|
| BrowseComp | 45.1 | verified |
Performance Over Time
Pricing
Per 1M tokens
Input$0.43
Output$1.75
Cache read$0.08
Added
September 30, 2025
Last synced 9/21/2026