Z.ai

GLM 4.5

by Z.aitext->text1 endpoint
Neura Intelligence Index
46.4/ 100
Rank
#194
Confidence
medium
Benchmarks
7(18% coverage)

Overview

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

Capabilities

Text generation
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
Top-P sampling

Modalities

Input
Text
Output
Text

Technical Specifications

Context Window
131.1K tokens
Max Output
98.3K tokens
Knowledge Cutoff
2024-12-31
Tokenizer
Other
Uptime (24h)
10000.0%
Parameters
355000000000B
License
mitOpen Weights
Country
🇨🇳CN
Release Date
July 28, 20251y

Supported Parameters

include_reasoningmax_tokensreasoningresponse_formattemperaturetool_choicetoolstop_ktop_p

Benchmark Performance

coding

BenchmarkScoreSource
SWE-bench Verified
64.2
verified
SciCode
41.7
verified
TerminalBench
37.5
verified

reasoning

BenchmarkScoreSource
GPQA Diamond
79.1
verified
Humanity's Last Exam
14.4
verified

agent

BenchmarkScoreSource
TAU-Bench Retail
79.7
verified
BrowseComp
26.4
verified

Performance Over Time

Pricing

Per 1M tokens
Input$0.60
Output$2.20
Cache read$0.11
Added
July 25, 2025
Last synced 9/19/2026