Qwen3.8 2.4T A95B
by Qwentext->text7 endpoints
Overview
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is suited for coding, research, complex reasoning, and agentic workflows.
Capabilities
Text generation
Tool use / Function calling
Structured output
Extended reasoning
Prompt caching
Long context
Top-P sampling
Stop sequences
Deterministic seed
Log probabilities
Modalities
Input
Text
Output
Text
Technical Specifications
Context Window
1.0M tokens
Max Output
131.1K tokens
Tokenizer
Qwen
Uptime (24h)
10000.0%
Parameters
2400000000000B
License
qwen3_8_maxOpen Weights
Country
🇨🇳CN
Release Date
August 12, 20261mo
Supported Parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Pricing
Per 1M tokens
Input$2.00
Output$6.00
Cache read$0.25
Added
August 12, 2026
Last synced 9/16/2026