gpt-oss-120b (batch)
by OpenAItext->text1 endpoint
Overview
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized to run on a single H100 GPU with native MXFP4 quantization. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation.
Capabilities
Text generation
Tool use / Function calling
Structured output
Extended reasoning
Long context
Top-P sampling
Stop sequences
Modalities
Input
Text
Output
Text
Technical Specifications
Context Window
131.1K tokens
Max Output
118.0K tokens
Knowledge Cutoff
2024-06-30
Tokenizer
GPT
Supported Parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Pricing
Per 1M tokens
Input$0.15
Output$0.60
Added
August 5, 2025
Last synced 9/19/2026