Nemotron 3.5 Content Safety
by NVIDIAtext+image->text1 endpoint
Overview
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting text and image input and returning text output: a safe/unsafe classification for the user prompt and the response, safety category labels, and an optional reasoning trace. It covers 12 languages with a context window of up to 128K tokens.
It is suited for prompt and response moderation, content classification, safety pipelines, and enterprise AI guardrails with policy enforcement, and includes a togglable reasoning mode. It is part of the NVIDIA Nemotron family of open models for agentic AI.
Capabilities
Text generation
Image understanding
Extended reasoning
Long context
Top-P sampling
Stop sequences
Deterministic seed
Modalities
Input
TextImage
Output
Text
Technical Specifications
Context Window
131.1K tokens
Max Output
118.0K tokens
Tokenizer
Other
Uptime (24h)
10000.0%
Supported Parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyseedstoptemperaturetop_ktop_p
Pricing
Per 1M tokens
Input$0.20
Output$0.20
Added
June 4, 2026
Last synced 9/21/2026