NVIDIA

Nemotron 3.5 Content Safety

by NVIDIAtext+image->text1 endpoint

Overview

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting text and image input and returning text output: a safe/unsafe classification for the user prompt and the response, safety category labels, and an optional reasoning trace. It covers 12 languages with a context window of up to 128K tokens.

It is suited for prompt and response moderation, content classification, safety pipelines, and enterprise AI guardrails with policy enforcement, and includes a togglable reasoning mode. It is part of the NVIDIA Nemotron family of open models for agentic AI.

Capabilities

Text generation
Image understanding
Extended reasoning
Long context
Top-P sampling
Stop sequences
Deterministic seed

Modalities

Input
TextImage
Output
Text

Technical Specifications

Context Window
131.1K tokens
Max Output
118.0K tokens
Tokenizer
Other
Uptime (24h)
10000.0%

Supported Parameters

frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyseedstoptemperaturetop_ktop_p

Pricing

Per 1M tokens
Input$0.20
Output$0.20
Added
June 4, 2026
Last synced 9/21/2026