Author's Description
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
Key Specifications
Supported Parameters
This model supports the following parameters:
Features
This model supports the following features:
Performance Summary
NVIDIA Nemotron 3.5 Content Safety demonstrates exceptional performance in terms of operational efficiency. It consistently ranks among the fastest models, achieving an "Infinityth percentile" across 8 benchmarks, indicating unparalleled speed. Similarly, its pricing is highly competitive, also securing an "Infinityth percentile" ranking across 8 benchmarks, making it an extremely cost-effective solution. However, the model exhibits significant limitations in its current benchmark evaluations for accuracy. Across all tested categories—Hallucinations, Instruction Following, General Knowledge, Coding, Email Classification, Reasoning, Ethics, and Mathematics—the model achieved 0.0% accuracy. This suggests that while the model is highly efficient and affordable, its ability to correctly answer or perform tasks in these domains is currently non-existent based on the provided data. Given its description as a "Content Safety" model, these benchmarks may not fully capture its intended function as a guardrail for LLMs and VLMs. Its strengths lie purely in its speed and cost-effectiveness, while its current benchmark results indicate a complete lack of general task performance.
Model Pricing
Current Pricing
| Feature | Price (per 1M tokens) |
|---|---|
| Prompt | $0.2 |
| Completion | $0.2 |
Price History
Available Endpoints
| Provider | Endpoint Name | Context Length | Pricing (Input) | Pricing (Output) |
|---|---|---|---|---|
|
DeepInfra
|
DeepInfra | nvidia/nemotron-3.5-content-safety-20260604 | 131K | $0.2 / 1M tokens | $0.2 / 1M tokens |
Benchmark Results
| Benchmark | Category | Reasoning | Strategy | Free | Executions | Accuracy | Cost | Duration |
|---|
Other Models by nvidia
|
|
Released | Params | Context |
|
Speed | Ability | Cost |
|---|---|---|---|---|---|---|---|
| NVIDIA: Nemotron 3.5 Lightning (free) | Aug 11, 2026 | ~30B | 262K |
Text input
Text output
|
★★ | ★★★ | $$$ |
| NVIDIA: Nemotron 3 Ultra (free) | Jun 03, 2026 | 550B | 262K |
Text input
Text output
|
★ | ★ | $$$$ |
| NVIDIA: Nemotron 3 Nano Omni (free) Unavailable | Apr 28, 2026 | 30B | N/A |
Video input
Image input
Text input
Audio input
Text output
|
— | — | — |
| NVIDIA: Nemotron 3 Super (free) | Mar 11, 2026 | 120B | 262K |
Text input
Text output
|
★★★ | ★★★ | $$$$ |
| NVIDIA: Nemotron 3 Nano 30B A3B | Dec 14, 2025 | 30B | 262K |
Text input
Text output
|
★★★ | ★★★★★ | $$$ |
| NVIDIA: Nemotron Nano 12B 2 VL (free) Unavailable | Oct 28, 2025 | 12B | 131K |
Video input
Image input
Text input
Text output
|
★ | ★★ | $$$$ |
| NVIDIA: Llama 3.3 Nemotron Super 49B V1.5 Unavailable | Oct 10, 2025 | 49B | 131K |
Text input
Text output
|
★★ | ★★★★ | $$$$ |
| NVIDIA: Nemotron Nano 9B V2 (free) Unavailable | Sep 05, 2025 | 9B | 128K |
Text input
Text output
|
★ | ★★ | $ |
| NVIDIA: Llama 3.3 Nemotron Super 49B v1 Unavailable | Apr 08, 2025 | 49B | 131K |
Text input
Text output
|
★★★★ | ★★ | $$ |
| NVIDIA: Llama 3.1 Nemotron Ultra 253B v1 Unavailable | Apr 08, 2025 | 253B | 131K |
Text input
Text output
|
★★ | ★★ | $$$$ |
| NVIDIA: Llama 3.1 Nemotron 70B Instruct Unavailable | Oct 14, 2024 | 70B | 131K |
Text input
Text output
|
★★★ | ★ | $$ |