NVIDIA: Nemotron 3.5 Content Safety (free)

Image input Text input Text output
Author's Description

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

Key Specifications
Cost
$$
Context
131K
Parameters
4B (Rumoured)
Released
Jun 04, 2026
Ability
★
Reliability
★
Supported Parameters

This model supports the following parameters:

Top P Max Tokens Frequency Penalty Min P Stop Temperature Seed Include Reasoning Logit Bias Presence Penalty Reasoning
Features

This model supports the following features:

Reasoning
Performance Summary

NVIDIA Nemotron 3.5 Content Safety demonstrates exceptional performance in terms of operational efficiency. It consistently ranks among the fastest models, achieving an "Infinityth percentile" across 8 benchmarks, indicating unparalleled speed. Similarly, its pricing is highly competitive, also securing an "Infinityth percentile" ranking across 8 benchmarks, making it an extremely cost-effective solution. However, the model exhibits significant limitations in its current benchmark evaluations for accuracy. Across all tested categories—Hallucinations, Instruction Following, General Knowledge, Coding, Email Classification, Reasoning, Ethics, and Mathematics—the model achieved 0.0% accuracy. This suggests that while the model is highly efficient and affordable, its ability to correctly answer or perform tasks in these domains is currently non-existent based on the provided data. Given its description as a "Content Safety" model, these benchmarks may not fully capture its intended function as a guardrail for LLMs and VLMs. Its strengths lie purely in its speed and cost-effectiveness, while its current benchmark results indicate a complete lack of general task performance.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $0.2
Completion $0.2

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
DeepInfra
DeepInfra | nvidia/nemotron-3.5-content-safety-20260604 131K $0.2 / 1M tokens $0.2 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by nvidia