Nemotron 3.5 Content Safety

NVIDIA · nvidia/nemotron-3.5-content-safety

← Back to leaderboard

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

Nemotron 3.5 Content Safety ranks #100 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.20 blended per million tokens, it is cheaper than 81% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).

open weightsimagetexttext+image->text

Context

Max context: 131072
Max output: 117964

Pricing

Input / 1M: 0.20
Output / 1M: 0.20
Blend / 1M: 0.20

Quality

Quality index:

Provider

Provider: NVIDIA
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.20per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.20
10,000,000$2.00
100,000,000$20.00

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from NVIDIA

Closest API price