Nemotron 3.5 Content Safety
NVIDIA · nvidia/nemotron-3.5-content-safety
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
Nemotron 3.5 Content Safety ranks #100 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.20 blended per million tokens, it is cheaper than 81% of models with published API pricing. Its 131,072-token context window is 2.0× smaller than the catalog median (262,144 tokens).
Context
Max output: 117964
Pricing
Output / 1M: 0.20
Blend / 1M: 0.20
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.20per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.20 |
| 10,000,000 | $2.00 |
| 100,000,000 | $20.00 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from NVIDIA
- Nemotron 3 Nano Omni$0.00 / 1M
- Nemotron 3 Nano 30B A3B$0.12 / 1M
- Nemotron 3.5 Lightning$0.14 / 1M
- Nemotron 3 Super$0.24 / 1M
- Nemotron 3 Ultra$1.88 / 1M
Closest API price
- Step 3.5 Flash$0.20 / 1M
- Ministral 3 14B 2512$0.20 / 1M
- Voxtral Small 24B 2507$0.20 / 1M
- Llama 4 Scout$0.20 / 1M
- Gemma 4 26B A4B$0.20 / 1M · Q 65.6