Nemotron 3 Super
NVIDIA · nvidia/nemotron-3-super-120b-a12b
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
Nemotron 3 Super ranks #108 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.24 blended per million tokens, it is cheaper than 79% of models with published API pricing. Its 262,144-token context window is about the catalog median (262,144 tokens).
Context
Max output: 16384
Pricing
Output / 1M: 0.40
Blend / 1M: 0.24
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.24per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.24 |
| 10,000,000 | $2.42 |
| 100,000,000 | $24.25 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from NVIDIA
- Nemotron 3 Nano Omni$0.00 / 1M
- Nemotron 3 Nano 30B A3B$0.12 / 1M
- Nemotron 3.5 Lightning$0.14 / 1M
- Nemotron 3.5 Content Safety$0.20 / 1M
- Nemotron 3 Ultra$1.88 / 1M
Closest API price
- Seed-2.0-Mini$0.25 / 1M
- Gemini 2.5 Flash Lite$0.25 / 1M
- GPT-4.1 Nano$0.25 / 1M
- GLM 4.7 Flash$0.23 / 1M
- GPT-5 Nano$0.22 / 1M