NVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)
- Released
- Dec 15, 2025
- Tracked since
- 2026-07-23
Intelligence index6.8Ranks 519 of 664 models with this metric
Coding index—Not tested
Agentic index—Not tested
Capability profile
5 axes vs. site averageNVIDIA Nemotron 3 Nano 30B A3B (Non-reasoning)
Site average
Some dimensions have no data yet — marked on the chart
Pricing
per 1M tokens- Input
- $0.050
- Output
- $0.200
- Cache read
- $0.050
Performance
- Output throughput
- 245 tokens/sec
- Time to first token
- 700ms
- First answer token
- 700ms
- End-to-end response
- 2.75s
History
2Intelligence index
7.4→6.8
Output
Unchanged across 62 days of records
Output throughput
103→245
Cheaper at this level
Models with a comparable intelligence index and a lower output price- Llama 3.1 Instruct 8B
MetaIntelligence 6.9$0.050−75%
- Mistral Small 3
MistralIntelligence 6.7$0.080−60%
- Gemma 4 E4B (Reasoning)
GoogleIntelligence 8.9$0.100−50%
- Gemma 4 E4B (Non-reasoning)
GoogleIntelligence 7.5$0.100−50%
- Granite 4.1 8B
IBMIntelligence 6.6$0.100−50%