DeepSeek V4.1 Flash (Non-Reasoning)
- Released
- Sep 10, 2026
- Tracked since
- 2026-09-25
Intelligence index24.7Ranks 148 of 664 models with this metric
Coding index—Not tested
Agentic index—Not tested
Capability profile
5 axes vs. site averageDeepSeek V4.1 Flash (Non-Reasoning)
Site average
Some dimensions have no data yet — marked on the chart
Pricing
per 1M tokens- Input
- $0.300
- Output
- $1.20
- Cache read
- $0.010
Performance
- Output throughput
- 272 tokens/sec
- Time to first token
- 1.19s
- First answer token
- 1.19s
- End-to-end response
- 3.03s
History
1Intelligence index
Unchanged across 62 days of records
Output
Unchanged across 62 days of records
Output throughput
243→272
Cheaper at this level
Models with a comparable intelligence index and a lower output price- Ling 3.0 Flash
InclusionAIIntelligence 24.9$0.220−82%
- Ling-3.0-flash-VL
InclusionAIIntelligence 24.6$0.220−82%
- DeepSeek V4 Flash 0420 (Reasoning, High Effort)
DeepSeekIntelligence 26$0.280−77%
- MiMo-V2.5
XiaomiIntelligence 25.2$0.280−77%
- DeepSeek V4 Flash 0420 (Reasoning, Max Effort)
DeepSeekIntelligence 24.2$0.280−77%