综合智能↑ Top 9%
40.3第 51 / 567
编程指数↑ Top 25%
56.2第 47 / 190
首 Token 延迟− Top 27%
1.23s端到端 52.25s
输出价格↑ Cost-effective
$0.280/M输入 $0.140/M
DeepSeek V4 Flash (Reasoning, Max Effort) 由 DeepSeek 发布。更偏向成本敏感或特定任务场景。输出约 $0.280/M tokens,性价比突出。
模型信息
- 发布日期
- 2026年4月24日
- 研发机构
- DeepSeek
- 智能体指数
- 31.1
- Slug
- deepseek-v4-flash
经济与性能
- 输入价格 (1M)
- $0.140
- 输出价格 (1M)
- $0.280
- 缓存命中 (1M)
- $0
- 生成速度 (TPS)
- 120 tok/s
- 端到端响应
- 52.25s
深度分析
相对全库 580 款模型的能力分解、性价比位置与相近对照。
三项核心指数
综合智能排名 #51 / 567
综合智能40.3/ 中位 14.8
编程指数56.2/ 中位 36.3
智能体指数31.1/ 中位 18.2
金色竖线为全库中位数,便于判断该模型相对行业基准的位置。
DeepSeek 其他模型
8 款
DeepSeek V4 Pro (Reasoning, Max Effort)44.3DeepSeek V4 Pro (Reasoning, High Effort)43.1DeepSeek V4 Flash (Reasoning, High Effort)37.5DeepSeek V3.2 (Reasoning)32DeepSeek V4 Pro (Non-reasoning)31.2DeepSeek V3.1 Terminus (Reasoning)30.4DeepSeek V4 Flash (Non-reasoning)28.7DeepSeek V3.2 Exp (Reasoning)25.4
评测任务成本$0.0223每任务均摊
评测总成本$74Intelligence Index 全量
输出单价$0.280/M输入 $0.140/M
输出价格最低对比
Top 9 + 当前模型 · 金色为当前模型
1Gemma 3n E4B Instruct
$0.040
2Llama 3.1 Instruct 8B
$0.090
3Gemma 4 E4B (Reasoning)
$0.100
4Gemma 4 E4B (Non-reasoning)
$0.100
5Ministral 3 3B
$0.100
6Granite 4.1 8B
$0.100
7Sarvam 30B (high)
$0.110
8HyperNova 60B 2605
$0.140
9Nova Micro
$0.140
10DeepSeek V4 Flash (Reasoning, Max Effort)
$0.280
生成速度最高对比
Top 9 + 当前模型 · 金色为当前模型
1Mercury 2
958 TPS
2Gemini 3.5 Flash-Lite
450 TPS
3HyperNova 60B 2605
421 TPS
4Granite 4.0 H Small
411 TPS
5LFM2.5-VL-1.6B
400 TPS
6Step 3.7 Flash
391 TPS
7Granite 3.3 8B (Non-reasoning)
344 TPS
8LFM2.5-8B-A1B
331 TPS
9gpt-oss-120b (low)
314 TPS
10DeepSeek V4 Flash (Reasoning, Max Effort)
120 TPS
年度智能峰值
按发布时间排列 · 金色为当前模型
GPT-3.5 Turbo2022 年发布
3.6
GPT-4 Turbo2023 年发布
7.9
o12024 年发布
23.4
GPT-5.2 (xhigh)2025 年发布
42.2
DeepSeek V4 Flash (Reasoning, Max Effort)2026 年发布
40.3
Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)2026 年发布
59.9
相近智能水平模型
智能指数 ±8 · 基准 40.3
| 模型 | 智能 | 编程 | 输出价 | 首 Token |
|---|---|---|---|---|
DeepSeek V4 Flash (Reasoning, Max Effort)当前 | 40.3 | 56.2 | $0.280 | 1.23s |
| MiMo-V2-Pro | 40.3 | — | — | — |
| GLM-5.1 (Reasoning) | 40.2 | 55.8 | $4.4 | 1.56s |
| GPT-5.2 Codex (xhigh) | 40.1 | — | $14 | 35.66s |
| GPT-5.6 Terra (low) | 40.5 | 58.1 | $15 | 1.33s |
| Qwen3.6 Max Preview | 40 | — | $7.8 | 3.51s |
| GPT-5.4 mini (xhigh) | 40 | 56.1 | $4.5 | 6.29s |
| Inkling (xhigh) | 40.7 | 52.1 | $4.68 | 1.87s |
| Claude Opus 4.5 (Reasoning) | 40.8 | — | $25 | 14.17s |
| Grok Build 0.1 0616 | 39.8 | 51.5 | $2 | 550ms |
| Gemini 3 Pro Preview (high) | 39.6 | — | $12 | — |