摘要
DeepSeek 将 V4-Pro 从预览转为 GA,覆盖 app、web 和 API(模型名 deepseek-v4-pro)。它原生支持 OpenAI Responses API,并提供一键 Codex 配置。V4-Pro 和 V4-Flash 新增思考强度(thinking-effort)三档:low、high、max。峰谷定价于 2026-08-16T16:00Z 生效,闲时价为峰时的 50%。
为什么重要
对用 OpenAI 原生技术栈搭 agent 的团队,DeepSeek 现在接近即插即用,迁移成本很低。但这不是降价而是涨价:2026-08-16T16:00Z 起峰谷定价生效,V4-Pro output 从 $0.87 涨到峰时 $3.96/Mtok、闲时 $1.98。预算敏感的 agent 负载要按时段调度,或者重新算账。
技术细节
| Terminal-Bench 2.1 | 87.9 | ||||||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| HLE(无工具) | 42.7 | ||||||||||||
| HLE(带工具) | 60 | ||||||||||||
| DeepSWE | 62.7 | ||||||||||||
| CyberGym | 83.3 | ||||||||||||
| 兼容 OpenAI Responses API | ✓ | ||||||||||||
| 峰谷定价生效时间 | 2026-08-16T16:00Z | ||||||||||||
| 定价(美元/Mtok) |
|
后续更新
2026-08-16 Exact peak/off-peak rates confirmed from official pricing docs during 2026-08-16T00:00Z run. Net effect is a substantial price increase, not a discount: even off-peak V4-Pro output is ~2.3x the current flat rate, and peak is ~4.6x. Peak windows align with China business hours (09:00-12:00 / 14:00-18:00 Beijing). Concurrency limits unchanged: flash 2500, pro 500.
标签
deepseekv4-progaresponses-apipricingagents