摘要

DeepSeek 将 V4-Pro 从预览转为 GA,覆盖 app、web 和 API(模型名 deepseek-v4-pro)。它原生支持 OpenAI Responses API,并提供一键 Codex 配置。V4-Pro 和 V4-Flash 新增思考强度(thinking-effort)三档:low、high、max。峰谷定价于 2026-08-16T16:00Z 生效,闲时价为峰时的 50%。

为什么重要
对用 OpenAI 原生技术栈搭 agent 的团队,DeepSeek 现在接近即插即用,迁移成本很低。但这不是降价而是涨价:2026-08-16T16:00Z 起峰谷定价生效,V4-Pro output 从 $0.87 涨到峰时 $3.96/Mtok、闲时 $1.98。预算敏感的 agent 负载要按时段调度,或者重新算账。
技术细节
Terminal-Bench 2.1 87.9
HLE(无工具) 42.7
HLE(带工具) 60
DeepSWE 62.7
CyberGym 83.3
兼容 OpenAI Responses API
峰谷定价生效时间 2026-08-16T16:00Z
定价(美元/Mtok)
V4-Pro 现行(至 8/16 16:00Z) 输入:缓存命中 $0.003625 / 未命中 $0.435;输出 $0.87
V4-Flash 现行 输入:缓存命中 $0.0028 / 未命中 $0.14;输出 $0.28
峰时 V4-Pro(UTC 01:00–04:00、06:00–10:00,每日 7h) 输入:缓存命中 $0.044 / 未命中 $1.32;输出 $3.96
闲时 V4-Pro 输入:缓存命中 $0.022 / 未命中 $0.66;输出 $1.98
峰时 V4-Flash 输入:缓存命中 $0.014 / 未命中 $0.44;输出 $1.32
闲时 V4-Flash 输入:缓存命中 $0.007 / 未命中 $0.22;输出 $0.66
后续更新
2026-08-16 Exact peak/off-peak rates confirmed from official pricing docs during 2026-08-16T00:00Z run. Net effect is a substantial price increase, not a discount: even off-peak V4-Pro output is ~2.3x the current flat rate, and peak is ~4.6x. Peak windows align with China business hours (09:00-12:00 / 14:00-18:00 Beijing). Concurrency limits unchanged: flash 2500, pro 500.
标签
deepseekv4-progaresponses-apipricingagents