DeepSeek-V4-Pro
1.6T参数MoE旗舰,原生百万token上下文,开源模型新标杆。
仅 safetensors · 无 pickle 加载风险
社区实测
社区普遍认为 DeepSeek V4 Pro 以极低成本提供了接近前沿模型的性能,性价比突出;人格感受类似 Claude Opus 4.6,日常编码够用,但在大型复杂代码库和模糊提示下表现下降,Arena 用户偏好基准上口碑不如能力基准亮眼。
- 极高性价比:重度使用 4 小时(含密集子代理循环)仅花费 5 美分
- SWE-bench Verified 得分 80.6%,与 Claude Opus 4.6 仅差 0.2 分
- 输出 token 价格 $3.48/百万,远低于 Claude 的 $25/百万
- 开源权重,可在自有基础设施部署或通过第三方 API 使用
- 给定明确指令时能较好地按指示执行
- 适合复杂代理工作流、深度代码库分析和多步推理任务
- 适合多模型路由方案,将低成本任务分流给 V4 Pro,高难度任务留给闭源前沿模型
- 在聚焦单一端点的逐层深度分析中表现可用
- 对系统指令极为字面化且敏感,容易因提示措辞产生意外行为
- 细节丰富程度不如 Claude Opus
- 代码库规模和复杂度增大后,对模糊提示的容错性显著下降,要求精确描述
- 在 Arena 众包用户偏好基准上表现不佳,用户主观偏好低于能力基准得分
- 在编码工具链中存在低效的随机搜索和 grep 行为,部分归因于 harness
- 基准测试中输出 token 消耗远高于同类开源模型中位数(190M vs 47M),运行成本达 $1,071
- 编码能力存在分歧,部分用户认为 Kimi K2.6 编码更强
Deepseek V4 is mindblowing : r/opencodeCLI - RedditDeepSeek V4 – almost on the frontier | Hacker NewsTo anyone saying deepseek v4 pro is better than opus 4.7, it's a lie.DeepSeek V4 Pro just dropped — is anyone actually using Chinese ...Is DeepSeek V4 Pro already good enough for everyday coding?DeepSeek V4 Alters Everything We Knew About Price-Performance ...DeepSeek V4 Pro: Model Overview, Features & Performance GuideDeepSeek V4: The Open-Source Model That Rivals Closed Frontier ...DeepSeek V4 Pro vs DeepSeek V4 Flash | by Mehul Gupta - MediumDeepSeek V4 Pro underwhelms on Arena (crowdsourced user ...DeepSeek V4 Flash (Reasoning, Max Effort) vs ... - Artificial Analysis
截至 2026-06-22