AGENT PULSESJCPal Special EditionAI Industry Evidence & Trends
Aug 17, 2026 · DeepSeek

DeepSeek V4 Pro 0813 vs Claude Fable 5 on DeepSWE: Cost, Coding, and Routing

What Happened

Together AI 在 2026 年 8 月 17 日发布博客,对比 DeepSeek V4 Pro 0813 与 Claude Fable 5 在 DeepSWE 基准上的表现。他们运行了 904 次 rollout,Fable 在 pass@1 上领先,但成本是 Pro 的 90 倍;Pro 在 pass@4 上胜出,Pro-first 级联达到 82.7%。

EVENT STORY

Development

  1. First ReportDeepSeek V4 Pro 0813 vs Claude Fable 5 on DeepSWE: Cost, Coding, and RoutingTogether AI
  2. Industry ResponseHW for DeepSeek-V4-Flash-0731Reddit r/LocalLLaMA
  3. Industry ResponseGLM-5.3 vs. Claude Fable 5 on DeepSWE: Cost, Coding, and RoutingTogether AI
  4. Industry Response撞名Anthropic的“外挂”刷屏:让“DeepSeek V4‐Pro碾压 Fable 5”但无人能复现,Token开销反而翻倍InfoQ China AI
  5. Industry Responsedeepseek-v4-flash on single DGXReddit r/LocalLLaMA
  6. Industry ResponseQwen3.8-Flash-Next better then DeepSeek V4 ProReddit r/LocalLLaMA
  7. Industry ResponseDeepSeek-V4-Flash vs. GLM-5.3-Flash on 2× DGX SparkReddit r/LocalLLaMA
  8. Industry ResponseDeepSeek-V4-Flash-Vision Q8 vs Qwen3.8-Flash-Next Q8Reddit r/LocalLLaMA
  9. Current AssessmentDeepSeek 的模型在成本上具有显著优势,可能推动更多企业采用其 API 或本地部署。Reddit 用户对 DeepSeek V4 Flash 的积极反馈表明开源模型在本地运行上的吸引力,可能影响闭源模型的市场份额。Agent Pulse · analysis
What Changed

Together AI 的博客报告了 DeepSeek V4 Pro 0813 与 Claude Fable 5 在 DeepSWE 上的对比。Fable 在 pass@1 上领先,但成本是 Pro 的 90 倍;Pro 在 pass@4 上胜出,Pro-first 级联达到 82.7%。Reddit 用户称赞 DeepSeek V4 Flash 与 antirez Dwarfstar 4 的组合,在 Mac Studio M3U 上达到 35 tok/s,但需要 390 GB RAM。

How the Capability Boundary Shifted

DeepSeek V4 Pro 0813 在 pass@4 上优于 Claude Fable 5,表明其采样多样性或探索能力更强,而 Fable 的 pass@1 优势可能来自更好的单次生成质量。Pro-first 级联达到 82.7%,暗示路由策略可以平衡成本与性能。

Why It Matters

DeepSeek 的模型在成本上具有显著优势,可能推动更多企业采用其 API 或本地部署。Reddit 用户对 DeepSeek V4 Flash 的积极反馈表明开源模型在本地运行上的吸引力,可能影响闭源模型的市场份额。

Who It Affects

对于依赖代码生成的企业,DeepSeek V4 Pro 提供了更低成本的替代方案,可能降低软件开发成本。Pro-first 级联策略展示了如何通过路由优化成本,具有商业价值。

What to Watch Next

未来可关注 DeepSeek 是否发布更多关于路由策略的细节,以及 Claude Fable 5 是否通过降价或优化来应对成本劣势。