跳到主内容
精选85Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)模型发布/更新

DeepSeek首次披露推理成本:V4-Pro服务成本至少降3倍

Importantly, DeepSeek has disclosed their inference economics – for the first ti…

原文
推荐理由

做推理部署的同学重点关注,DeepSeek首次公开成本数据,V4-Pro性价比大幅提升,建议评估是否替换现有方案。

Importantly, DeepSeek has disclosed their inference economics – for the first time since Open Source Week. Then, they were doing 14.8K generation on 8xH800 node (1,85K/GPU) at 20-22 tps. Whatever these GPUs are now, V4-Pro is *at least 3x cheaper to serve*. 2K/GPU needs >60 tps.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近