跳到主内容
@wquguru
精选85Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)模型发布/更新

DeepSeek首次披露推理成本:V4-Pro服务成本至少降3倍

Importantly, DeepSeek has disclosed their inference economics – for the first ti…

原文
发到 X
推荐理由

做推理部署的同学重点关注,DeepSeek首次公开成本数据,V4-Pro性价比大幅提升,建议评估是否替换现有方案。

Importantly, DeepSeek has disclosed their inference economics – for the first time since Open Source Week. Then, they were doing 14.8K generation on 8xH800 node (1,85K/GPU) at 20-22 tps. Whatever these GPUs are now, V4-Pro is *at least 3x cheaper to serve*. 2K/GPU needs >60 tps.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

关联信息,但可能不是同一事件