精选85Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)模型发布/更新
DeepSeek首次披露推理成本:V4-Pro服务成本至少降3倍
Importantly, DeepSeek has disclosed their inference economics – for the first ti…
推荐理由
做推理部署的同学重点关注,DeepSeek首次公开成本数据,V4-Pro性价比大幅提升,建议评估是否替换现有方案。
Importantly, DeepSeek has disclosed their inference economics – for the first time since Open Source Week. Then, they were doing 14.8K generation on 8xH800 node (1,85K/GPU) at 20-22 tps. Whatever these GPUs are now, V4-Pro is *at least 3x cheaper to serve*. 2K/GPU needs >60 tps.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力