跳到主内容
@wquguru
精选60Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)行业动态

后训练算力消耗成谜:V4-Pro 级模型预训练占比引猜测

Does anyone have a clue of how much post-training compute is being used right no…

原文
发到 X

Does anyone have a clue of how much post-training compute is being used right now? V4-Pro is ≈1e25 class model (as are its peers). Over 2 months, could they have spent another 1e25 on rollouts? More? What is the pretraining share at this point?

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近