跳到主内容
@wquguru
精选75Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)模型发布/更新

摩尔线程完成236B MoE模型万卡集群预训练

«MoE-236B model, on 25T+ tokens, completed the entire process from pre-training…

原文
发到 X

«MoE-236B model, on 25T+ tokens, completed the entire process from pre-training to long-context training from scratch on Moore Threads' domestic 10,000-card-class cluster, with the effective training time accounting for over 90%.» Moore Threads getting more serious

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近