精选75Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)模型发布/更新
摩尔线程完成236B MoE模型万卡集群预训练
«MoE-236B model, on 25T+ tokens, completed the entire process from pre-training…
«MoE-236B model, on 25T+ tokens, completed the entire process from pre-training to long-context training from scratch on Moore Threads' domestic 10,000-card-class cluster, with the effective training time accounting for over 90%.» Moore Threads getting more serious
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力