跳到主内容
@wquguru
精选70Przemek Chojecki | PC模型发布/更新

Meta Muse Spark 1.2 数学基准超 GPT-5.5

Muse Spark 1.2 is better than GPT-5.5 xhigh and only slightly worse than Kimi K3…

原文
发到 X

Muse Spark 1.2 is better than GPT-5.5 xhigh and only slightly worse than Kimi K3 on our ErdosBench.

We've tested the new model from Meta on 226 research-level math problems and it solved 40 / 226 problems and gave many interesting partial solutions.

Muse Spark 1.2 is a strong entrant: good proof hygiene, high B-grade review yield, no rejected strong claims, but fewer decisive A-grade closures.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

相似阅读

另一事件,读法相近