精选70Przemek Chojecki | PC模型发布/更新
Meta Muse Spark 1.2 数学基准超 GPT-5.5
Muse Spark 1.2 is better than GPT-5.5 xhigh and only slightly worse than Kimi K3…
Muse Spark 1.2 is better than GPT-5.5 xhigh and only slightly worse than Kimi K3 on our ErdosBench.
We've tested the new model from Meta on 226 research-level math problems and it solved 40 / 226 problems and gave many interesting partial solutions.
Muse Spark 1.2 is a strong entrant: good proof hygiene, high B-grade review yield, no rejected strong claims, but fewer decisive A-grade closures.
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力