跳到主内容
@wquguru
精选60Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)模型发布/更新

Grok 4.20 超越 Opus 4.8,Kimi K2.5 超 GLM 5.2,基准测试引争议

Honestly this makes the whole benchmark look even more absurd. Grok 4.20 over Op…

原文
发到 X

Honestly this makes the whole benchmark look even more absurd. Grok 4.20 over Opus 4.8 (max), Kimi K2.5 > GLM 5.2 and Opus 4.7, Opus 4.6 down in the dumps below Grok 4… what is going on here? Sounds like it's super sensitive to lab priorities in this domain.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源

相似阅读

另一事件,读法相近