跳到主内容
@wquguru
精选75The Decoder(RSS)模型发布/更新

OpenAI称GPT-5.6 Sol在ARC-AGI-3上超越Opus 5

OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 but only with its own custom test harness

原文
发到 X

OpenAI counters Anthropic's ARC-AGI-3 record: GPT-5.6 Sol scores 38.3 percent, but only through its own API with retained reasoning and context compaction. In the official test environment, the model managed just 7.8 percent. Opus 5 hit its 30.2 percent without such aids. The article OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 but only with its own custom test harness appeared first on The Decoder.

更进一步:量化金融体系

看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力

进入量化体系 →

关联讨论

同一事件的更多信源
Anthropic质疑OpenAI ARC-AGI-3得分异常
Peter Steinberger 🦞原文
向Anthropic喊话:轮到你了
Chubby♨️原文

相似阅读

另一事件,读法相近