小红书AI模型IMO满分,超越谷歌DeepMind
A Chinese social app's AI model scored a perfect 42 at the International Mathema…
IMO满分是AI推理能力的里程碑,直接碾压去年谷歌和OpenAI的成绩。做数学推理或竞赛类AI的同学必须关注,这套自我审查和智能体循环的思路值得借鉴。
A Chinese social app's AI model scored a perfect 42 at the International Mathematical Olympiad.
RedNote, the company behind Xiaohongshu, built the system, called dots-note-3.0, still in beta.
Google DeepMind and OpenAI reached 35 out of 42 last year, a gold-level showing.
Only 7 of 666 human contestants in Shanghai matched that result this year.
The IMO asks for written proofs, not final answers, and human graders read every line. A skipped case or an unstated condition costs points, even when the final answer is right.
Most AI maths benchmarks only check the last number, so a lucky guess still scores full credit. That gap is why 42 out of 42 here means more than a high score on a normal test.
This time the model read the original contest documents directly, with no human rewriting. It ran an agentic loop mixing natural-language reasoning with Python code it executed itself.
Drafts got tested, broken, and repaired before anything went to the official graders.
Self-review carries the weight here, since the system questioned its own assumptions repeatedly.
Contest rules barred hints, corrections, and every other form of help during the run.
RedNote made the promise, that the model will be open sourced in due course, but has not committed to any time yet.
The winning model variant is also the smallest in the dots3 family, which includes two larger versions – jazz and aria – tailored for different use cases and compute costs.
---
scmp. com/tech/article/3361482/worlds-first-ai-model-earn-perfect-score-maths-olympiad-comes-chinas-rednote
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力