Signal Brief

RedNote AI 模型获 IMO 满分 42/42

RedNote(小红书母公司)旗下 AI 模型 dots-note-3.0 在国际数学奥林匹克竞赛(IMO)中获得 42/42 满分,成为首个达成此成就的 AI 系统。该模型直接阅读原始竞赛文档,通过结合自然语言推理与 Python 代码执行的代理循环,经草稿测试、破坏、修复及自我审查后提交答案,全...

twitter关注列表 Rohan Paul (@rohanpaul_ai) 发布 2026-07-23 收录 2026-07-23 值得跟进

一句话判断

首个在 IMO 正式赛制下获满分的 AI,其无人工改写原始文档、代理循环自我修复的方法论区别于传统仅验证最终答案的基准,值得阅读原文了解技术细节并跟踪开源进展。

核心信息

RedNote(小红书母公司)旗下 AI 模型 dots-note-3.0 在国际数学奥林匹克竞赛(IMO)中获得 42/42 满分,成为首个达成此成就的 AI 系统。该模型直接阅读原始竞赛文档,通过结合自然语言推理与 Python 代码执行的代理循环,经草稿测试、破坏、修复及自我审查后提交答案,全程无人工改写或提示。对比之下,Google DeepMind 和 OpenAI 去年最佳成绩为 35/42,本届上海 666 名人类参赛者仅 7 人满分。RedNote 承诺后续开源该模型,获胜变体为 dots3 系列中最小版本,另有 jazz 和 aria 两个更大版本。

原始内容

A Chinese social app's AI model scored a perfect 42 at the International Mathematical Olympiad. RedNote, the company behind Xiaohongshu, built the system, called dots-note-3.0, still in beta. Google DeepMind and OpenAI reached 35 out of 42 last year, a gold-level showing. Only 7 of 666 human contestants in Shanghai matched that result this year. The IMO asks for written proofs, not final answers, and human graders read every line. A skipped case or an unstated condition costs points, even when the final answer is right. Most AI maths benchmarks only check the last number, so a lucky guess still scores full credit. That gap is why 42 out of 42 here means more than a high score on a normal test. This time the model read the original contest documents directly, with no human rewriting. It ran an agentic loop mixing natural-language reasoning with Python code it executed itself. Drafts got tested, broken, and repaired before anything went to the official graders. Self-review carries the weight here, since the system questioned its own assumptions repeatedly. Contest rules barred hints, corrections, and every other form of help during the run. RedNote made the promise, that the model will be open sourced in due course, but has not committed to any time yet. The winning model variant is also the smallest in the dots3 family, which includes two larger versions – jazz and aria – tailored for different use cases and compute costs. --- scmp. com/tech/article/3361482/worlds-first-ai-model-earn-perfect-score-maths-olympiad-comes-chinas-rednote ![photo](https://pbs.twimg.com/media/HN5PQ35aQAACP2u.png) > **引用原帖 Rohan Paul (@rohanpaul_ai):** > And this comes from a theoretical physicist. https://t.co/0FRml3vJkh > https://x.com/rohanpaul_ai/status/2080185447788224602

相关动态

01

GPT-5.6 Sol解决6个Erdős问题

一位研究员使用GPT-5.6 Sol在5天内解决了6个开放Erdős问题(尝试13个,成功率46%)。他采用合同式提示,明确证明要求、排除弱结果,并指定搜索策略,包括同时探索多种路径和对抗性检查。

twitter关注列表2026-07-23#技术突破#模型#研究
观察
03

RedNote AI模型IMO满分42分

RedNote(小红书)的AI模型dots-note-3.0在国际数学奥林匹克竞赛中获得满分42分,而Google DeepMind和OpenAI去年仅得35分。该模型直接阅读原题,使用自然语言推理与自执行Python代码的智能体循环,并具备自我审查机制,最终在禁止外部帮助的条件下取得满分。模型为dots3系列最小版本,承诺未来开源。

twitter关注列表2026-07-23#技术突破#大模型#AI模型
观察