Signal Brief

RedNote AI模型IMO满分42分

RedNote(小红书)的AI模型dots-note-3.0在国际数学奥林匹克竞赛中获得满分42分,而Google DeepMind和OpenAI去年仅得35分。该模型直接阅读原题,使用自然语言推理与自执行Python代码的智能体循环,并具备自我审查机制,最终在禁止外部帮助的条件下取得满分。模型为d...

twitter关注列表 Rohan Paul (@rohanpaul_ai) 发布 2026-07-23 收录 2026-07-23 观察

一句话判断

该模型首次在IMO中取得满分,远超此前DeepMind和OpenAI的35分,其自检查机制和直接处理原题能力是技术亮点,值得关注开源进展。

核心信息

RedNote(小红书)的AI模型dots-note-3.0在国际数学奥林匹克竞赛中获得满分42分,而Google DeepMind和OpenAI去年仅得35分。该模型直接阅读原题,使用自然语言推理与自执行Python代码的智能体循环,并具备自我审查机制,最终在禁止外部帮助的条件下取得满分。模型为dots3系列最小版本,承诺未来开源。

原始内容

https://x.com/rohanpaul_ai/status/2080188865470693745 > **引用原帖 Rohan Paul (@rohanpaul_ai):** > A Chinese social app's AI model scored a perfect 42 at the International Mathematical Olympiad. > RedNote, the company behind Xiaohongshu, built the system, called dots-note-3.0, still in beta. > Google DeepMind and OpenAI reached 35 out of 42 last year, a gold-level showing. > Only 7 of 666 human contestants in Shanghai matched that result this year. > The IMO asks for written proofs, not final answers, and human graders read every line. A skipped case or an unstated condition costs points, even when the final answer is right. > Most AI maths benchmarks only check the last number, so a lucky guess still scores full credit. That gap is why 42 out of 42 here means more than a high score on a normal test. > This time the model read the original contest documents directly, with no human rewriting. It ran an agentic loop mixing natural-language reasoning with Python code it executed itself. > Drafts got tested, broken, and repaired before anything went to the official graders. > Self-review carries the weight here, since the system questioned its own assumptions repeatedly. > Contest rules barred hints, corrections, and every other form of help during the run. > RedNote made the promise, that the model will be open sourced in due course, but has not committed to any time yet. > The winning model variant is also the smallest in the dots3 family, which includes two larger versions – jazz and aria – tailored for different use cases and compute costs. > --- > scmp. com/tech/article/3361482/worlds-first-ai-model-earn-perfect-score-maths-olympiad-comes-chinas-rednote > https://x.com/rohanpaul_ai/status/2080188865470693745

相关动态

01

GPT-5.6 Sol解决6个Erdős问题

一位研究员使用GPT-5.6 Sol在5天内解决了6个开放Erdős问题(尝试13个,成功率46%)。他采用合同式提示,明确证明要求、排除弱结果,并指定搜索策略,包括同时探索多种路径和对抗性检查。

twitter关注列表2026-07-23#技术突破#模型#研究
观察
03

RedNote AI 模型获 IMO 满分 42/42

RedNote(小红书母公司)旗下 AI 模型 dots-note-3.0 在国际数学奥林匹克竞赛(IMO)中获得 42/42 满分,成为首个达成此成就的 AI 系统。该模型直接阅读原始竞赛文档,通过结合自然语言推理与 Python 代码执行的代理循环,经草稿测试、破坏、修复及自我审查后提交答案,全程无人工改写或提示。对比之下,Googl...

twitter关注列表2026-07-23#AI#技术突破#大模型
值得跟进