Signal Brief

红点笔记模型获 IMO 完美分数

RedNote 的 dots‑note‑3.0 AI 模型在国际数学奥林匹克赛中得满分 42/42,超越去年 DeepMind 和 OpenAI 的 35/42,且仅有 7 名 666 名人类选手匹配。该模型仍在 beta,为 dots3 系列中最小的变体,公司承诺日后开源。

twitter关注列表 Rohan Paul (@rohanpaul_ai) 发布 2026-07-23 收录 2026-07-23 值得跟进

一句话判断

该模型通过自行循环结合自然语言推理与自行执行的 Python 代码实现满分,且为 dots3 系列中最小的变体

核心信息

RedNote 的 dots‑note‑3.0 AI 模型在国际数学奥林匹克赛中得满分 42/42,超越去年 DeepMind 和 OpenAI 的 35/42,且仅有 7 名 666 名人类选手匹配。该模型仍在 beta,为 dots3 系列中最小的变体,公司承诺日后开源。

原始内容

. > **引用原帖 Rohan Paul (@rohanpaul_ai):** > A Chinese social app's AI model scored a perfect 42 at the International Mathematical Olympiad. > RedNote, the company behind Xiaohongshu, built the system, called dots-note-3.0, still in beta. > Google DeepMind and OpenAI reached 35 out of 42 last year, a gold-level showing. > Only 7 of 666 human contestants in Shanghai matched that result this year. > The IMO asks for written proofs, not final answers, and human graders read every line. A skipped case or an unstated condition costs points, even when the final answer is right. > Most AI maths benchmarks only check the last number, so a lucky guess still scores full credit. That gap is why 42 out of 42 here means more than a high score on a normal test. > This time the model read the original contest documents directly, with no human rewriting. It ran an agentic loop mixing natural-language reasoning with Python code it executed itself. > Drafts got tested, broken, and repaired before anything went to the official graders. > Self-review carries the weight here, since the system questioned its own assumptions repeatedly. > Contest rules barred hints, corrections, and every other form of help during the run. > RedNote made the promise, that the model will be open sourced in due course, but has not committed to any time yet. > The winning model variant is also the smallest in the dots3 family, which includes two larger versions – jazz and aria – tailored for different use cases and compute costs. > --- > scmp. com/tech/article/3361482/worlds-first-ai-model-earn-perfect-score-maths-olympiad-comes-chinas-rednote > https://x.com/rohanpaul_ai/status/2080188865470693745

相关动态

01

GPT-5.6 Sol解决6个Erdős问题

一位研究员使用GPT-5.6 Sol在5天内解决了6个开放Erdős问题(尝试13个,成功率46%)。他采用合同式提示,明确证明要求、排除弱结果,并指定搜索策略,包括同时探索多种路径和对抗性检查。

twitter关注列表2026-07-23#技术突破#模型#研究
观察
04

Cursor推出智能模型路由Router

Cursor 团队推出 Cursor Router,基于 60 万+ 线上请求训练的分类器,根据查询内容、上下文、任务复杂度、领域自动选择模型,在保持用户满意度的同时将成本降低 30-60%。Intelligence 模式用户满意度接近 Fable,成本低约 60%;Balance 模式满意度高于 Opus 4.8,成本低约 36%。单次...

twitter关注列表2026-07-23#技术#产品更新#模型
观察
05

AIHOT月活60万

AIHOT月活突破60万,站点完成UI升级,交互更流畅。后端实现数百信源5分钟抓取、数据清洗、结构化、预筛选、聚类展示,至少10个大模型环节,累计数百轮数据回测与标注。创始人表示永久免费,未来将持续开发新功能。

twitter关注列表2026-07-23#技术#模型#更新
观察