Signal Brief

Anthropic 研究:Claude 内部推理工作空间

Anthropic 研究发现 Claude 模型在训练中自然涌现出名为 J-Space 的内部工作空间,不同于链式思维,可以用于直接观察和操纵模型的内部推理过程。

twitter关注列表 elvis (@omarsar0) 发布 2026-07-06 收录 2026-07-06 观察

一句话判断

首次定位模型内部推理的具体位置,而非仅从输出文本推断,值得阅读原文了解方法论。

核心信息

Anthropic 研究发现 Claude 模型在训练中自然涌现出名为 J-Space 的内部工作空间,不同于链式思维,可以用于直接观察和操纵模型的内部推理过程。

原始内容

Must-read research by Anthropic. Here is the simple explanation and why this is a big deal. We suspect LLMs perform "internal reasoning". But little is known or do good methods exist to understand it. Anthropic claims that J-Space (which differs from chain-of-thought or scratchpad), emerged on its own through training and provides a window into how Claude "reasons" internally. In other words, this shows that Claude has a sort of internal workspace where information gets held, combined, and passed between different parts of the model. They can read from it, and they can steer the model by changing it. As it is the case with these reports, the consciousness angle will get all the attention. However, the bigger story is that for the first time you can point to a specific place inside the model where reasoning is staged, rather than guessing at it from the text that comes out. This, of course, changes what interpretability can be. We spent years inferring what a model was doing from what it said. Now there's a mechanism to observe directly, and a direct lever to move. This could enable even more advanced levels of "reasoning" in LLMs and bridges gaps in frontier intelligence and world models. If you can see where a model holds an idea, you can also verify it, audit it, and catch it working toward a goal you never gave it. You can implement better guardrails and predict dangerous/unwanted scenarios better. > **引用原帖 Anthropic (@AnthropicAI):** > New Anthropic research: A global workspace in language models. > Of everything happening in your brain right now, only a tiny fraction is consciously accessible—thoughts you can describe, hold in mind, and reason with. > We found a strikingly similar divide inside Claude. https://t.co/aLUPBifxth > https://x.com/AnthropicAI/status/2074185348142280912

相关动态

02

Kimi K3非对称竞争分析

Kimi K3在Arena前端/WebDev人类盲测中以1679 Elo排名第一,领先Fable 5(1631)和GPT-5.6 Sol(1618),但在综合智能指数中仅排第四(57.1分),落后Fable 5(59.9)和GPT-5.6 Sol(58.9)。K3定价15美元/百万输出tokens,远低于Fable 5的50美元,并计划7...

twitter关注列表2026-07-18#AI#模型#技术突破
观察
03

二年级学生Jo Nagai发现蝴蝶可遗传记忆

日本东京都二年级学生Jo Nagai注意到养育的食蚜 caterpillars 在蝴蝶化后仍保持对薰衣草的回避行为,经Georgetown大学Entomologist Dr. Martha Weiss合作完成实验:70%训练过的蝴蝶及其后代均表现出对薰衣草的遗传性回避,记忆在全变态发育中保存并遗传。

twitter关注列表2026-07-18#研究#技术突破#信息
值得跟进
05

BestBlogs 早报 · 07-18

月之暗面发布Kimi K3,2.8万亿参数,896选16的Stable LatentMoE,上下文100万token,接近Fable-5但仍落后最强闭源模型,完整权重7月27日前开源;VentureBeat调查显示54%企业已发生AI代理安全事件;xAI开源Grok Build(84万行Rust代码)并残留上传用户代码痕迹;Cursor评...

twitter关注列表2026-07-17#AI#模型发布#开源
观察