Anthropic Discovers 'J-Space', a Window Into Claude's Internal Reasoning
Anthropic has published research identifying a component of Claude's internal processing called the J-space, which appears to represent the steps the model works through before generating a response. Despite accounting for only 6–7% of a concept's representation and never exceeding about 10% of model activity at any layer, the J-space plays a critical role in multi-step reasoning. When the J-space is disabled, Claude's complex reasoning breaks down almost entirely, though fluent speech and simple recall remain unaffected. The finding is being described as one of the closest looks yet at how a large language model processes information internally. Anthropic accompanied the research with a video explanation, and the full paper has drawn significant attention from the AI community.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in