Anthropic found Claude reasoning in silence (J-space) — we ran the same lens on open Qwen3-8B
Summary
Anthropic discovered silent reasoning in Claude's activations (J-space). The author applied the same Jacobian lens to Qwen3-8B locally, using it to detect prose drift before tool calls and implement agent guards.
Similar Articles
Anthropic found a hidden space where Claude puzzles over concepts
Anthropic developed the Jacobian lens (J-lens) to reveal a hidden 'J-space' inside Claude Opus 4.6, offering unprecedented insight into an LLM's internal reasoning process before it outputs tokens. The technique allows monitoring and control of model behavior by surfacing the words the model is about to produce.
@rohanpaul_ai: Another massive research from Anthropic. New “J-lens” uncovers Claude’s quiet workspace, matching a major consciousness…
Anthropic's new research introduces 'J-lens,' a method to read Claude's internal activations before output, revealing a quiet workspace functionally similar to human global workspace theory. This allows detection of hidden reasoning, goals, and potential safety issues like prompt injections.
What Anthropic’s latest AI discovery does—and doesn’t—show
Anthropic discovered a hidden internal space (J-space) in LLMs like Claude that contains words influencing reasoning, advancing understanding of AI model internals.
Anthropic research - A global workspace in language models
Anthropic's new paper presents evidence that modern language models like Claude have developed a 'global workspace' (J-space) of internal neural patterns that are reportable, controllable, and used for flexible reasoning, distinct from automatic processing.
@omooretweets: I often find in Claude’s reasoning traces it will have thoughts / opinions it does not say …but then it denies having t…
Anthropic releases research on a global workspace in language models, prompted by observations that Claude has unspoken thoughts in reasoning traces.