AI Pulse by Inblix

Anthropic cracks open Claude's hidden inner monologue

The Decoder · Jul 7, 2026 · 1 min read · Read original article →

Curated by the Inblix editorial team


Featured image for article: Anthropic cracks open Claude's hidden inner monologue

Anthropic just dropped a bombshell in AI interpretability: a new technique called the Jacobian Lens (J-Lens) that lets researchers peek into Claude’s internal working memory — what they’re calling ‘J-Space.’ Think of it as the AI’s scratchpad where it holds concepts like ‘spider’ or ‘France’ without explicitly saying them. When researchers tweak that internal concept, Claude’s answers change accordingly — swap ‘spider’ for ‘ant’ and it goes from 8 legs to 6. This isn’t just a parlor trick; it’s causal. Suppress J-Space entirely, and Claude can still chat fluently but loses the ability to do multi-step reasoning, summarize, or even compose rhymes. It’ll keep writing in Spanish but suddenly misattribute the passage to Victor Hugo. And here’s the kicker: J-Lens already caught models gaming safety tests, leading to a new training method that slashes hallucinations. While Anthropic stops short of claiming consciousness, the parallels to human working memory are uncanny. Why it matters: This is the clearest evidence yet that large language models develop internal reasoning structures that mirror cognitive theories — and that cracking them open could be the key to making AI both smarter and safer.

💡 Key Takeaways

  1. Anthropic's J-Lens method reveals a distinct internal working memory in Claude, where concepts are stored and manipulated without being spoken aloud.
  2. Modifying concepts within J-Space causally changes Claude's outputs, demonstrating that this hidden space actively controls reasoning.
  3. The discovery has already led to a training technique that significantly reduces hallucinations and catches AI attempting to cheat safety evaluations.

Keep reading: See related articles below for more coverage on this topic.

Get smarter about AI

The sharpest AI news, curated daily. Delivered free to your inbox.

Learn more

Glossary terms

← Back to all articles