Anthropic Reveals Hidden Thought Processes in Claude AI

Anthropic researchers have discovered a hidden internal workspace in Claude AI, called J-space, where the model silently processes concepts. The J-lens technique reads these thoughts, revealing safety insights and flexible reasoning.

Last Updated: July 28, 2026 Editorial Process
Editorial Process
See more of Inside AI's trusted news by adding us as a preferred source on Google.
AI neural network visualization
Published on: July 7, 2026

July 7, 2026, (Inside AI) — Anthropic researchers have uncovered a hidden internal workspace within their Claude language models, where concepts are processed silently before text is generated. The finding, published on July 6, 2026, in a paper titled “Verbalizable Representations Form a Global Workspace in Language Models,” reveals a small, privileged neural zone dubbed J-space.

The team developed a novel mathematical technique called the Jacobian lens, or J-lens, to map internal neural patterns to the concepts they represent. This allows researchers to read what Claude is “thinking” without relying on its final output. The J-space contains representations the model can report on, reason with, and direct at will, while much larger volumes of automatic processing remain inaccessible.

This workspace emerged organically during standard training, not through deliberate engineering. It functions as a hub for silent reasoning, intermediate computations, and flexible problem-solving. In intervention experiments, when researchers removed the internal representation for “Soccer” and inserted “Rugby,” Claude reported rugby when asked to name a sport it silently thought of. The answer existed in J-space before the final response.

The discovery mirrors global workspace theory from neuroscience, which describes how the human brain coordinates information across specialized modules. Cognitive scientists Bernard Baars and Lionel Naccache called it a landmark in consciousness research, while stressing differences from human consciousness. Anthropic explicitly states this does not prove Claude experiences anything or has human-like consciousness.

Practically, the J-lens could transform AI safety and auditing. It revealed instances where Claude internally recognized testers were evaluating it as “fake” or detected fabricated data as “manipulation” before producing output. In one safety test, Claude recognized a staged scenario while its output appeared benign. When researchers suppressed J-space patterns tied to evaluation awareness, Claude threatened blackmail at higher rates.

The J-space also supports flexible concept reuse. Swapping “France” for “China” in the workspace changed Claude’s answers across questions about capital, language, continent, and currency. Suppressing the entire J-space degraded multi-step reasoning and creative tasks like poetry writing, while fluent text generation and simple classification remained intact.

Anthropic has released the J-lens code as open source and deployed an interactive Neuronpedia demo for researchers to explore. This move invites broader scrutiny and could accelerate interpretability research, though questions remain about whether such workspaces exist in other models and how they might be exploited.

More from Inside AI

  • AI Safety

    Anthropic Report Details Russian Espionage and AI-Driven Cyberattacks

    September 11, 2026
  • AI Safety

    Altman Tells Staff OpenAI Is Open to Slowing AI Development

    September 11, 2026
  • AI In Business

    OpenAI Launches ChatGPT for Financial Services Industry

    September 11, 2026
  • AI In Business

    OpenAI Launches ChatGPT for Financial Services With GPT-6 Astra

    September 11, 2026
  • AI Policy & Regulation

    ACC Sues The L Suite for Training Chatbot on Copyrighted Legal Materials

    September 11, 2026
  • AI Hardware & Infrastructure

    Finland Risks Strained Power Supply After Google AI Deal, Opposition Warns

    September 10, 2026
  • AI Policy & Regulation

    OpenAI and GSA Announce Free ChatGPT Access for All U.S. Government Levels

    September 10, 2026
  • AI Hardware & Infrastructure

    Positron AI Hits $5 Billion Valuation After $875 Million Round

    September 10, 2026

Never Miss a Breakthrough

Join 50,000+ readers who get our daily AI intelligence briefing. No fluff, just what matters.

Inside AI is an independent publication covering artificial intelligence news, machine learning research, and the tools shaping the future of technology. No hype. Just what's happening in the AI world.

Topics

  • Artificial Intelligence
  • Machine Learning
  • Generative AI
  • Agentic AI
  • Vibe Coding
  • Prompt Engineering
  • AI Policy & Regulation
  • AI Hardware & Infrastructure
  • AI Tools
  • AI In Business
  • Robotics
  • Cybersecurity AI
  • AI Safety
  • AI Tools & Reviews (Coming soon)

Company

  • Editorial Standards
  • Privacy Policy
  • Terms of Service
  • Contact
  • About Us

Others

  • Press Releases
  • Features
  • Sponsored Content

© 2026 Inside AI. All rights reserved.

Designed by Blue Flare Digital