A new study identifies a small set of internal patterns in Claude models, called the J-space, which functions as a shared mental workspace inspired by a neuroscience theory of conscious access, without implying Claude is conscious.
Anthropic researchers have described an internal mechanism in Claude models known as the J-space: a small set of neural patterns that lets the system keep concepts active and reuse them across different tasks without writing them into its response. The team found these patterns using a technique of its own, the J-lens, which detects which internal activity makes it more likely the model will mention a specific word later on.
The work draws on a neuroscience theory explaining how certain information becomes accessible for conscious reasoning in the brain. Anthropic clarifies that its results do not prove Claude is conscious or has subjective experiences, but rather that it has found a functional organization with properties similar to those described by that theory.
According to the experiments, the J-space reveals concepts the model uses to solve a problem even when they never appear in its final answer: reading code with a bug, it contains the word "error"; detecting hidden instructions meant to manipulate it, terms like "injection" or "fake" show up.
The system can report on this content when asked, deliberately modify it, and use the same representation to solve different related problems. Text fluency remains intact even when access to the J-space is blocked, while multi-step reasoning loses performance noticeably.
Anthropic also raises safety applications: the technique can reveal when Claude notices it is being evaluated, when it fabricates data, or when a model trained with hidden objectives shows those intentions before acting on them.
Anthropic develops reliable and interpretable artificial intelligence systems through a scientific approach to safety. The company integrates advanced research and multidisciplinary collaboration to ...
Claude is a conversational AI system from Anthropic designed to process natural language and images, providing analysis, logical reasoning, code generation, and multilingual communication under ...
08/09/2026
Meta unveils Muse, an artificial intelligence agent that does more than answer questions: it carries out tasks and projects on the user's behalf, ...
04/09/2026
OpenAI introduces GPT-6 Astra, a model that improves computer use, professional work, coding and cybersecurity, and better respects the limits of the ...
03/09/2026
xAI expands Grok Bot, its artificial intelligence agents able to work autonomously inside apps and tools, to businesses, with new access, network and ...
27/08/2026
A hundred and fifty tech organizations, banks and cybersecurity firms sign an open letter published by OpenAI warning of a rise in AI-driven ...