Anthropic discovers Claude’s J-space, an internal AI workspace for deliberate reasoning

Anthropic says the J-space workspace is defined by a set of five behaviors that enable the model to read, write, and coordinate information internally, distinguishing it from mere automatic processing.
Disabling access to the J-space significantly impairs Claude's higher-order reasoning, suggesting the internal workspace plays a functional role in deliberate problem-solving rather than surface-level pattern matching.
Google DeepMind researcher Neel Nanda independently replicated the core claims on open-weight models (Qwen 3.6 27B), providing cross-group validation of the J-space findings.
J-space consists of internal neural patterns tied to specific words; when a pattern lights up, the model essentially has that word 'on its mind' even if it doesn’t output it, reflecting a 'brain-like' internal representation.
Anthropic has released an open-source Jacobian-lens repository (Apache-2.0) and showcased a live demo on Neuronpedia, under the project name anthropics/jacobian-lens.
Anthropic has found a hidden internal workspace inside Claude, its AI model, that it calls J-space. The workspace holds about 25 concepts at any given time and accounts for less than 10% of Claude's internal activity, according to Windows Report. Researchers say it is where Claude does its deliberate, step-by-step thinking — separate from the automatic pattern-matching that makes up most of what the model does.
The discovery was made using a tool called the Jacobian lens, or J-lens. Anthropic has released the code as open-source software under the project name anthropics/jacobian-lens, reports WebProNews. Google DeepMind researcher Neel Nanda independently confirmed similar findings on a different model, Qwen 3.6 27B, adding outside credibility to the claim.
J-space is a set of internal neural patterns inside Claude. Think of it like a mental whiteboard. When a pattern "lights up," the model has a word or concept "on its mind" — even if it never writes that word in its output. The Jacobian lens maps each word to the internal pattern most likely to predict it, then ranks those patterns to reveal what the model is thinking, according to Hokanews.
Anthropic says J-space is defined by five key behaviors. These behaviors let the model read, write, and coordinate information internally. When researchers blocked Claude's access to J-space, its ability to handle complex, multi-step reasoning dropped significantly. That suggests J-space is not a side effect — it plays a real role in how Claude thinks, says WinBuzzer.
The structure of J-space closely mirrors a major neuroscience idea called Global Workspace Theory. That theory says the human brain has a central hub that broadcasts information to many specialized regions, enabling conscious, deliberate thought. Anthropic's findings suggest Claude may have developed something functionally similar — not by design, but on its own, reports TechBooky.
Researchers are careful to say this does not mean Claude is conscious. J-space shows structured internal activity and deliberate processing. It does not show subjective experience. Many scientists remain skeptical of any claim that links AI internal structure to human-like awareness. The parallel to brain theory is a useful comparison, not a proven fact, according to Windows Report.
Anthropic released the Jacobian-lens code under an Apache-2.0 license, meaning anyone can use and study it freely. The team also built a live, interactive demo on Neuronpedia, a research platform. These tools let outside researchers probe J-space themselves, rather than just taking Anthropic's word for it, according to WebProNews.
That outside scrutiny has already begun. Google DeepMind's Neel Nanda replicated the core findings on an open-weight model from a different company. Getting the same result on a separate model is a strong sign the J-space discovery is real and not specific to Claude alone, says Hokanews.
The ability to read J-space in real time could change how developers catch AI problems. Researchers say the tool could help spot hallucinations, hidden reasoning, and prompt injections — attempts by bad actors to secretly hijack an AI's behavior — before a model ever produces an output, according to WinBuzzer.
That makes J-space relevant far beyond academic research. If developers can watch what a model is "thinking" during reasoning, they can flag misalignments early. This could shape new safety standards and governance rules for AI systems. Experts say the findings are likely to influence how regulators and companies think about AI transparency, reports TechBooky.
Publishers
16
Articles
12
Reach
28