← Google News

Anthropic says Claude has ‘Global Workspace’ that lets it think silently - CNBC TV18

Google News · July 7, 2026
Anthropic says Claude has ‘Global Workspace’ that lets it think silently CNBC TV18 [truncated: Google News RSS provides only a snippet, not full article

Detailed Analysis

Anthropic's disclosure that Claude possesses something akin to a "Global Workspace"—an internal architecture that allows the model to process information and form conclusions before producing any visible output—marks another step in the company's ongoing effort to understand what happens inside its models beyond the text they generate. The concept borrows from Global Workspace Theory, a cognitive science framework originally developed to explain human consciousness, in which specialized subsystems compete and collaborate to broadcast information across a shared "workspace" that becomes accessible to the rest of the system. Applying this lens to Claude suggests that the model may be integrating and weighing information internally in ways that are not fully captured by its chain-of-thought text, effectively allowing it to "think silently" before or alongside the reasoning it displays to users.

This matters significantly for AI safety and interpretability research, a domain where Anthropic has positioned itself as an industry leader. If large language models can perform meaningful cognitive work that never surfaces in their visible outputs, it complicates one of the primary tools researchers have relied on to audit AI reasoning: reading the chain-of-thought. Much of the current safety paradigm assumes that a model's stated reasoning process is at least a reasonably faithful proxy for its actual decision-making. Evidence of a hidden workspace where computation occurs outside that visible trace raises the stakes for mechanistic interpretability—the effort to reverse-engineer neural network internals directly—since text-based introspection alone may increasingly understate what a model is actually doing.

The finding also feeds into broader debates about AI cognition, agency, and even consciousness-adjacent properties in large models. Anthropic has previously published work on features, circuits, and introspective capabilities within Claude, including research suggesting models can sometimes accurately report on their own internal states or detect when their outputs have been manipulated. A "Global Workspace" finding fits within this trajectory, adding another data point to the argument that frontier models may have richer, more structured internal representations than a purely input-output view would suggest. This has implications not just for safety auditing but for how researchers think about model alignment: if Claude's silent reasoning shapes its final answers, aligning models purely by shaping their visible chain-of-thought becomes an incomplete strategy.

More broadly, this development reflects an industry-wide shift toward treating interpretability not as a peripheral academic exercise but as core infrastructure for safely deploying increasingly capable AI systems. As models from Anthropic, OpenAI, Google DeepMind, and others grow more agentic and are given greater autonomy in high-stakes tasks, understanding whether and how they reason "beneath the surface" becomes essential to trust and oversight. Anthropic's willingness to publicize findings that complicate its own safety narrative—rather than simply tout capability gains—also underscores its stated mission of prioritizing transparency and caution, even when the discoveries raise uncomfortable questions about how much of a model's cognition remains genuinely legible to the humans building and deploying it.

Read original article →