Claude's Secret Workshop: Why Anthropic's J-Space Discovery Changes AI Safety Forever
Anthropic found a hidden internal workspace inside Claude where it silently plans, reasons, and catches bugs before responding. The consciousness headlines are flashy. What this actually means for AI safety is far more important.

The Signal
On Sunday, a 16-author team at Anthropic published a paper that quietly upended how we understand what happens inside a large language model when it's not talking to us. Titled "Verbalizable Representations Form a Global Workspace in Language Models," the study reveals that Claude has spontaneously developed an internal structure the researchers call "J-space" — a tiny, privileged zone where the model holds concepts it can actively reason with and report on, surrounded by a vast...