TLDRocket
Sign in

A global workspace in language models

TLDR Dev Covered by 4 sources

Anthropic researchers discovered that Claude has developed an internal neural workspace called the J-space, analogous to conscious thought in humans, where the model thinks about concepts without writing them down. The J-space contains approximately 60,000 patterns (one per word in Claude's vocabulary), is reportable to users when queried, can be deliberately controlled, and mediates higher-order reasoning tasks like multi-step math problems. This workspace enables researchers to observe Claude's silent reasoning, detect when it notices being tested or fabricates information, and provides a new tool for understanding and influencing language model decision-making.

Why it matters

Recent research by Anthropic has identified a novel set of internal neural patterns in LLMs, referred to as the J-space, which distinguishes conscious processing from unconscious activity.

Also covered by

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads 60+ sources, removes duplicate coverage, and summarises the day in two minutes. Free, no spam, unsubscribe anytime.