TLDRocket
Sign in

A global workspace in language models

Anthropic Covered by 4 sources

Anthropic researchers discovered that Claude has developed an internal neural workspace called the J-space, analogous to conscious thought in humans, where the model thinks about concepts without writing them down. The J-space contains approximately 60,000 patterns (one per word in Claude's vocabulary), is reportable to users when queried, can be deliberately controlled, and mediates higher-order reasoning tasks like multi-step math problems. This workspace enables researchers to observe Claude's silent reasoning, detect when it notices being tested or fabricates information, and provides a new tool for understanding and influencing language model decision-making.

Why it matters

Recent research by Anthropic has identified a novel set of internal neural patterns in LLMs, referred to as the J-space, which distinguishes conscious processing from unconscious activity.

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.