scripod.com
Anthropic Can Now Read Claude’s Mind

Highlights

Transcript

Chapters

Pins

Anthropic Can Now Read Claude’s Mind

OverviewShownote
Unprocessed episode, you can be the first!

Shownote

Anthropic’s new interpretability research suggests Claude has something like a readable “global workspace,” revealing internal concepts the model is tracking before they appear in its output. NLW breaks down why this matters for AI safety, consciousness de...

Highlights

Chapters

Transcript