What Anthropic’s latest AI discovery does—and doesn’t—show
By James O'Donnell
An MIT Technology Review analysis examines Anthropic's ongoing focus on mechanistic interpretability, exploring how the company investigates the internal mathematical structures of its models to understand their outputs. The piece highlights Anthropic's distinct research culture relative to other top-tier AI labs.