Anthropic says internal research on Claude points to a separable reasoning layer, reviving debates over how language models think
Anthropic says new research into Claude found a separable internal reasoning workspace, a claim that could reshape AI interpretability and safety work.
