How Emerging Multiagent AI Systems Are Shaping The Future Of Artificial Intelligence
Anthropic has published a report examining patterns and problems in multiagent AI systems, highlighting their behavior, coordination challenges, and potential risks.
Uncovering AI’s Ability To Recognize Hidden Words Like ‘Bread’ In Neural Data
Anthropic researchers report that Claude Opus can sometimes recognize externally inserted concepts like ‘bread’ in its neural activations, with limited detection accuracy.