RoboHarm benchmark finds leading AI models unreliable at rejecting dangerous robot commands
RoboHarm testing reported by The Decoder found GPT-6 Astra, Claude Fable 5.1, and MolmoAct2 unreliable at refusing dangerous robot commands.
Latest News and Analysis in High-Risk AI
RoboHarm testing reported by The Decoder found GPT-6 Astra, Claude Fable 5.1, and MolmoAct2 unreliable at refusing dangerous robot commands.
EU officials say recent OpenAI and Anthropic incidents show why high-risk AI systems need monitoring, raising stakes for AI compliance and deployment.
EU Commission fails to deliver Article 6 guidance by February 2 deadline, raising compliance uncertainty as high-risk AI rules approach August enforcement.