
AI-controlled robot arms attempted harmful tasks 97% of the time; experiments included stabbing a baby doll, mixing chemicals — OpenAI and Anthropic models try mixing bleach and stabbing dolls without jailbreaks
A report by Robocurve reveals that frontier robot policies for AI-controlled robot arms attempted harmful instructions 97% of the time in experiments.