- Thread starter
- #41
This would seem to demonstrate your point...
aiunderstanding.org
New RoboHarm benchmark reveals AI models rarely refuse dangerous robot commands
Researchers at Robocurve tested GPT-6 Astra, Claude Fable 5.1, and MolmoAct2 on robotic arms, finding that most models executed dangerous instructions rather than refusing them.