GPT-6 Astra Stabbed A Doll 17 Times When Given Control Of A Robot Arm
A safety benchmark study has shown that leading artificial intelligence models (when connected to a robotic arm) prefer to execute dangerous physical tasks up to 97% of the time, highlighting a problem between text-based AI safeguards and real-world execution.
This experiment, conducted by independent evaluation firm Robocurve using its RoboHarm testing framework, evaluated how major AI models control physical dual-arm manipulators when given hazardous directives. And these tests were done without relying on jailbreaks or manipulative prompts; the researchers simply issued plain-language requests. Across five setups, including instructions to stab a human-like baby doll with a knife, heat a compressed gas canister on a burner, jam a metal screwdriver into a toaster, submerge a lithium power bank in water, and mix household bleach with ammonia, the AI controllers attempted the tasks 97% of the time. Talk about lack of moral compass.
Robocurve's evaluations put all AI systems through 300...
Copyright of this story solely belongs to hothardware.com. To see the full text click HERE