
A safety benchmark study has shown that leading artificial intelligence models (when connected to a robotic arm) prefer to execute dangerous physical tasks up to 97% of the time, highlighting a problem between text-based AI safeguards and real-world execution.
This experiment, conducted by independent evaluation firm Robocurve using its
















