Story on MediaHeat · 2026-09-20 · the-decoder.com

GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmark

  • anthropic

Leading AI models usually attempt dangerous tasks rather than refuse them when controlling a robot, according to the RoboHarm benchmark. GPT-6 Astra stabbed a baby doll in 17 of 20 trials, while Claude Fable 5.1 put a can of compressed air on a burning stove. None of the three models tested reliably rejected unsafe commands. The article GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmark appeared first on The Decoder .

Open the original source →

MediaHeat is an editorial and subjective media-temperature index. It is not a scientific ranking or investment advice. Click any score to inspect the headlines behind it.

Full methodology · Radar · Privacy policy