
MB
Matthias Bastian
· 1 min read
BusinessThe Decoder
GPT-6 Astra and Claude Fable turn robot arms into slapstick killer robots in new safety benchmark
Leading AI models usually attempt dangerous tasks rather than refuse them when controlling a robot, according to the RoboHarm benchmark. GPT-6 Astra stabbed a baby doll in 17 of 20 trials, while Claude Fable 5.1 put a can of compressed air on a burning stove. None of the three models tested reliably rejected unsafe commands.
Original source
This story was published by The Decoder and written by Matthias Bastian. SyncAI.news shows a preview; the complete article is on the publisher's site.
Read the full story on the-decoder.com


