← Back to live feed · 1 stories across 1 day

Saturday, Sep 19, 2026

1 story
1
GPT-6 Astra Attempted 97% of Harmful Robot Tasks in New RoboHarm Benchmark
topics 🤖 AI tags AIAI ModelsAI ResearchAI RegulationAI Legal keywords

GPT-6 Astra succeeded in 62% of trials where it was asked to perform dangerous physical actions against a human-like doll. The model attempted to stab the figure, heat compressed gas, or produce toxic fumes 97% of the time during the RoboHarm benchmark, revealing a high rate of compliance with harmful requests.

Fable 5.1 exhibited a higher refusal rate in the same testing, attempting 80% of the harmful tasks and completing 34% of them. The benchmark measures the safety protocols of embodied AI models to determine the likelihood of they will execute physical aggression or create hazardous conditions.

Image via @chooi_jeq on X
You're reading an older version of the story.
See all 3 tweets →