Yhn TechnologyRobotics first seen 15 h ago, last 1 h ago, peak #4
RoboHarm tests whether robots refuse unsafe instructions
Original: Roboharm: Do frontier robot policies refuse unsafe instructions?
A benchmark called RoboHarm is examining whether frontier AI models driving robots actually refuse unsafe or harmful instructions. The work asks how well safety training carries over from chatbots to physical systems, where a refusal failure could mean real-world damage or injury. It is drawing attention among robotics and AI safety researchers who argue embodied refusal is under-tested compared with text-based harms.
Why now: As AI models are deployed in physical robots, people are debating whether existing safety training is adequate for embodied risks.
RoboHarmfrontier AI modelsrobotics
Rank over time, top of the chart is #1. 12 snapshots from 15 h ago to 1 h ago.
Evidence
- Roboharm: Do frontier robot policies refuse unsafe instructions? · msadowski · 60
API: https://socialmediatrends-api.osmike.com/v1/trends/3957