MikeTrendsTrends right now

Yhn TechnologyRobotics first seen 14 h ago, last 7 min ago, peak #4

RoboHarm tests whether robots refuse unsafe instructions

Original: Roboharm: Do frontier robot policies refuse unsafe instructions?

A benchmark called RoboHarm is examining whether frontier AI models driving robots actually refuse unsafe or harmful instructions. The work asks how well safety training carries over from chatbots to physical systems, where a refusal failure could mean real-world damage or injury. It is drawing attention among robotics and AI safety researchers who argue embodied refusal is under-tested compared with text-based harms.

Why now: As AI models are deployed in physical robots, people are debating whether existing safety training is adequate for embodied risks.

RoboHarmfrontier AI modelsrobotics

Open on hn →

Rank over time, top of the chart is #1. 2 snapshots from 1 h ago to 7 min ago.

Evidence

API: https://socialmediatrends-api.osmike.com/v1/trends/3957