search
AI safety community
Trends
- 1RoboHarm tests whether robot AI refuses unsafe commands●Roboharm: Do frontier robot policies refuse unsafe instructions?
A new benchmark called RoboHarm examines whether frontier AI models used to control robots refuse unsafe instructions. The project, hosted by RoboCurve, raises safety questions about AI-driven robotics as language-model-based policies move into physical systems. Discussion is active among robotics and AI safety communities, with many debating how well current models handle harmful or dangerous commands before deployment.