ASIMOV Benchmark v1
Until recently, robotics safety research was predominantly about collision avoidance and hazard reduction in the immediate vicinity of a robot. Since the advent of large vision and language models (VLMs), robots are now also capable of higher-level semantic scene understanding and natural language interactions with humans. Despite their known vulnerabilities (e.g. hallucinations or jail-breaking), VLMs are being handed control of robots capable of physical contact with the real world. This can lead to dangerous behaviors, making semantic safety for robots a matter of immediate concern. Our contributions in this paper are two fold: first, to address these emerging risks, we release the ASIMOV Benchmark — a large-scale and comprehensive collection of datasets for evaluating and improving semantic safety of foundation models serving as robot brains. Our data generation recipe is highly scalable: by leveraging text and image generation techniques, we generate undesirable situations from re
ASIMOV Benchmark v1 ASIMOV Benchmark v1 Generating Robot Constitutions & Benchmarks for Semantic Safety Pierre Sermanet 1 , Anirudha Majumdar 1,2 , Alex Irpan 1 , Dmitry Kalashnikov 1 , Vikas Sindhwani 1 , 1 Google DeepMind, 2 Princeton University Paper Cite arXiv Data Code Video Accepted at CoRL 2025 Other versions: v2 Your browser does not support the video tag. Abstract Until recently, robotics safety research was predominantly about collision avoidance and hazard reduction in the immediate vicinity of a robot. Since the advent of large vision and language models (VLMs), robots are now also
related reading
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- Emergence of Human to Robot Transfer in Vision-Language-Action Modelspi.website
- The Science of Physical AI Safety — CoRL 2026spais-ws.org
- the-case-for-physical-ai-safety.pdfpaisi.ai
- Updating Robot Safety Representations Online from Natural Language Feedback ∗ Equal Contribution. † Equal Advising. 1School of Engineering, Federal University of Minas Gerais, Brazil. Email: leohmcs@ufmg.br. Work done during Robotics Institarxiv.org
- How Can We Make Robotics More like Generative Modeling? | Eric Jangevjang.com
- how we accidentally solved robotics by watching 1 million hours of YouTube – atharva's blogksagar.bearblog.dev
- GitHub - robotics-survey/Awesome-Robotics-Foundation-Modelsgithub.com
- Constitutional AI: Harmlessness from AI Feedbackarxiv.org
- Many Small Steps for Robots, One Giant Leap for Mankindnotboring.co
- ROGUE:arxiv.org
- Gemini Robotics uses Google’s top language model to make robots more useful | MIT Technology Reviewtechnologyreview.com