Safety Evaluation Enters the Foundation-Model Stack
Google paired its robotics release with layered safeguards and an agentic-robot safety benchmark.
Google paired its robotics release with layered safeguards and an agentic-robot safety benchmark.
Researchers propose more granular evaluation of spatial reasoning needed for physical manipulation and navigation.
A benchmark evaluates decision-level safety and reliability across multi-robot and aerial scenarios.