← Back to brief
ResearchOfficialPreprintarXiv Robotics

SafeRelBench: A Spatial-Relation-Aware Benchmark for Process-Level Safety in VLM-Driven Embodied Agents

Researchers have introduced SafeRelBench, a benchmark comprising 507 samples designed to evaluate process-level safety in vision-language-model-driven embodied agents. The benchmark focuses on spatial relations such as support and containment, which are critical for safe interaction in household environments. Testing seven different models revealed that agents frequently achieve task completion while violating safety constraints during multi-step actions, highlighting a significant gap between task success and safety compliance.

Why it matters: SafeRelBench exposes the need for embodied agents to reason more effectively about spatial relations to ensure safety, rather than prioritizing task completion alone.

Full story at: arXiv Robotics