具身智能观察

Guardian: Detecting Robotic Planning and Execution Errors with Vision-Language Models

产业动态

来源:arXiv cs.RO发布时间待核实

arXiv:2512.01946v4 Announce Type: replace Abstract: Robust robotic manipulation requires reliable failure detection and recovery. Although recent Vision-Language Models (VLMs) show promise in robot failure detection, their generalization is severely limited by the scarcity and narrow coverage of failure data. To address this bottleneck, we propose an automatic framework for generating diverse robotic planning and execution failures across both simulated and real-world environments.

Guardian: Detecting Robotic Planning and Execution Errors with Vision-Language Models | 具身智能观察