STEP:多模态 LLM 的状态感知人机协作任务估计与规划
Original title: STEP: State-Aware Task Estimation and Planning with Multi-Modal LLMs for Human-Robot Collaboration
IndustryAI 72
Source: arXiv cs.ROPublish time unverified
arXiv:2608.27225v1 Announce Type: new Abstract: Effective human-robot collaboration in industrial settings requires robots to understand human intentions and assist with task planning, reducing workload. Recent works have explored the use of Multi-modal Large Language Models (MM-LLMs) for task planning in such data-scarce scenarios, leveraging in-context learning to interpret user actions and generate long-horizon action plans in natural language.