Skip to main content
Aggregate arXiv cs.AI 人工智能 17 Aug 2026 - 12:30

Active Perception for Embodied Disambiguation

RSS 官方收录 · 可信分层展示

关键摘要

arXiv:2608.…

  • 13605v1 Announce Type: new Abstract: Natural language provides robots …
  • Existing interactive disambiguation methods primarily obtain additiona…
  • We propose an active-perception framework for embodied target disambig…

摘要引擎:抽取

正文提要

arXiv:2608.13605v1 Announce Type: new Abstract: Natural language provides robots with a flexible task interface, but target ambiguity in embodied environments arises not only from user intent; it can also result from missing taskrelevant physical evidence in the current observation. Existing interactive disambiguation methods primarily obtain additional information by asking the user, whereas occlusion, restricted viewpoints, unreadable text, and unobserved targets require the robot to actively change its observation. We propose an active-perception framework for embodied target disambiguation that uses active observation as the backbone for information acquisition and uses a vision-language model to decide, on the basis of accumulated visual evidence and interaction information, whether to continue observing, request clarification, or complete target selection. Active observation can both directly recover missing discriminative evidence and reveal object names, labels, and semantic attributes, thereby improving user clarification when it remains necessary. Real-robot experiments show that the framework combines physical information acquisition and userintent clarification within a unified embodied disambiguation process.

来源:https://arxiv.org/abs/2608.13605

打开官方原文 站点原文页 可信分区 本信源更多 今日简报 分享图 RSS 稍后再看列表