arXiv cs.CV· Yunzhe Xu, Zhe Liu·· 4 小时前AI 评分24
多智能体视觉语言导航 MAVLN:首个系统形式化、基准与 TRISS 方法
Systematic Multi-Agent Vision-and-Language Navigation: Formulation, Benchmark, and Method
AI 导读
研究者提出多智能体视觉语言导航(VLN)的首个系统形式化,将其建模为带依赖与资源约束(presence locks、holding chains)的受限协调问题,并构建 MAVLN 基准,含 145 个场景、11,724 个 episode,最多支持 4 个智能体与三种指令模式。
来源:arXiv cs.CV · arxiv.org