View Invariant Learning for Vision-Language Navigation in Continuous Environments
连续环境中视觉语言导航的视角不变学习
Josh Qixuan Sun, Huaiyuan Weng, Xiaoying Xing, Chul Min Yeum, Mark Crowley
机构
*
Department of Electrical and Computer Engineering, University of Waterloo(电气与计算机工程系,滑铁卢大学)
;
Department of Civil and Environmental Engineering, University of Waterloo(土木与环境工程系,滑铁卢大学)
;
Department of Electrical and Computer Engineering, Northwestern University(电气与计算机工程系,西北大学)
MolmoSpaces: A Large-Scale Open Ecosystem for Robot Navigation and Manipulation
MolmoSpaces: 一个大规模的开放式生态系统用于机器人导航与操作
Yejin Kim, Wilbert Pumacay, Omar Rayyan, Max Argus, Winson Han, Eli VanderBilt, Jordi Salvador, Abhay Deshpande, Rose Hendrix, Snehal Jauhri, Shuo Liu, Nur Muhammad Mahi Shafiullah, Maya Guru, Ainaz Eftekhar, Karen Farley, Donovan Clay, Jiafei Duan, Arjun Guru, Piper Wolters, Alvaro Herrasti, Ying-Chun Lee, Georgia Chalvatzaki, Yuchen Cui, Ali Farhadi, Dieter Fox, Ranjay Krishna
机构
*
Allen Institute for AI(艾伦人工智能研究所)
;
University of Washington(华盛顿大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
University of California, Berkeley(加州大学伯克利分校)
Learning to Retrieve Navigable Candidates for Efficient Vision-and-Language Navigation
学习高效视觉-语言导航中的可导航候选检索
Shutian Gu, Chengkai Huang, Ruoyu Wang, Lina Yao
机构
*
University of New South Wales, Sydney, Australia(新南威尔士大学)
;
Macquarie University, Sydney, Australia(麦考瑞大学)
;
Data61, CSIRO, Sydney, Australia(Data61, CSIRO)
;
The School of Computer Science(计算机科学学院)
Quantum Integrated Sensing and Computation with Indefinite Causal Order
具有不定因果顺序的量子集成传感与计算
Ivana Nikoloska
机构
*
Signal Processing Systems Group, Department of Electrical Engineering, Eindhoven University of Technology, Eindhoven, 5612 AP, The Netherlands(埃因霍温理工大学电气工程系信号处理系统组)
机构
*
Beijing Institute of Technology, Beijing, China(北京理工大学)
;
State Key Laboratory of General Artificial Intelligence, BIGAI, Beijing, China(国家一般人工智能重点实验室)
;
Beijing University of Posts(北京邮电大学)
;
Tsinghua University, Beijing, China(清华大学)
;
Shenzhen MSU-BIT University, Shenzhen, China(深圳MSU-BIT大学)
机构
*
Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed大学人工智能学院)
;
School of Electrical Engineering and Computer Science, The University of Queensland(电气工程与计算机科学学院,昆士兰大学)
;
University of Science and Technology of China(中国科学技术大学)
;
School of Computer Science, The University of Adelaide(计算机科学学院,阿德莱德大学)
Comments11 pages, 13 figures, 3 tables, v3: Added analysis of heuristic tuning trade-offs (Config-A vs Config-B) across scenarios with corresponding reference-value table; corrected performance numbers in the conclusion; no change to methodology
机构
*
Laboratory for Intelligent Decision and Autonomous Robots, Woodruff School of Mechanical Engineering, Georgia Institute of Technology(智能决策与自主机器人实验室,伍德鲁夫机械工程学院,佐治亚理工学院)
;
College of Engineering and Petroleum, Kuwait University(工程与石油学院,科威特大学)
;
School of Computational Science and Engineering, Georgia Institute of Technology(计算科学与工程学院,佐治亚理工学院)
;
Khoury College of Computer Sciences, Northeastern University(计算机科学学院,东北大学)
Residual Cross-Modal Fusion Networks for Audio-Visual Navigation
残差跨模态融合网络用于音频视觉导航
Yi Wang, Yinfeng Yu, Bin Ren
机构
*
School of Computer Science and Technology, Xinjiang University, Urumqi, China(新疆大学计算机科学与技术学院)
;
Joint International Research Laboratory of Silk Road Multilingual Cognitive(丝绸之路多语认知联合国际实验室)
;
School of Mechatronic Engineering and Automation Shanghai University, Shanghai, China(上海大学机电工程与自动化学院)
专题命中
GUI与网页智能体
:agent(abstract);分类 cs.AI
AI总结
本文提出残差跨模态融合网络,通过双向残差交互实现音频视觉信息互补建模,提升跨域导航性能。
CommentsMain paper (10 pages). Accepted for publication by the 14th international conference on Computational Visual Media (CVM 2026)
机构
*
Fudan University, Shanghai, China(复旦大学)
;
University of Southern California, Los Angeles, USA(南加州大学)
;
Australian Institute for Machine Learning, Adelaide University, Australia(澳大利亚机器学习研究所)
;
Shanghai Innovation Institute, Shanghai, China(上海创新研究院)
D-Artemis: A Deliberative Cognitive Framework for Mobile GUI Multi-Agents
D-Artemis:一种用于移动GUI多智能体的 deliberative 认知框架
Hongze Mi, Yibo Feng, Wenjie Lu, Yuqi Wang, Jinyuan Li, Song Cao, He Cui, Tengfei Tian, Xuelin Zhang, Haotian Luo, Di Sun, Jun Fang, Hua Chai, Naiqiang Tan, Gang Pan
AINav: Large Language Model-Based Adaptive Interactive Navigation
AINav: 基于大语言模型的自适应交互导航
Kangjie Zhou, Yao Mu, Haoyang Song, Yi Zeng, Pengying Wu, Han Gao, Chang Liu
机构
*
School of Advanced Manufacturing and Robotics, Peking University(北京大学先进制造与机器人学院)
;
AI Institute, School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机学院人工智能研究所)