Interpreting Context-Aware Human Preferences for Multi-Objective Robot Navigation
解释情境感知的人类偏好以实现多目标机器人导航
Tharun Sethuraman, Subham Agrawal, Nils Dengler, Jorge de Heuvel, Teena Hassan, Maren Bennewitz
机构
*
Hochschule Bonn-Rhein-Sieg, Germany(波恩-莱茵-锡格应用科学大学,德国)
;
University of Bonn, Germany(波恩大学,德国)
;
Lamarr Institute for Machine Learning and Artificial Intelligence, Germany(拉马尔机器学习与人工智能研究所,德国)
机构
*
Research Intern, Department of Mechanical and Aerospace Engineering, George Washington University(乔治华盛顿大学机械与航空航天工程系研究实习生)
;
Ph.D. Student, Department of Mechanical and Aerospace Engineering, George Washington University(乔治华盛顿大学机械与航空航天工程系博士生)
;
Undergraduate Student, Aerospace Program, University of California, Berkeley(加州大学伯克利分校航空航天项目本科生)
;
Full Professor, Department of Electrical Engineering and Computer Science, University of California, Berkeley(加州大学伯克利分校电气工程与计算机科学系教授)
;
Ph.D. Student, Department of Computer Science, George Washington University(乔治华盛顿大学计算机科学系博士生)
;
Associate Professor, Department of Mechanical and Aerospace Engineering, George Washington University(乔治华盛顿大学机械与航空航天工程系副教授)
When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs
当观察不足时:视觉注意结构揭示大语言模型中的幻觉
Fanpu Cao, Xin Zou, Xuming Hu, Hui Xiong
机构
*
Thrust of Artificial Intelligence, HKUST (Guangzhou)(人工智能前沿 thrust,香港科技大学(广州))
;
Department of Computer Science and Engineering, HKUST(计算机科学与工程系,香港科技大学)
专题命中
幻觉与鲁棒性
:visual reasoning(abstract);multimodal large language model(abstract);分类 cs.CV、cs.AI
Probing Cross-modal Information Hubs in Audio-Visual LLMs
探测音频-视觉大语言模型中的跨模态信息枢纽
Jihoo Jung, Chaeyoung Jung, Ji-Hoon Kim, Joon Son Chung
机构
*
Department of Electrical Engineering, Korea Advanced Institute of Science
;
The Graduate School of Advanced Imaging Science, Multimedia \& Film, Chung-Ang University, Seoul, Republic of Korea
专题命中
幻觉与鲁棒性
:vision language model(abstract);分类 cs.AI
Safety-Oriented Evaluation of Language Understanding Systems for Air Traffic Control
面向安全的语言理解系统在空交通管制中的评估
Yujing Chang, Yash Guleria, Duc-Thinh Pham, Nhut-Huy Pham, Ningli Wang, Vu N. Duong, Sameer Alam
机构
*
ATMRI, Nanyang Technological University (NTU), Singapore(航空交通管理研究所,南洋理工大学(NTU),新加坡)
;
School of Management, Indian Institute of Technology Mandi, India(管理学院,印度理工学院曼迪分校,印度)
;
Centre of AI Research, VinUniversity, Vietnam(人工智能研究中心,文大学,越南)
Cluster-Aware Neural Collapse Prompt Tuning for Long-Tailed Generalization of Vision-Language Models
面向长尾泛化的眼视觉-语言模型集群感知神经崩溃提示微调
Boyang Guo, Liang Li, Lin Peng, Yuhan Gao, Xichun Sheng, Chenggang Yan
机构
*
Hangzhou Dianzi University(杭州电子科技大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Macao Polytechnic University(澳门理工学院)
;
Zhejiang Provincial Key Laboratory of Low Altitude Ubiquitous Networking Technology, HDU(浙江省低空 ubiquitous 网络技术重点实验室,HDU)
See the past: Time-Reversed Scene Reconstruction from Thermal Traces Using Visual Language Models
回望过去:利用视觉语言模型从热迹中进行时间反转场景重建
Kebin Contreras, Luis Toscano-Palomino, Mauro Dalla Mura, Jorge Bacca
机构
*
Physics School, Universidad Industrial de Santander, Colombia(圣安德烈大学物理系,哥伦比亚)
;
Department of Computer Science, Universidad Industrial de Santander, Colombia(圣安德烈大学计算机科学系,哥伦比亚)
;
GIPSA-Lab, Université Grenoble Alpes, CNRS, Grenoble INP, Grenoble, France(格拉斯实验室,格勒诺布尔阿尔卑斯大学,CNRS,格勒诺布尔INP,法国)
;
Institut Universitaire de France (IUF), France(法国国家科学院(IUF))
专题命中
VLM训练与架构
:visual language model(title);VLM(abstract,abstract_cn);分类 cs.CV、cs.AI
Agent-Based Post-Hoc Correction of Agricultural Yield Forecasts
基于代理的农业产量预测事后修正
Matthew Beddows, Aiden Durrant, Georgios Leontidis
机构
*
School of Natural and Computing Sciences(自然与计算科学学院)
;
University of Aberdeen(阿伯丁大学)
;
School of Computing Sciences(计算科学学院)
;
University of East Anglia(东安格利亚大学)
;
Department of Physics and Technology(物理与技术系)
;
UiT The Arctic University of Norway(北欧大学(UiT))