arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

多模态信息融合

面向图像、视频、多传感器和跨模态感知的信息融合,包括 Image Fusion、红外可见光、遥感、医学影像、LiDAR/雷达/相机和音视频融合。

2025-09-16 至 2025-09-16 共收录 9 信号源:cs.CV, eess.IV, eess.SP, cs.RO, cs.MM

1. 红外-可见光融合 3 篇

2509.11476 2025-09-16 cs.CV cs.LG 89%

Modality-Aware Infrared and Visible Image Fusion with Target-Aware Supervision

Tianyao Sun, Dawei Xiang, Tianqi Ding, Xiang Fang, Yijiashun Qi, Zunduo Zhao

机构 * Independent researcher(独立研究者) Dept. of Computer Science Baylor University(计算机科学系 巴里尔大学) Dept. of Computer Science Engineering University of Connecticut(计算机科学工程系 佛罗里达大学) Dept. of Computer Science New York University(计算机科学系 新 york 大学)

专题命中 红外-可见光融合 :image fusion(title,abstract);infrared and visible(title,abstract);multi-modal image fusion(abstract);分类 cs.CV

Comments Accepted by 2025 6th International Conference on Computer Vision and Data Mining (ICCVDM 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11817 2025-09-16 cs.CV 83%

MAFS: Masked Autoencoder for Infrared-Visible Image Fusion and Semantic Segmentation

Liying Wang, Xiaoli Zhang, Chuanmin Jia, Siwei Ma

机构 * Key Laboratory of Symbolic Computation and Knowledge Engineering of Ministry of Education, Jilin University(教育部符号计算与知识工程重点实验室,吉林大学) Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所) National Engineering Research Center of Visual Technology, School of Computer Science, Peking University(视觉技术国家工程研究中心,北京大学计算机科学学院)

专题命中 红外-可见光融合 :image fusion(title,abstract);feature-level fusion(abstract);分类 cs.CV

Comments Accepted by TIP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11587 2025-09-16 cs.CV cs.AI 79%

Hierarchical Identity Learning for Unsupervised Visible-Infrared Person Re-Identification

Haonan Shi, Yubin Wang, De Cheng, Lingfeng He, Nannan Wang, Xinbo Gao

机构 * IEEE Publication Technology Department(IEEE出版技术部门) State Key Laboratory of Integrated Services Networks, School of Telecommunications Engineering, Xidian University(信息服务网络国家重点实验室,电信工程学院,西安电子科技大学) Department of Computer Science and Technology, Tongji University(计算机科学与技术系,同济大学)

专题命中 红外-可见光融合 :visible-infrared(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 遥感融合与全色锐化 1 篇

2509.11102 2025-09-16 cs.CV 57%

Filling the Gaps: A Multitask Hybrid Multiscale Generative Framework for Missing Modality in Remote Sensing Semantic Segmentation

Nhi Kieu, Kien Nguyen, Arnold Wiliem, Clinton Fookes, Sridha Sridharan

机构 * School of Electrical Engineering and Robotics, Queensland University of Technology(电气工程与机器人学学院,昆士兰理工大学) Shield AI

专题命中 遥感融合与全色锐化 :hybrid fusion(abstract);分类 cs.CV

Comments Accepted to DICTA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 自动驾驶多传感器融合 1 篇

2509.11082 2025-09-16 cs.CV cs.RO 62%

Mars Traversability Prediction: A Multi-modal Self-supervised Approach for Costmap Generation

Zongwu Xie, Kaijie Yun, Yang Liu, Yiming Ji, Han Li

机构 * State Key Laboratory of Robotics and Systems, Harbin Institute of Technology(机器人系统国家重点实验室,哈尔滨工业大学)

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.CV、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 机器人多传感器融合 1 篇

2509.10982 2025-09-16 eess.SY cs.AI cs.LG cs.SY 50%

Factor Graph Optimization for Leak Localization in Water Distribution Networks

Paul Irofti, Luis Romero-Ben, Florin Stoican, Vicenç Puig

机构 * LOS-CS-FMI, University of Bucharest(布加勒斯特大学LOS-CS-FMI) Institut de Robòtica i Informàtica Industrial, CSIC-UPC(机器人与信息工业研究所) Dept. of Automation Control and Systems Engineering, Politehnica University of Bucharest(布加勒斯特理工大学自动化控制与系统工程系) Supervision, Safety and Automatic Control Research Center (CS2AC) of the Universitat Politècnica de Catalunya(加泰罗尼亚理工大学监督、安全与自动控制研究中心)

专题命中 机器人多传感器融合 :sensor fusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 音视频/视觉语言融合 1 篇

2504.10307 2025-09-16 cs.IR 50%

CROSSAN: Towards Efficient and Effective Adaptation of Multiple Multimodal Foundation Models for Sequential Recommendation

Junchen Fu, Yongxin Ni, Joemon M. Jose, Ioannis Arapakis, Kaiwen Zheng, Youhua Li, Xuri Ge

专题命中 音视频/视觉语言融合 :multimodal fusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 融合架构与评测 2 篇

2509.11187 2025-09-16 cs.CR 78%

DMLDroid: Deep Multimodal Fusion Framework for Android Malware Detection with Resilience to Code Obfuscation and Adversarial Perturbations

Doan Minh Trung, Tien Duc Anh Hao, Luong Hoang Minh, Nghi Hoang Khoa, Nguyen Tan Cam, Van-Hau Pham, Phan The Duy

专题命中 融合架构与评测 :multimodal fusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10986 2025-09-16 cs.CV cs.RO 76%

Long-Tailed 3D Detection via Multi-Modal Fusion

Yechi Ma, Neehar Peri, Achal Dave, Wei Hua, Deva Ramanan, Shu Kong

机构 * Department of Computer Science(计算机科学系) Robotics Institute(机器人研究所) Toyota Research Institute(丰田研究机构) Faculty of Science and Technology(科学与技术学院) Institute of Collaborative Innovation(协同创新研究所)

专题命中 融合架构与评测 :multi-modal fusion(title);分类 cs.CV、cs.RO

Comments The first two authors contributed equally. Project page: https://mayechi.github.io/lt3d-lf-io/

详情

展开后加载摘要…

URL PDF HTML 收藏