arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

多模态信息融合

面向图像、视频、多传感器和跨模态感知的信息融合,包括 Image Fusion、红外可见光、遥感、医学影像、LiDAR/雷达/相机和音视频融合。

2025-09-25 至 2025-09-25 共收录 8 信号源:cs.CV, eess.IV, eess.SP, cs.RO, cs.MM

1. 通用Image Fusion 2 篇

2411.10679 2025-09-25 cs.CV 88%

SMLNet: A SPD Manifold Learning Network for Infrared and Visible Image Fusion

Huan Kang, Hui Li, Tianyang Xu, Xiao-Jun Wu, Rui Wang, Chunyang Cheng, Josef Kittler

机构 * School of Artificial Intelligence and Computer Science(人工智能与计算机科学学院) Jiangnan University(江南大学) Centre for Vision, Speech and Signal Processing(视觉、语音与信号处理中心) University of Surrey(Surrey大学)

专题命中 通用Image Fusion :image fusion(title,abstract);infrared and visible(title);multi-modal image fusion(abstract);分类 cs.CV

Comments 23 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16549 2025-09-25 cs.CV 79%

Efficient Rectified Flow for Image Fusion

Zirui Wang, Jiayi Zhang, Tianwei Guan, Yuhan Zhou, Xingyuan Li, Minjing Dong, Jinyuan Liu

机构 * City University of Hong Kong(香港城市大学) Dalian University of Technology(大连理工大学) Chinese University of Hong Kong(香港中文大学) Zhejiang University(浙江大学)

专题命中 通用Image Fusion :image fusion(title,abstract);分类 cs.CV

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 多聚焦/多曝光融合 1 篇

2509.19779 2025-09-25 cs.CV 79%

EfficienT-HDR: An Efficient Transformer-Based Framework via Multi-Exposure Fusion for HDR Reconstruction

Yu-Shen Huang, Tzu-Han Chen, Cheng-Yen Hsiao, Shaou-Gang Miaou

专题命中 多聚焦/多曝光融合 :multi-exposure(title,abstract);分类 cs.CV

Comments 10 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 医学影像融合 1 篇

2509.20022 2025-09-25 cs.CV 57%

PS3: A Multimodal Transformer Integrating Pathology Reports with Histology Images and Biological Pathways for Cancer Survival Prediction

Manahil Raza, Ayesha Azam, Talha Qaiser, Nasir Rajpoot

机构 * University of Warwick, UK(沃里克大学)

专题命中 医学影像融合 :multimodal fusion(abstract);分类 cs.CV

Comments Accepted at ICCV 2025. Copyright 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 自动驾驶多传感器融合 1 篇

2501.16803 2025-09-25 cs.RO cs.CV cs.NI eess.IV 67%

RG-Attn: Radian Glue Attention for Multi-modality Multi-agent Cooperative Perception

Lantao Li, Kang Yang, Wenqi Zhang, Xiaoxue Wang, Chen Sun

机构 * Sony (China) Limited(索尼(中国)有限公司) Renmin University of China(中国人民大学)

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.CV、eess.IV、cs.RO

Comments Accepted by ICCV 2025 DriveX workshop (Final Version)

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 机器人多传感器融合 1 篇

2509.19521 2025-09-25 cs.RO 79%

A Bimanual Gesture Interface for ROS-Based Mobile Manipulators Using TinyML and Sensor Fusion

Najeeb Ahmed Bhuiyan, M. Nasimul Huq, Sakib H. Chowdhury, Rahul Mangharam

机构 * †‡Department of Mechatronics Engineering, Rajshahi University of Engineering \& Technology, Kazla, Rajshahi-6204, Bangladesh §Department of Electrical \& Systems Engineering, School of Engineering \& Applied Science, University of Pennsylvania, Philadelphia, PA 19104, United States Email: , †, ‡, §

专题命中 机器人多传感器融合 :sensor fusion(title,abstract);分类 cs.RO

Comments 12 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 音视频/视觉语言融合 1 篇

2509.19719 2025-09-25 cs.CV 74%

Frequency-domain Multi-modal Fusion for Language-guided Medical Image Segmentation

Bo Yu, Jianhua Yang, Zetao Du, Yan Huang, Chenglong Li, Liang Wang

机构 * School of Computer Science and Technology, Anhui University(安徽大学计算机科学与技术学院) NLPR, MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) School of Artificial Intelligence, Anhui University(安徽大学人工智能学院)

专题命中 音视频/视觉语言融合 :multi-modal fusion(title);分类 cs.CV

Comments Accepted by MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

7. 融合架构与评测 1 篇

2509.20240 2025-09-25 cs.LG cs.AI 50%

A HyperGraphMamba-Based Multichannel Adaptive Model for ncRNA Classification

Xin An, Ruijie Li, Qiao Ning, Hui Li, Qian Ma, Shikai Guo

专题命中 融合架构与评测 :multimodal fusion(abstract)

Comments 9 pages, 17 figures (including subfigures), 1 table. Xin An and Ruijie Li contributed equally to this work and should be considered co-first authors

详情

展开后加载摘要…

URL PDF HTML 收藏