OVGrasp: Open-Vocabulary Grasping Assistance via Multimodal Intent Detection
Chen Hu, Shan Luo, Letizia Gionfrida
机构
*
Department of Informatics, King's College London(伦敦国王学院信息学院)
;
Department of Engineering, King's College London(伦敦国王学院工程学院)
;
John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院)
机构
*
Department of Biomedical Engineering(生物医学工程系)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Department of Machine Learning(机器学习系)
;
Department of Radiation Oncology(放射肿瘤科)
;
Emory University School of Medicine(埃默里大学医学院)
;
University of Southern California(南加州大学)
Denoising GER: A Noise-Robust Generative Error Correction with LLM for Speech Recognition
Yanyan Liu, Minqiang Xu, Yihao Chen, Liang He, Lei Fang, Sian Fang, Lin Liu
机构
*
School of Computer Science and Technology(计算机科学与技术学院)
;
Xinjiang University(新疆大学)
;
Hefei iFly Digital Technology Co. Ltd.(合肥iFly数字技术有限公司)
;
University of Science and Technology of China(中国科学技术大学)
;
Tsinghua University(清华大学)
COBRA: Multimodal Sensing Deep Learning Framework for Remote Chronic Obesity Management via Wrist-Worn Activity Monitoring
Zhengyang Shen, Bo Gao, Mayue Shi
机构
*
Department of Electrical and Electronic Engineering, Imperial College London, London SW7 2AZ, UK(帝国理工学院电子与电气工程系)
;
Institute of Biomedical Engineering, Department of Engineering Science, University of Oxford, Oxford OX3 7DQ, UK(牛津大学生物医学工程研究所)
专题命中
视频多模态
:multimodal(title,abstract)
Comments19 pages, 4 figures. *Correspondence: m.shi16@imperial.ac.uk. Accepted by the IUPESM World Congress on Medical Physics and Biomedical Engineering 2025
TauGenNet: Plasma-Driven Tau PET Image Synthesis via Text-Guided 3D Diffusion Models
Yuxin Gong, Se-in Jang, Wei Shao, Yi Su, Kuang Gong
机构
*
J. Crayton Pruitt Family Department of Biomedical Engineering, University of Florida(J. Crayton Pruitt家族生物医学工程系,佛罗里达大学)
;
Department of Radiology & Biomedical Imaging, Yale University(放射学与生物医学成像系,耶鲁大学)
;
Department of Medicine, University of Florida(医学系,佛罗里达大学)
;
Banner Alzheimer’s Institute(Banner阿尔茨海默症研究所)
专题命中
多模态生成
:multimodal(abstract);分类 cs.CV
Comments9 pages, 4 figures, submitted to IEEE Transactions on Radiation and Plasma Medical Sciences
Straighter Flow Matching via a Diffusion-Based Coupling Prior
Siyu Xing, Jie Cao, Huaibo Huang, Haichao Shi, Xiao-Yu Zhang
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
Promptception: How Sensitive Are Large Multimodal Models to Prompts?
Mohamed Insaf Ismithdeen, Muhammad Uzair Khattak, Salman Khan
机构
*
Mohamed Bin Zayed University of Artificial Intelligence(莫罕默德·本·扎耶德人工智能大学)
;
Swiss Federal Institute of Technology Lausanne (EPFL)(洛桑联邦理工学院)
;
Australian National University(澳大利亚国立大学)
Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks
Nils Hoehing, Mayug Maniparambil, Ellen Rushe, Noel E. O'Connor, Anthony Ventresque
机构
*
School of Computer Science(计算机科学学院)
;
University College Dublin(都柏林大学)
;
School of Computing(计算机科学学院)
;
Dublin City University(都柏林城市大学)
;
School of Electronic Engineering(电子工程学院)
;
Trinity College Dublin(都柏林三一学院)
Multimodal Feature Fusion Network with Text Difference Enhancement for Remote Sensing Change Detection
Yijun Zhou, Yikui Zhai, Zilu Ying, Tingfeng Xian, Wenlve Zhou, Zhiheng Zhou, Xiaolin Tian, Xudong Jia, Hongsheng Zhang, C. L. Philip Chen
机构
*
College of Electronics and Information Engineering, Wuyi University(威怡大学电子与信息工程学院)
;
School of Electronic and Information Engineering and the Key Laboratory of Big Data and Intelligent Robot, Ministry of Education, South China University of Technology(电子与信息工程学院和大数据与智能机器人重点实验室,华南理工大学)
;
State Key Laboratory of Lunar and Planetary Sciences, Macau University of Science and Technology(澳门大学地球和行星科学国家重点实验室)
;
College of Engineering and Computer Science, California State University, Northridge(工程与计算机科学学院,加州大学北岭分校)
;
Department of Geography, The University of Hong Kong(地理系,香港大学)
;
Faculty of Computer Science and Engineering, S(计算机科学与工程学院,S)
Focus Through Motion: RGB-Event Collaborative Token Sparsification for Efficient Object Detection
Nan Yang, Yang Wang, Zhanwen Liu, Yuchao Dai, Yang Liu, Xiangmo Zhao
机构
*
School of Information Engineering, Chang’an University(信息工程学院,长安大学)
;
School of Electronics and Information, Northwestern Polytechnical University(电子信息学院,西北工业大学)
;
School of Vehicle and Mobility, Tsinghua University(车辆与移动学院,清华大学)