MK-Pose: Category-Level Object Pose Estimation via Multimodal-Based Keypoint Learning
Yifan Yang, Peili Song, Enfan Lan, Dong Liu, Jingtai Liu
机构
*
Institute of Robotics and Automatic Information System(机器人与自动信息系统研究所)
;
Tianjin Key Laboratory of Intelligent Robotics(天津智能机器人重点实验室)
;
TBI center(TBI中心)
;
Nankai University(南开大学)
TalkFashion: Intelligent Virtual Try-On Assistant Based on Multimodal Large Language Model
Yujie Hu, Xuanyu Zhang, Weiqi Li, Jian Zhang
机构
*
School of Electronic and Computer Engineering, Peking University, China(北京大学电子与计算机工程学院)
;
Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology, Shenzhen Graduate School, Peking University(北京大学深圳研究生院)
Leveraging Multimodal Data and Side Users for Diffusion Cross-Domain Recommendation
Fan Zhang, Jinpeng Chen, Huan Li, Senzhang Wang, Yuan Cao, Kaimin Wei, JianXiang He, Feifei Kou, Jinqing Wang
机构
*
School of Computer Science (National Pilot Software Engineering School), Beijing University of Posts and Telecommunications(北京邮电大学计算机学院(国家级试点软件工程学院))
;
The State Key Laboratory of Blockchain and Data Security, Zhejiang University(浙江大学区块链与数据安全国家重点实验室)
;
Central South University(中南大学)
;
National University of Singapore(新加坡国立大学)
;
Jinan University(暨南大学)
TaleForge: Interactive Multimodal System for Personalized Story Creation
Minh-Loi Nguyen, Quang-Khai Le, Tam V. Nguyen, Minh-Triet Tran, Trung-Nghia Le
机构
*
1 University of Science, VNU-HCM, Ho Chi Minh City, Vietnam
;
2 Vietnam National University, Ho Chi Minh City, Vietnam
;
3 Department of Computer Science \ of Dayton Ohio, United States
机构
*
Indian Institute of Technology, Bombay(印度理工学院班加罗尔分校)
;
Indian Institute of Technology, Roorkee(印度理工学院罗尔基分校)
;
Microsoft Research(微软研究院)
;
Stanford University(斯坦福大学)
;
Adobe MDSR(Adobe MDSR实验室)
;
Carnegie Mellon University(卡内基梅隆大学)
Multimodal Coherent Explanation Generation of Robot Failures
Pradip Pramanick, Silvia Rossi
机构
*
Interdepartmental Center for Advances in Robotic Surgery - ICAROS, University of Naples Federico II(跨部门先进机器人手术中心 - ICAROS,那不勒斯费德里科二世大学)
;
Department of Electrical Engineering and Information Technologies - DIETI, University of Naples Federico II(电气工程与信息科技系 - DIETI,那不勒斯费德里科二世大学)
专题命中
多模态生成
:multimodal(title,abstract);分类 cs.AI
Journal ref2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching
Yepeng Liu, Zhichao Sun, Baosheng Yu, Yitian Zhao, Bo Du, Yongchao Xu, Jun Cheng
机构
*
National Engineering Research Center for Multimedia Software(多媒体软件国家工程研究中心)
;
Institute of Artificial Intelligence(人工智能研究院)
;
School of Computer Science(计算机科学学院)
;
Hubei Key Laboratory of Multimedia and Network Communication Engineering(湖北省多媒体与网络通信工程重点实验室)
;
Lee Kong Chian School of Medicine(李光耀医学院)
;
Nanyang Technological University(南洋理工大学)
;
Ningbo Institute of Materials Technology and Engineering(宁波材料技术与工程研究所)
;
Chinese Academy of Sciences(中国科学院)
;
Institute for Infocomm Research (I 2 R)(信息与通信研究所以(I 2 R))
;
Agency for Science, Technology and Research (A*STAR)(科技研究局(A*STAR))
COSMMIC: Comment-Sensitive Multimodal Multilingual Indian Corpus for Summarization and Headline Generation
Raghvendra Kumar, S. A. Mohammed Salman, Aryan Sahu, Tridib Nandi, Pragathi Y. P., Sriparna Saha, Jose G. Moreno
机构
*
Department of Computer Science and Engineering, Indian Institute of Technology Patna, India(印度理工学院帕纳分校计算机科学与工程系)
;
Department of Metallurgical and Materials Engineering, National Institute of Technology Tiruchirappalli, India(印度理工学院 Tiruchirappalli 金属与材料工程系)
;
Department of Computer Science and Information Systems, BITS Pilani – Goa Campus, India(比斯·帕尼学院 Goa 分校计算机科学与信息系统系)
;
Department of Computer Science and Engineering, Indian Institute of Information Technology Vadodara, India(印度信息科技学院瓦达拉分校计算机科学与工程系)
;
Department of Computer Science and Engineering, B.M.S. College of Engineering, Bangalore, India(班加罗尔 B.M.S. 工程学院计算机科学与工程系)
;
Université de Toulouse, IRIT UMR 5505 CNRS, France(图卢兹大学 IRIT UMR 5505 CNRS 实验室)
BrainMAP: Multimodal Graph Learning For Efficient Brain Disease Localization
Nguyen Linh Dan Le, Jing Ren, Ciyuan Peng, Chengyao Xie, Bowen Li, Feng Xia
机构
*
School of Computing Technologies, RMIT University(计算技术学院,拉筹纳斯大学)
;
Institute of Innovation, Science and Sustainability, Federation University Australia(创新、科学与可持续性研究所,联邦大学澳大利亚)
seg2med: a bridge from artificial anatomy to multimodal medical images
Zeyu Yang, Zhilin Chen, Yipeng Sun, Anika Strittmatter, Anish Raj, Ahmad Allababidi, Johann S. Rink, Frank G. Zöllner
机构
*
Computer Assisted Clinical Medicine, Medical Faculty Mannheim, Heidelberg University(计算机辅助临床医学,曼海姆医学院,海德堡大学)
;
Pattern Recognition Lab, Friedrich-Alexander-University Erlangen-Nuremberg(模式识别实验室,埃尔兰根-纽伦堡弗里德里希-亚历山大大学)
;
Department of Radiology and Nuclear Medicine, University Medical Center Mannheim(放射学与核医学系,曼海姆大学医学中心)
;
Mannheim Institute for Intelligent Systems in Medicine, Medical Faculty Mannheim, Heidelberg University(曼海姆智能医学研究所,曼海姆医学院,海德堡大学)
;
Optical Bioimaging Laboratory, Department of Biomedical Engineering, College of Design and Engineering, National University of Singapore(光学生物成像实验室,生物医学工程系,设计与工程学院,新加坡国立大学)