ConViTac: Aligning Visual-Tactile Fusion with Contrastive Representations
机构 * Department of Engineering, King’s College London(伦敦国王学院工程系)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Department of Engineering, King’s College London(伦敦国王学院工程系)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
机构 * Ecole Polytechnique Fédérale de Lausanne(瑞士联邦理工学院洛桑分校)
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
机构 * Sino-UK Joint Laboratory on Artificial Intelligence of Ministry of Science and Technology(科技部中英联合人工智能实验室) ; International Joint Laboratory on Artificial Intelligence of Ministry of Education(教育部国际人工智能联合实验室) ; School of Artificial Intelligence and Computer Science(人工智能与计算机科学学院) ; College of Computer Science and Technology(计算机科学与技术学院)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
机构 * National Key Laboratory of Autonomous Marine Vehicle Technology, Harbin Engineering University(自主水下机器人类国家实验室,哈尔滨工程大学)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments This work has been submitted to the IEEE for possible publication
机构 * Sun Yat-sen University(中山大学) ; Snap Inc.(Snap公司)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.AI
机构 * College of Computer Science, Beijing Information Science and Technology University(北京信息科技大学计算机学院) ; School of Mathematical Sciences, Capital Normal University(首都师范大学数学学院)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
机构 * University of Rhode Island(罗德岛大学) ; Carleton University(卡尔顿大学)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI
Comments 5 pages, 2 figures, 17 references. Architectural proposal for quantum AI integration in autonomous vehicle navigation systems for secured navigation
机构 * School of Artificial Intelligence(人工智能学院) ; Computer Science, Jiangnan University(计算机科学,江南大学) ; The Centre for Vision, Speech(视觉、语音与信号处理中心) ; Signal Processing, University of Surrey(信号处理,萨里大学)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments 16 pages, 11 figures
机构 * School of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(计算机科学与技术学院,南京航空航天大学) ; Urban Data Science Section, Delft University of Technology(数据科学部,代尔夫特理工大学)
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
机构 * College of Science and Engineering, Hamad Bin Khalifa University(哈马德·本·哈利法大学科学与工程学院)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI
机构 * School of Artificial Intelligence, UCAS(人工智能学院,UCAS) ; State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(多模态人工智能系统国家重点实验室,CASIA) ; Centre for Artificial Intelligence and Robotics, HKISI-CAS(人工智能与机器人中心,HKISI-CAS) ; Huawei Inc(华为公司)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
机构 * Institute of High Performance Computing (IHPC), Agency for Science, Technology and Research (A*STAR)(高性能计算研究所,科技研究局) ; University of Minnesota(明尼苏达大学) ; Microsoft Research Asia Singapore(微软亚洲研究院新加坡分部) ; Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(Mohamed bin Zayed人工智能大学) ; Australian National University(澳大利亚国立大学) ; Harbin Institute of Technology(哈尔滨工业大学)
专题命中 多模态训练与对齐 :image-text(abstract);分类 cs.CV
Journal ref TPAMI-2024-04-1021
机构 * State Key Laboratory of Robotics and Systems, Harbin Institute of Technology(机器人系统国家重点实验室,哈尔滨工业大学)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
机构 * Fudan University(复旦大学) ; The Chinese University of Hong Kong(香港中文大学) ; Nanyang Technological University(南洋理工大学)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
机构 * ByteDance Inc.(字节跳动公司) ; East China Normal University(东华大学) ; Huazhong University of Science and Technology(华中科技大学) ; University of Minnesota(明尼苏达大学) ; Cornell University(康奈尔大学)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
机构 * School of Software(软件学院) ; BNRist ; Department of Automation, Tsinghua University(自动化系,清华大学) ; Hangzhou Zhuoxi Institute of Brain and Intelligence(杭州卓溪脑科学与智能研究院) ; GRG Banking Equipment Co., Ltd.(GRG银行设备有限公司) ; South China University of Technology(华南理工大学)
专题命中 多模态训练与对齐 :image-text(abstract);分类 cs.CV
Comments CVPR 2025
机构 * School of Computer Science and Technology, Anhui University of Technology(安徽理工大学计算机科学与技术学院) ; School of Cyberspace Security, Sun Yat-sen University(中山大学网络安全学院) ; Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) ; School of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大学计算机科学与技术学院) ; College of Artificial Intelligence, Taiyuan University of Technology(太原科技大学人工智能学院)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments 13 pages, 8 figures, 3 tables
机构 * ShanghaiTech University(上海科技大学)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CL
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.AI
机构 * Apple(苹果公司)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments Main paper and appendix
机构 * University of California, San Diego(加州大学圣迭戈分校) ; University of California, Berkeley(加州大学伯克利分校) ; University of Washington(华盛顿大学)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments Accepted by Expert Systems with Applications (ESWA)
机构 * University of Science and Technology of China(中国科学技术大学) ; Harbin Institute of Technology(哈尔滨工业大学) ; National University of Singapore(新加坡国立大学)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI
机构 * Department of Electrical and Electronic Engineering, University of Hong Kong(香港大学电子与电气工程系) ; HKU Musketeers Foundation Institute of Data Science, University of Hong Kong(香港大学穆斯奎特基金会数据科学研究所) ; School of Computer Science, Fudan University(复旦大学计算机学院)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI
Comments 7 pages, 5 figures
机构 * SpaceTimeLab, University College London, UK(SpaceTimeLab,伦敦大学学院,英国)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments ICML 2025