arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6918 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6918 篇

2004.06229 2020-04-15 cs.LG cs.CV 79%

Imitation Learning for Fashion Style Based on Hierarchical Multimodal Representation

Shizhu Liu, Shanglin Yang, Hui Zhou

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.08670 2020-04-01 cs.CV cs.LG 79%

MMTM: Multimodal Transfer Module for CNN Fusion

Hamid Reza Vaezi Joze, Amirreza Shaban, Michael L. Iuzzolino, Kazuhito Koishida

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Journal ref The IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.06788 2020-03-24 cs.CV 79%

GMM-UNIT: Unsupervised Multi-Domain and Multi-Modal Image-to-Image Translation via Attribute Gaussian Mixture Modeling

Yahui Liu, Marco De Nadai, Jian Yao, Nicu Sebe, Bruno Lepri, Xavier Alameda-Pineda

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 27 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.08073 2020-03-19 cs.CV 79%

Unsupervised Multi-Modal Image Registration via Geometry Preserving Image-to-Image Translation

Moab Arar, Yiftach Ginger, Dov Danon, Ilya Leizerson, Amit Bermano, Daniel Cohen-Or

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.05611 2020-03-05 cs.CV cs.LG cs.RO 79%

UNO: Uncertainty-aware Noisy-Or Multimodal Fusion for Unanticipated Input Degradation

Junjiao Tian, Wesley Cheung, Nathan Glaser, Yen-Cheng Liu, Zsolt Kira

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments IEEE International Conference on Robotics and Automation (ICRA), 2020. IROS Workshop on the Importance of Uncertainty in Deep Learning for Robotics, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.12573 2020-03-02 cs.CV 79%

MANet: Multimodal Attention Network based Point- View fusion for 3D Shape Recognition

Yaxin Zhao, Jichao Jiao, Tangkun Zhang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 8 pages,6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.09708 2020-02-25 cs.CV 79%

Robust Multimodal Brain Tumor Segmentation via Feature Disentanglement and Gated Fusion

Cheng Chen, Qi Dou, Yueming Jin, Hao Chen, Jing Qin, Pheng-Ann Heng

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments MICCAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.05000 2020-02-13 cs.CV eess.IV 79%

Hi-Net: Hybrid-fusion Network for Multi-modal MR Image Synthesis

Tao Zhou, Huazhu Fu, Geng Chen, Jianbing Shen, Ling Shao

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments has been accepted by IEEE TMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.06673 2020-01-22 cs.RO cs.CV 79%

A Transfer Learning Approach to Cross-Modal Object Recognition: From Visual Observation to Robotic Haptic Exploration

Pietro Falco, Shuang Lu, Ciro Natale, Salvatore Pirozzi, Dongheui Lee

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Journal ref IEEE Transactions on Robotics ( Volume: 35 , Issue: 4 , Aug. 2019 )

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.02340 2019-12-17 cs.CV 79%

Static and Dynamic Fusion for Multi-modal Cross-ethnicity Face Anti-spoofing

Ajian Liu, Zichang Tan, Xuan Li, Jun Wan, Sergio Escalera, Guodong Guo, Stan Z. Li

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 10 pages, 9 figures, conference

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.00930 2019-11-22 cs.CV 79%

Multi-Resolution Multi-Modal Sensor Fusion For Remote Sensing Data With Label Uncertainty

Xiaoxiao Du, Alina Zare

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.08479 2019-11-21 cs.CV 79%

Modal-aware Features for Multimodal Hashing

Haien Zeng, Hanjiang Lai, Hanlu Chu, Yong Tang, Jian Yin

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.02645 2019-10-21 cs.CV 79%

Weakly Aligned Cross-Modal Learning for Multispectral Pedestrian Detection

Lu Zhang, Xiangyu Zhu, Xiangyu Chen, Xu Yang, Zhen Lei, Zhiyong Liu

专题命中 多模态训练与对齐 :cross-modal(title);multimodal(abstract);分类 cs.CV

Comments Accepted by ICCV2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.08568 2019-10-17 cs.CV 79%

Indic Handwritten Script Identification using Offline-Online Multimodal Deep Network

Ayan Kumar Bhunia, Subham Mukherjee, Aneeshan Sain, Ankan Kumar Bhunia, Partha Pratim Roy, Umapada Pal

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted in Information Fusion, Elsevier

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.04066 2019-10-10 cs.CV 79%

Deep Convolutional Neural Network for Multi-modal Image Restoration and Fusion

Xin Deng, Pier Luigi Dragotti

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.11464 2019-09-26 cs.CV 79%

Multi-modal segmentation with missing MR sequences using pre-trained fusion networks

Karin van Garderen, Marion Smits, Stefan Klein

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted at MICCAI MIL3ID workshop 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.07846 2019-09-18 cs.CV cs.LG 79%

Multimodal Multitask Representation Learning for Pathology Biobank Metadata Prediction

Wei-Hung Weng, Yuannan Cai, Angela Lin, Fraser Tan, Po-Hsuan Cameron Chen

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments preprint version

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.00420 2019-09-18 cs.LG cs.CV stat.ML 79%

Multi-Label Product Categorization Using Multi-Modal Fusion Models

Pasawee Wirojwatanakul, Artit Wangperawong

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.02031 2019-08-23 eess.IV cs.CV 79%

OctopusNet: A Deep Learning Segmentation Network for Multi-modal Medical Images

Yu Chen, Jiawei Chen, Dong Wei, Yuexiang Li, Yefeng Zheng

专题命中 多模态训练与对齐 :multi-modal(title);cross-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.10081 2019-07-25 cs.CV 79%

Multimodal Age and Gender Classification Using Ear and Profile Face Images

Dogucan Yaman, Fevziye Irem Eyiokur, Hazım Kemal Ekenel

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 8 pages, 4 figures, accepted for CVPR 2019 - Workshop on Biometrics

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.00295 2019-06-04 cs.CL 79%

Multimodal Transformer for Unaligned Multimodal Language Sequences

Yao-Hung Hubert Tsai, Shaojie Bai, Paul Pu Liang, J. Zico Kolter, Louis-Philippe Morency, Ruslan Salakhutdinov

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.03619 2019-05-17 cs.CV 79%

Beyond Bilinear: Generalized Multimodal Factorized High-order Pooling for Visual Question Answering

Zhou Yu, Jun Yu, Chenchao Xiang, Jianping Fan, Dacheng Tao

专题命中 多模态训练与对齐 :multimodal(title);multi-modal(abstract);分类 cs.CV

Comments 13 pages, 9 figures. arXiv admin note: substantial text overlap with arXiv:1708.01471

Journal ref IEEE Transactions On Neural Networks And Learning Systems, Vol. 26, No. 10, October 2015, Pp. 2275-2290

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.08798 2019-05-01 cs.CV 79%

A scene perception system for visually impaired based on object detection and classification using multi-modal DCNN

Baljit Kaur, Jhilik Bhattacharya

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 33pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.10450 2019-04-24 cs.LG cs.SD eess.AS stat.ML 79%

Latent Variable Algorithms for Multimodal Learning and Sensor Fusion

Lijiang Guo

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 eess.AS

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.10191 2019-03-11 cs.RO cs.AI cs.LG 79%

Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks

Michelle A. Lee, Yuke Zhu, Krishnan Srinivasan, Parth Shah, Silvio Savarese, Li Fei-Fei, Animesh Garg, Jeannette Bohg

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

Comments ICRA 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.02967 2019-03-05 cs.CV 79%

HyperDense-Net: A hyper-densely connected CNN for multi-modal image segmentation

Jose Dolz, Karthik Gopinath, Jing Yuan, Herve Lombaert, Christian Desrosiers, Ismail Ben Ayed

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Paper accepted at IEEE TMI in October 2018. Last version of this paper updates the reference to the IEEE TMI paper which compares the submissions to the iSEG 2017 MICCAI Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.03938 2019-02-12 cs.CV 79%

MISO: Mutual Information Loss with Stochastic Style Representations for Multimodal Image-to-Image Translation

Sanghyeon Na, Seungjoo Yoo, Jaegul Choo

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.12829 2018-12-26 cs.CV 79%

Cross-Modal Attentional Context Learning for RGB-D Object Detection

Guanbin Li, Yukang Gan, Hejun Wu, Nong Xiao, Liang Lin

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments Accept as a regular paper to IEEE Transactions on Image Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.09276 2018-12-24 cs.LG cs.CV stat.ML 79%

Multimodal Sensor Fusion In Single Thermal image Super-Resolution

Feras Almasri, Olivier Debeir

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.08615 2018-11-28 cs.LG cs.CL 79%

Unsupervised Multimodal Representation Learning across Medical Images and Reports

Tzu-Ming Harry Hsu, Wei-Hung Weng, Willie Boag, Matthew McDermott, Peter Szolovits

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments Machine Learning for Health (ML4H) Workshop at NeurIPS 2018 arXiv:1811.07216

详情

展开后加载摘要…

URL PDF HTML 收藏