arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6918 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6918 篇

1811.08305 2018-11-21 cs.CV 79%

IVD-Net: Intervertebral disc localization and segmentation in MRI with a multi-modal UNet

Jose Dolz, Christian Desrosiers, Ismail Ben Ayed

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Manuscript submitted to the Proceedings of the MICCAI 2018 IVD Challenge. arXiv admin note: text overlap with arXiv:1810.07003

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.04697 2018-11-13 cs.CL 79%

CUNI System for the WMT18 Multimodal Translation Task

Jindřich Helcl, Jindřich Libovický, Dušan Variš

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments Published at WMT18

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.02745 2018-11-08 cs.CV 79%

Y^2Seq2Seq: Cross-Modal Representation Learning for 3D Shape and Text by Joint Reconstruction and Prediction of View and Word Sequences

Zhizhong Han, Mingyang Shang, Xiyang Wang, Yu-Shen Liu, Matthias Zwicker

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments To be pubilished at AAAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.06233 2018-11-05 cs.CV 79%

Robust Deep Multi-modal Learning Based on Gated Information Fusion Network

Jaekyum Kim, Junho Koh, Yecheol Kim, Jaehyung Choi, Youngbae Hwang, Jun Won Choi

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments 2018 Asian Conference on Computer Vision (ACCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.01220 2018-10-22 cs.CV 79%

Multi-Modal Multi-Scale Deep Learning for Large-Scale Image Annotation

Yulei Niu, Zhiwu Lu, Ji-Rong Wen, Tao Xiang, Shih-Fu Chang

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Submited to IEEE TIP

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.02001 2018-10-05 cs.CV 79%

Image and Encoded Text Fusion for Multi-Modal Classification

Ignazio Gallo, Alessandro Calefati, Shah Nawaz, Muhammad Kamran Janjua

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted to DICTA 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.08993 2018-09-28 cs.CV 79%

Improved Semantic Stixels via Multimodal Sensor Fusion

Florian Piewak, Peter Pinggera, Markus Enzweiler, David Pfeiffer, Marius Zöllner

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.04732 2018-08-16 cs.CV cs.LG stat.ML 79%

Multimodal Unsupervised Image-to-Image Translation

Xun Huang, Ming-Yu Liu, Serge Belongie, Jan Kautz

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments Accepted by ECCV 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.04836 2018-07-17 cs.CV 79%

Disjoint Mapping Network for Cross-modal Matching of Voices and Faces

Yandong Wen, Mahmoud Al Ismail, Weiyang Liu, Bhiksha Raj, Rita Singh

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV

Comments Tech report

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.03232 2018-07-10 eess.SP cs.CV physics.med-ph 79%

Robust Heartbeat Detection from Multimodal Data via CNN-based Generalizable Information Fusion

B S Chandra, C S Sastry, S Jana

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.00864 2018-07-04 cs.CV 79%

Semi-supervised Learning: Fusion of Self-supervised, Supervised Learning, and Multimodal Cues for Tactical Driver Behavior Detection

Athma Narayanan, Yi-Ting Chen, Srikanth Malla

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.00064 2018-06-04 cs.AI cs.LG stat.ML 79%

Efficient Low-rank Multimodal Fusion with Modality-Specific Factors

Zhun Liu, Ying Shen, Varun Bharadhwaj Lakshminarasimhan, Paul Pu Liang, Amir Zadeh, Louis-Philippe Morency

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.AI

Comments * Equal contribution. 10 pages. Accepted by ACL 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.08660 2018-05-23 cs.CL 79%

Multimodal Affective Analysis Using Hierarchical Attention Strategy with Word-Level Alignment

Yue Gu, Kangning Yang, Shiyu Fu, Shuhong Chen, Xinyu Li, Ivan Marsic

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments Accepted by ACL 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.00532 2018-01-03 cs.CL 79%

Learning Multimodal Word Representation via Dynamic Fusion Methods

Shaonan Wang, Jiajun Zhang, Chengqing Zong

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments To be appear in AAAI-18

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.08681 2017-11-27 cs.NE cs.CV 79%

Beyond RGB: Very High Resolution Urban Remote Sensing With Multimodal Deep Networks

Nicolas Audebert, Bertrand Le Saux, Sébastien Lefèvre

专题命中 多模态训练与对齐 :multimodal(title);multi-modal(abstract);分类 cs.CV

Comments ISPRS Journal of Photogrammetry and Remote Sensing, Elsevier, A Para{î}tre

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.05516 2017-11-23 cs.CL 79%

Investigating Inner Properties of Multimodal Representation and Semantic Compositionality with Brain-based Componential Semantics

Shaonan Wang, Jiajun Zhang, Nan Lin, Chengqing Zong

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments To appear in AAAI-18

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.00003 2017-11-02 cs.CV 79%

Common Representation Learning Using Step-based Correlation Multi-Modal CNN

Gaurav Bhatt, Piyush Jha, Balasubramanian Raman

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments Accepted in Asian Conference of Pattern Recognition (ACPR-2017)

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.02116 2017-08-09 cs.MM 79%

CCL: Cross-modal Correlation Learning with Multi-grained Fusion by Hierarchical Network

Yuxin Peng, Jinwei Qi, Xin Huang, Yuxin Yuan

专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.MM

Comments 16 pages, accepted by IEEE Transactions on Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.01471 2017-08-07 cs.CV 79%

Multi-modal Factorized Bilinear Pooling with Co-Attention Learning for Visual Question Answering

Zhou Yu, Jun Yu, Jianping Fan, Dacheng Tao

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

Comments ICCV 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.07250 2017-07-25 cs.CL 79%

Tensor Fusion Network for Multimodal Sentiment Analysis

Amir Zadeh, Minghai Chen, Soujanya Poria, Erik Cambria, Louis-Philippe Morency

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL

Comments Accepted as full paper in EMNLP 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.04256 2017-06-15 cs.CV 79%

Online Convolutional Dictionary Learning for Multimodal Imaging

Kevin Degraux, Ulugbek S. Kamilov, Petros T. Boufounos, Dehong Liu

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.08624 2017-05-25 cs.CV 79%

VANETs Meet Autonomous Vehicles: A Multimodal 3D Environment Learning Approach

Yassine Maalej, Sameh Sorour, Ahmed Abdel-Rahim, Mohsen Guizani

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 7 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.06676 2017-05-23 cs.CV 79%

MUTAN: Multimodal Tucker Fusion for Visual Question Answering

Hedi Ben-younes, Rémi Cadene, Matthieu Cord, Nicolas Thome

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.07548 2017-04-26 cs.AI cs.LG stat.ML 79%

Semi-supervised Bayesian Deep Multi-modal Emotion Recognition

Changde Du, Changying Du, Jinpeng Li, Wei-long Zheng, Bao-liang Lu, Huiguang He

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.04853 2017-03-16 cs.CV 79%

Face Recognition using Multi-Modal Low-Rank Dictionary Learning

Homa Foroughi, Moein Shakeri, Nilanjan Ray, Hong Zhang

专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.05119 2017-03-13 cs.CV 79%

Deep Impression: Audiovisual Deep Residual Networks for Multimodal Apparent Personality Trait Recognition

Yağmur Güçlütürk, Umut Güçlü, Marcel A. J. van Gerven, Rob van Lier

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1701.08918 2017-02-01 cs.CV 79%

Feature Selection based on PCA and PSO for Multimodal Medical Image Fusion using DTCWT

Padmavathi K, Mahima Bhat, Maya V Karki

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1701.06121 2017-01-24 cs.CV 79%

Multimodal Fusion via a Series of Transfers for Noise Removal

Chang-Hwan Son, Xiao-Ping Zhang

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1509.00244 2016-11-17 cs.CV 79%

Robust Face Recognition via Multimodal Deep Face Representation

Changxing Ding, Dacheng Tao

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments To appear in IEEE Trans. Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.04322 2016-10-17 cs.CV 79%

Learning and Fusing Multimodal Features from and for Multi-task Facial Computing

Wei Li, Zhigang Zhu

专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV

Comments An experiment to feature fusion in deep learning

详情

展开后加载摘要…

URL PDF HTML 收藏