arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6929 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6929 篇

1902.00038 2019-02-13 cs.CV 57%

BLOCK: Bilinear Superdiagonal Fusion for Visual Question Answering and Visual Relationship Detection

Hedi Ben-younes, Rémi Cadene, Nicolas Thome, Matthieu Cord

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.01936 2019-01-24 cs.CV 57%

Disentangled Variational Representation for Heterogeneous Face Recognition

Xiang Wu, Huaibo Huang, Vishal M. Patel, Ran He, Zhenan Sun

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments AAAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.11903 2018-11-30 cs.CV 57%

Visual Question Answering as Reading Comprehension

Hui Li, Peng Wang, Chunhua Shen, Anton van den Hengel

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.07392 2018-11-20 cs.CV 57%

Facial Expression and Peripheral Physiology Fusion to Decode Individualized Affective Experience

Yu Yin, Mohsen Nabian, Miolin Fan, ChunAn Chou, Maria Gendron, Sarah Ostadabbas

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 2nd IJCAI Workshop on Artificial Intelligence in Affective Computing

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.02447 2018-11-07 cs.CV 57%

Multi-Level Sensor Fusion with Deep Learning

Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux, Frédéric Jurie

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:1808.07275

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.01701 2018-11-06 cs.AI q-bio.NC 57%

Role of Awareness and Universal Context in a Spiking Conscious Neural Network (SCNN): A New Perspective and Future Directions

Ahsan Adeel

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI

Comments Neural Computation | MIT Press Journals (In-Process)

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.01206 2018-10-25 cs.LG cs.CV cs.IR 57%

A Survey of Multi-View Representation Learning

Yingming Li, Ming Yang, Zhongfei Zhang

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Knowledge and Data Engineering

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.11024 2018-10-03 cs.CV 57%

Adversarial Image Registration with Application for MR and TRUS Image Fusion

Pingkun Yan, Sheng Xu, Ardeshir R. Rastinehad, Brad J. Wood

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments Presented at the workshop on MLMI 2018, LNCS, volume 11046, pages 197 to 204

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.02560 2018-10-02 cs.CV cs.GR cs.LG 57%

A Deeper Look at 3D Shape Classifiers

Jong-Chyi Su, Matheus Gadelha, Rui Wang, Subhransu Maji

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments Accepted to Second Workshop on 3D Reconstruction Meets Semantics, ECCV 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.07941 2018-09-24 cs.CV 57%

LIDAR-Camera Fusion for Road Detection Using Fully Convolutional Neural Networks

Luca Caltagirone, Mauro Bellone, Lennart Svensson, Mattias Wahde

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.04558 2018-09-13 cs.LG cs.AI cs.RO 57%

Coordinated Heterogeneous Distributed Perception based on Latent Space Representation

Timo Korthals, Jürgen Leitner, Ulrich Rückert

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

Comments IROS 2018 Second Workshop on Multi-robot Perception-Driven Control and Planning

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.06230 2018-08-24 cs.CV cs.RO 57%

Robust Fusion of LiDAR and Wide-Angle Camera Data for Autonomous Mobile Robots

Varuna De Silva, Jamie Roche, Ahmet Kondoz

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1808.03195 2018-08-10 cs.CV 57%

Overcoming Missing and Incomplete Modalities with Generative Adversarial Networks for Building Footprint Segmentation

Benjamin Bischke, Patrick Helber, Florian König, Damian Borth, Andreas Dengel

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1808.01752 2018-08-07 cs.CV 57%

Deep Transfer Learning for EEG-based Brain Computer Interface

Chuanqi Tan, Fuchun Sun, Wenchang Zhang

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments In Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 2018, 15-20 April 2018, Alberta, Canada

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.09072 2018-07-25 cs.CV 57%

Feature Fusion through Multitask CNN for Large-scale Remote Sensing Image Segmentation

Shihao Sun, Lei Yang, Wenjie Liu, Ruirui Li

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.00382 2018-07-23 cs.CV 57%

Automatic Brain Tumor Segmentation using Cascaded Anisotropic Convolutional Neural Networks

Guotai Wang, Wenqi Li, Sebastien Ourselin, Tom Vercauteren

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments 12 pages, 5 figures. MICCAI Brats Challenge 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.02199 2018-04-09 cs.CV stat.ML 57%

Mix and match networks: encoder-decoder alignment for zero-pair image translation

Yaxing Wang, Joost van de Weijer, Luis Herranz

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments Accepted CVPR 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.09374 2018-04-09 cs.LG cs.CV stat.ML 57%

Generalized Hadamard-Product Fusion Operators for Visual Question Answering

Brendan Duke, Graham W. Taylor

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 8 pages, 3 figures. To appear in CRV, 2018, 15th Canadian Conference on Computer and Robot Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.09789 2018-03-28 cs.AI 57%

On Chatbots Exhibiting Goal-Directed Autonomy in Dynamic Environments

Biplav Srivastava

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

Comments 3 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.07490 2018-03-14 cs.RO cs.CV cs.GR 57%

ViTac: Feature Sharing between Vision and Tactile Sensing for Cloth Texture Recognition

Shan Luo, Wenzhen Yuan, Edward Adelson, Anthony G. Cohn, Raul Fuentes

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments 6 pages, 5 figures, Accepted for 2018 IEEE International Conference on Robotics and Automation

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.05861 2017-11-17 cs.CV 57%

Modal Regression based Atomic Representation for Robust Face Recognition

Yulong Wang, Yuan Yan Tang, Luoqing Li, Hong Chen

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 10 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.04528 2017-11-13 cs.CV 57%

A deep learning model integrating FCNNs and CRFs for brain tumor segmentation

Xiaomei Zhao, Yihong Wu, Guidong Song, Zhenye Li, Yazhuo Zhang, Yong Fan

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments This version was accepted in the journal Medical Image Analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
1709.08203 2017-09-26 cs.CV 57%

Survey of Recent Advances in Visual Question Answering

Supriya Pandhre, Shagun Sodhani

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments 7 pages, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.08117 2017-08-29 cs.CV 57%

Part-to-whole Registration of Histology and MRI using Shape Elements

Jonas Pichat, Juan Eugenio Iglesias, Sotiris Nousias, Tarek Yousry, Sebastien Ourselin, Marc Modat

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments Paper accepted at ICCV Workshop (Bio-Image Computing)

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.04733 2017-08-18 cs.LG cs.AI stat.ML 57%

Geometric Enclosing Networks

Trung Le, Hung Vu, Tu Dinh Nguyen, Dinh Phung

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1603.06668 2017-08-15 cs.CV 57%

Learning Representations for Automatic Colorization

Gustav Larsson, Michael Maire, Gregory Shakhnarovich

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments ECCV 2016 (Project page: http://people.cs.uchicago.edu/~larsson/colorization/)

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.00584 2017-08-03 cs.CV 57%

A Simple Loss Function for Improving the Convergence and Accuracy of Visual Question Answering Models

Ilija Ilievski, Jiashi Feng

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments accepted at CVPR 2017 VQA workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.02596 2017-05-09 cs.CV 57%

Simultaneous Super-Resolution and Cross-Modality Synthesis of 3D Medical Images using Weakly-Supervised Joint Convolutional Sparse Coding

Yawen Huang, Ling Shao, Alejandro F. Frangi

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 10 pages, 6 figures. Accepted by CVPR 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.06084 2017-04-21 cs.AI cs.LG stat.ML 57%

Knowledge Fusion via Embeddings from Text, Knowledge Graphs, and Images

Steffen Thoma, Achim Rettinger, Fabian Both

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.08810 2016-10-05 cs.CL 57%

Effective Combination of Language and Vision Through Model Composition and the R-CCA Method

Hagar Loeub, Roi Reichart

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CL

Comments 6 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏