arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 6929 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态训练与对齐 6929 篇

1909.02410 2020-02-28 cs.CV 57%

Semantic-Aware Scene Recognition

Alejandro López-Cifuentes, Marcos Escudero-Viñolo, Jesús Bescós, Álvaro García-Martín

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments Paper submitted for publication to Elsevier Pattern Recognition journal

Journal ref Pattern Recognition Volume 102, June 2020, 107256

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.11430 2020-02-27 cs.CV 57%

Deform-GAN:An Unsupervised Learning Model for Deformable Registration

Xiaoyue Zhang, Weijian Jian, Yu Chen, Shihting Yang

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.10373 2020-02-25 cs.AI 57%

Symbolic Learning and Reasoning with Noisy Data for Probabilistic Anchoring

Pedro Zuidberg Dos Martires, Nitesh Kumar, Andreas Persson, Amy Loutfi, Luc De Raedt

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.08670 2020-02-21 cs.CV 57%

Stroke Constrained Attention Network for Online Handwritten Mathematical Expression Recognition

Jiaming Wang, Jun Du, Jianshu Zhang

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.03688 2020-02-11 eess.IV cs.CV cs.LG 57%

Knowledge Distillation for Brain Tumor Segmentation

Dmitrii Lachinov, Elena Shipunova, Vadim Turlapov

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.05512 2020-01-08 cs.CV 57%

MIMA: MAPPER-Induced Manifold Alignment for Semi-Supervised Fusion of Optical Image and Polarimetric SAR Data

Jingliang Hu, Danfeng Hong, Xiao Xiang Zhu

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.00225 2019-12-12 eess.IV cs.CV 57%

A Semantic-based Medical Image Fusion Approach

Fanda Fan, Yunyou Huang, Lei Wang, Xingwang Xiong, Zihan Jiang, Zhifei Zhang, Jianfeng Zhan

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.07472 2019-11-19 cs.CV 57%

Learning to Synthesize Fashion Textures

Wu Shi, Tak-Wai Hui, Ziwei Liu, Dahua Lin, Chen Change Loy

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.07807 2019-11-19 eess.IV cs.CV cs.LG 57%

A fully 3D multi-path convolutional neural network with feature fusion and feature weighting for automatic lesion identification in brain MRI images

Yunzhe Xue, Meiyan Xie, Fadi G. Farhat, Olga Boukrina, A. M. Barrett, Jeffrey R. Binder, Usman W. Roshan, William W. Graves

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments Machine Learning for Health (ML4H) at NeurIPS 2019 - Extended Abstract

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.05625 2019-11-14 cs.CV cs.LG eess.SP 57%

Twins Recognition Using Hierarchical Score Level Fusion

Cihan Akin, Umit Kacar, Murvet Kirci

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 4 pages, 5 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.02754 2019-10-08 cs.CL cs.LG 57%

On Leveraging the Visual Modality for Neural Machine Translation

Vikas Raunak, Sang Keun Choe, Quanyang Lu, Yi Xu, Florian Metze

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CL

Comments Accepted to INLG 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.01803 2019-10-07 cs.LG cs.AI stat.ML 57%

Unsupervised Representation for EHR Signals and Codes as Patient Status Vector

Sajad Darabi, Mohammad Kachuee, Majid Sarrafzadeh

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.01763 2019-10-07 cs.CV cs.LG cs.NE 57%

NeurReg: Neural Registration and Its Application to Image Segmentation

Wentao Zhu, Andriy Myronenko, Ziyue Xu, Wenqi Li, Holger Roth, Yufang Huang, Fausto Milletari, Daguang Xu

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments WACV 2020 first round early accept; supplementary https://drive.google.com/file/d/1kzTLQn8cpoQNAYWUDJMtN5HcqhbWbl7G/view?usp=sharing; code will be released soon under NVIDIA open source; demos https://www.youtube.com/watch?v=GYLD7t7dSAg&t=3s

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09252 2019-09-24 cs.LG cs.CV cs.DC stat.ML 57%

HyperLearn: A Distributed Approach for Representation Learning in Datasets With Many Modalities

Devanshu Arya, Stevan Rudinac, Marcel Worring

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09309 2019-09-23 cs.CV 57%

CNN-based RGB-D Salient Object Detection: Learn, Select and Fuse

Hao Chen, Youfu Li

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments submitted to a journal in 12-October-2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.03958 2019-09-20 eess.IV cs.CV cs.LG 57%

Structural Similarity based Anatomical and Functional Brain Imaging Fusion

Nishant Kumar, Nico Hoffmann, Martin Oelschlägel, Edmund Koch, Matthias Kirsch, Stefan Gumhold

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments Accepted at MICCAI-MBIA 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.07623 2019-09-18 cs.CV 57%

Deep End-to-End Alignment and Refinement for Time-of-Flight RGB-D Module

Di Qiu, Jiahao Pang, Wenxiu Sun, Chengxi Yang

专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV

Comments ICCV2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.06585 2019-09-18 cs.RO cs.CV 57%

Deep Robotic Prediction with hierarchical RGB-D Fusion

Yaoxian Song, Jun Wen, Yuejiao Fei, Changbin Yu

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 8 pages, 8 figures, submit to ICRA2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.10818 2019-08-29 cs.MM cs.SI 57%

False News Detection on Social Media

Juan Cao, Qiang Sheng, Peng Qi, Lei Zhong, Yanyan Wang, Xueyao Zhang

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.MM

Comments 4 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.05252 2019-08-27 cs.CV eess.IV 57%

Dynamic Fusion with Intra- and Inter- Modality Attention Flow for Visual Question Answering

Gao Peng, Zhengkai Jiang, Haoxuan You, Pan Lu, Steven Hoi, Xiaogang Wang, Hongsheng Li

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments CVPR 2019 ORAL

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.07191 2019-08-21 cs.CV 57%

Make a Face: Towards Arbitrary High Fidelity Face Manipulation

Shengju Qian, Kwan-Yee Lin, Wayne Wu, Yangxiaokang Liu, Quan Wang, Fumin Shen, Chen Qian, Ran He

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments Accepted to ICCV 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.09077 2019-05-21 cs.CV 57%

Incorporating Luminance, Depth and Color Information by a Fusion-based Network for Semantic Segmentation

Shang-Wei Hung, Shao-Yuan Lo, Hsueh-Ming Hang

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments Accepted in IEEE International Conference on Image Processing (ICIP) 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.02859 2019-05-21 cs.CV 57%

Exploiting 2D Floorplan for Building-scale Panorama RGBD Alignment

Erik Wijmans, Yasutaka Furukawa

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.04510 2019-05-14 cs.CV 57%

Deep Zero-Shot Learning for Scene Sketch

Yao Xie, Peng Xu, Zhanyu Ma

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments 5 pages, 3 figures, IEEE International Conference on Image Processing (ICIP)

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.12428 2019-04-30 cs.CV 57%

Attribute Guided Unpaired Image-to-Image Translation with Semi-supervised Learning

Xinyang Li, Jie Hu, Shengchuan Zhang, Xiaopeng Hong, Qixiang Ye, Chenglin Wu, Rongrong Ji

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.07003 2019-04-30 cs.CV 57%

3D-SIS: 3D Semantic Instance Segmentation of RGB-D Scans

Ji Hou, Angela Dai, Matthias Nießner

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments video: https://youtu.be/IH9rNLD1-JE

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.00457 2019-03-15 cs.RO cs.CV 57%

AgriColMap: Aerial-Ground Collaborative 3D Mapping for Precision Farming

Ciro Potena, Raghav Khanna, Juan Nieto, Roland Siegwart, Daniele Nardi, Alberto Pretto

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments Published in IEEE Robotics and Automation Letters, 2019

Journal ref IEEE Robotics and Automation Letters, Vol: 4, Issue: 2, April 2019, pages 1085-1092

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.02930 2019-03-08 cs.CL 57%

Neural Language Modeling with Visual Features

Antonios Anastasopoulos, Shankar Kumar, Hank Liao

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.01659 2019-03-06 cs.RO cs.CV 57%

Vision-Depth Landmarks and Inertial Fusion for Navigation in Degraded Visual Environments

Shehryar Khattak, Christos Papachristos, Kostas Alexis

专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV

Comments 11 pages, 6 figures, Published in International Symposium on Visual Computing (ISVC) 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.01673 2019-02-26 cs.CV 57%

Recurrent Convolutional Fusion for RGB-D Object Recognition

Mohammad Reza Loghmani, Mirco Planamente, Barbara Caputo, Markus Vincze

专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV

Comments Under review at RA-L

详情

展开后加载摘要…

URL PDF HTML 收藏