arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 4965 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态生成 4965 篇

2006.10649 2020-06-19 cs.CV 57%

Multi-Density Sketch-to-Image Translation Network

Jialu Huang, Jing Liao, Zhifeng Tan, Sam Kwong

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments 2020 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.12318 2020-05-27 cs.CV 57%

Identity-Preserving Realistic Talking Face Generation

Sanjana Sinha, Sandika Biswas, Brojeshwar Bhowmick

专题命中 多模态生成 :audio-visual(abstract);分类 cs.CV

Comments Accepted in IJCNN 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.10687 2020-05-22 eess.IV cs.CV cs.LG 57%

Medical Image Generation using Generative Adversarial Networks

Nripendra Kumar Singh, Khalid Raza

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments 19 pages, 3 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.06993 2020-05-15 cs.LG cs.SD eess.AS 57%

deepSELF: An Open Source Deep Self End-to-End Learning Framework

Tomoya Koike, Kun Qian, Björn W. Schuller, Yoshiharu Yamamoto

专题命中 多模态生成 :multi-modal(abstract);分类 eess.AS

Comments 4 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.00999 2020-04-17 cs.CV cs.LG eess.IV 57%

Cycle In Cycle Generative Adversarial Networks for Keypoint-Guided Image Generation

Hao Tang, Dan Xu, Gaowen Liu, Wei Wang, Nicu Sebe, Yan Yan

专题命中 多模态生成 :cross-modal(abstract);分类 cs.CV

Comments 9 pages, 8 figures, accepted to ACM MM 2019

Journal ref ACM MM 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.03891 2020-04-09 cs.LG cs.CV stat.ML 57%

Normalizing Flows with Multi-Scale Autoregressive Priors

Shweta Mahajan, Apratim Bhattacharyya, Mario Fritz, Bernt Schiele, Stefan Roth

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments To appear in CVPR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.07009 2020-04-06 cs.CV 57%

C-Flow: Conditional Generative Flow Models for Images and 3D Point Clouds

Albert Pumarola, Stefan Popov, Francesc Moreno-Noguer, Vittorio Ferrari

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.03360 2020-04-03 cs.CV 57%

Episode-based Prototype Generating Network for Zero-Shot Learning

Yunlong Yu, Zhong Ji, Zhongfei Zhang, Jungong Han

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.01941 2020-03-05 cs.LG cs.CV stat.ML 57%

Gaussianization Flows

Chenlin Meng, Yang Song, Jiaming Song, Stefano Ermon

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments AISTATS 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.01320 2019-12-09 cs.CV 57%

Synthetic Video Generation for Robust Hand Gesture Recognition in Augmented Reality Applications

Varun Jain, Shivam Aggarwal, Suril Mehta, Ramya Hebbalaguppe

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments Presented at the ICCV 2019 Workshop: The 5th International Workshop on Observing And Understanding Hands In Action

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.05424 2019-12-06 cs.CL 57%

VizSeq: A Visual Analysis Toolkit for Text Generation Tasks

Changhan Wang, Anirudh Jain, Danlu Chen, Jiatao Gu

专题命中 多模态生成 :multimodal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.08212 2019-11-20 cs.CL 57%

Deep Poetry: A Chinese Classical Poetry Generation System

Yusen Liu, Dayiheng Liu, Jiancheng Lv

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CL

Comments Association for the Advancement of Artificial Intelligence, Demonstrations Program. AAAI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.07273 2019-11-19 cs.CV 57%

Coherent and Controllable Outfit Generation

Kedan Li, Chen Liu, David Forsyth

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.06571 2019-11-07 cs.LG cs.CV stat.ML 57%

Latent Dirichlet Allocation in Generative Adversarial Networks

Lili Pan, Shen Cheng, Jian Liu, Yazhou Ren, Zenglin Xu

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.09280 2019-09-17 cs.CV 57%

Points2Pix: 3D Point-Cloud to Image Translation using conditional Generative Adversarial Networks

Stefan Milz, Martin Simon, Kai Fischer, Maximillian Pöpperl

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.13590 2019-08-30 eess.IV cs.CV 57%

Unsupervised Domain Adaptation via Disentangled Representations: Application to Cross-Modality Liver Segmentation

Junlin Yang, Nicha C. Dvornek, Fan Zhang, Julius Chapiro, MingDe Lin, James S. Duncan

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.10534 2019-08-29 cs.CV 57%

Adversarial Representation Learning for Text-to-Image Matching

Nikolaos Sarafianos, Xiang Xu, Ioannis A. Kakadiaris

专题命中 多模态生成 :cross-modal(abstract);分类 cs.CV

Comments To appear in ICCV 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.02149 2019-07-05 cs.CV cs.LG 57%

Analyzing the Cross-Sensor Portability of Neural Network Architectures for LiDAR-based Semantic Labeling

Florian Piewak, Peter Pinggera, Marius Zöllner

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.01524 2019-06-05 cs.CV cs.GR cs.LG 57%

Text-based Editing of Talking-head Video

Ohad Fried, Ayush Tewari, Michael Zollhöfer, Adam Finkelstein, Eli Shechtman, Dan B Goldman, Kyle Genova, Zeyu Jin, Christian Theobalt, Maneesh Agrawala

专题命中 多模态生成 :audio-visual(abstract);分类 cs.CV

Comments A version with higher resolution images can be downloaded from the authors' website

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.09400 2019-05-24 cs.CV 57%

AttentionRNN: A Structured Spatial Attention Mechanism

Siddhesh Khandelwal, Leonid Sigal

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.01280 2019-05-21 cs.CV 57%

Learning to Explain with Complemental Examples

Atsushi Kanehira, Tatsuya Harada

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments Camera ready version of CVPR'19

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.13281 2019-05-01 eess.IV cs.CV cs.LG 57%

CT-To-MR Conditional Generative Adversarial Networks for Ischemic Stroke Lesion Segmentation

Jonathan Rubin, S. Mazdak Abulnaga

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments Seventh IEEE International Conference on Healthcare Informatics (ICHI 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.09326 2019-04-18 cs.CV 57%

Making History Matter: History-Advantage Sequence Training for Visual Dialog

Tianhao Yang, Zheng-Jun Zha, Hanwang Zhang

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.03692 2019-03-26 cs.CV cs.LG 57%

Mode matching in GANs through latent space learning and inversion

Deepak Mishra, Prathosh A. P., Aravind Jayendran, Varun Srivastava, Santanu Chaudhury

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.04352 2019-03-12 cs.CV 57%

Joint inference on structural and diffusion MRI for sequence-adaptive Bayesian segmentation of thalamic nuclei with probabilistic atlases

Juan Eugenio Iglesias, Koen Van Leemput, Polina Golland, Anastasia Yendiki

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.03477 2019-03-11 cs.CV 57%

Auto-Encoding Progressive Generative Adversarial Networks For 3D Multi Object Scenes

Vedant Singh, Manan Oza, Himanshu Vaghela, Pratik Kanani

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.00913 2019-03-05 cs.CV 57%

Unsupervised Bi-directional Flow-based Video Generation from one Snapshot

Lu Sheng, Junting Pan, Jiaming Guo, Jing Shao, Xiaogang Wang, Chen Change Loy

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments 11 pages, 12 figures. Technical report for a project in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.00746 2019-01-23 cs.CV cs.LG 57%

Bayesian Prediction of Future Street Scenes using Synthetic Likelihoods

Apratim Bhattacharyya, Mario Fritz, Bernt Schiele

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments To appear in ICLR 2019. arXiv admin note: text overlap with arXiv:1806.06939

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.00053 2018-11-02 cs.CV 57%

DEEPGONET: Multi-label Prediction of GO Annotation for Protein from Sequence Using Cascaded Convolutional and Recurrent Network

Sheikh Muhammad Saiful Islam, Md Mahedi Hasan

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments Accepted in ICCIT 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.04413 2018-09-25 cs.CV 57%

Enhancing clinical MRI Perfusion maps with data-driven maps of complementary nature for lesion outcome prediction

Adriano Pinto, Sergio Pereira, Raphael Meier, Victor Alves, Roland Wiest, Carlos A. Silva, Mauricio Reyes

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments Accepted at MICCAI 2018

详情

展开后加载摘要…

URL PDF HTML 收藏