arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 4965 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 多模态生成 4965 篇

2111.01619 2021-11-03 cs.CV cs.LG 57%

StyleGAN of All Trades: Image Manipulation with Only Pretrained StyleGAN

Min Jin Chong, Hsin-Ying Lee, David Forsyth

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00323 2021-10-25 cs.CV cs.CG 57%

Future Urban Scenes Generation Through Vehicles Synthesis

Alessandro Simoni, Luca Bergamini, Andrea Palazzi, Simone Calderara, Rita Cucchiara

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments Accepted at ICPR2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.12761 2021-09-29 cs.CL 57%

OpenViDial 2.0: A Larger-Scale, Open-Domain Dialogue Generation Dataset with Visual Contexts

Shuhe Wang, Yuxian Meng, Xiaoya Li, Xiaofei Sun, Rongbin Ouyang, Jiwei Li

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.03670 2021-08-31 cs.CV 57%

3D Shape Generation and Completion through Point-Voxel Diffusion

Linqi Zhou, Yilun Du, Jiajun Wu

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments Project page: https://alexzhou907.github.io/pvd

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.10410 2021-08-02 cs.LG cs.CV 57%

Deep Generative Learning via Schrödinger Bridge

Gefei Wang, Yuling Jiao, Qian Xu, Yang Wang, Can Yang

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Journal ref ICML, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.06964 2021-07-20 cs.CV 57%

Surgical Instruction Generation with Transformers

Jinglu Zhang, Yinyu Nie, Jian Chang, Jian Jun Zhang

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments Accepted to MICCAI 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.13416 2021-06-30 cs.CV cs.GR 57%

Diversifying Semantic Image Synthesis and Editing via Class- and Layer-wise VAEs

Yuki Endo, Yoshihiro Kanamori

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments Accepted to Pacific Graphics 2020, codes available at https://github.com/endo-yuki-t/DiversifyingSMIS

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.04447 2021-06-09 cs.CL 57%

Reading StackOverflow Encourages Cheating: Adding Question Text Improves Extractive Code Generation

Gabriel Orlanski, Alex Gittens

专题命中 多模态生成 :multimodal(abstract);分类 cs.CL

Comments To be published in ACL-IJCNLP NLP4Prog workshop. (The First Workshop on Natural Language Processing for Programming)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.03484 2021-06-08 cs.CL 57%

BERTGEN: Multi-task Generation through BERT

Faidon Mitzalis, Ozan Caglayan, Pranava Madhyastha, Lucia Specia

专题命中 多模态生成 :multimodal(abstract);分类 cs.CL

Comments Accepted to ACL 2021 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.13994 2021-05-31 cs.CV 57%

Linguistic Structures as Weak Supervision for Visual Scene Graph Generation

Keren Ye, Adriana Kovashka

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments To appear in CVPR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.14109 2021-04-30 cs.MM 57%

Towards Harmonized Regional Style Transfer and Manipulation for Facial Images

Cong Wang, Fan Tang, Yong Zhang, Weiming Dong, Tieru Wu

专题命中 多模态生成 :multi-modal(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.13365 2021-04-28 cs.CV eess.IV 57%

NTIRE 2021 Depth Guided Image Relighting Challenge

Majed El Helou, Ruofan Zhou, Sabine Susstrunk, Radu Timofte

专题命中 多模态生成 :any-to-any(abstract);分类 cs.CV

Comments Code and data available on https://github.com/majedelhelou/VIDIT

Journal ref IEEE Conference on Computer Vision and Pattern Recognition Workshops 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.10993 2021-04-26 eess.IV cs.CV 57%

METGAN: Generative Tumour Inpainting and Modality Synthesis in Light Sheet Microscopy

Izabela Horvath, Johannes C. Paetzold, Oliver Schoppe, Rami Al-Maskari, Ivan Ezhov, Suprosanna Shit, Hongwei Li, Ali Ertuerk, Bjoern H. Menze

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.08842 2021-04-20 cs.NE cs.AI 57%

A Rank based Adaptive Mutation in Genetic Algorithm

Avijit Basak

专题命中 多模态生成 :multimodal(abstract);分类 cs.AI

Journal ref August 2020 International Journal of Computer Applications 175

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.11471 2021-03-23 cs.CV 57%

Conditional Generative Adversarial Networks for Speed Control in Trajectory Simulation

Sahib Julka, Vishal Sowrirajan, Joerg Schloetterer, Michael Granitzer

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06878 2021-03-12 cs.CV cs.GR 57%

Diverse Semantic Image Synthesis via Probability Distribution Modeling

Zhentao Tan, Menglei Chai, Dongdong Chen, Jing Liao, Qi Chu, Bin Liu, Gang Hua, Nenghai Yu

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments Accepted By CVPR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.12255 2021-02-09 cs.CV 57%

TPNet: Trajectory Proposal Network for Motion Prediction

Liangji Fang, Qinhong Jiang, Jianping Shi, Bolei Zhou

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.06803 2021-01-19 cs.CL 57%

Narration Generation for Cartoon Videos

Nikos Papasarantopoulos, Shay B. Cohen

专题命中 多模态生成 :multimodal(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.06775 2021-01-19 eess.IV cs.CV 57%

Symmetric-Constrained Irregular Structure Inpainting for Brain MRI Registration with Tumor Pathology

Xiaofeng Liu, Fangxu Xing, Chao Yang, C. -C. Jay Kuo, Georges ElFakhri, Jonghye Woo

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments Published at MICCAI Brainles 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.07304 2020-12-15 cs.CV 57%

Multi Modal Adaptive Normalization for Audio to Video Generation

Neeraj Kumar, Srishti Goel, Ankur Narang, Brejesh Lall

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.02947 2020-12-08 cs.AI 57%

Neurosymbolic AI for Situated Language Understanding

Nikhil Krishnaswamy, James Pustejovsky

专题命中 多模态生成 :multimodal(abstract);分类 cs.AI

Comments 18 pages + refs, 16 figures, presented at the 8th Annual Conference on Advances in Cognitive Systems (ACS), 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.10727 2020-11-24 cs.CV 57%

Stochastic Talking Face Generation Using Latent Distribution Matching

Ravindra Yadav, Ashish Sardana, Vinay P Namboodiri, Rajesh M Hegde

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments InterSpeech 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.04009 2020-11-05 eess.IV cs.CV 57%

Learning joint segmentation of tissues and brain lesions from task-specific hetero-modal domain-shifted datasets

Reuben Dorent, Thomas Booth, Wenqi Li, Carole H. Sudre, Sina Kafiabadi, Jorge Cardoso, Sebastien Ourselin, Tom Vercauteren

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments MIDL 2019 special issue - Medical Image Analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.14953 2020-10-29 cs.CV 57%

Leveraging Visual Question Answering to Improve Text-to-Image Synthesis

Stanislav Frolov, Shailza Jolly, Jörn Hees, Andreas Dengel

专题命中 多模态生成 :image-text(abstract);分类 cs.CV

Comments Accepted to the LANTERN workshop at COLING 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.03132 2020-10-12 cs.LG cs.CV 57%

Conditional Generative Modeling via Learning the Latent Space

Sameera Ramasinghe, Kanchana Ranasinghe, Salman Khan, Nick Barnes, Stephen Gould

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.02315 2020-10-07 cs.CV 57%

SMILE: Semantically-guided Multi-attribute Image and Layout Editing

Andrés Romero, Luc Van Gool, Radu Timofte

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.05678 2020-09-15 cs.AI 57%

To Root Artificial Intelligence Deeply in Basic Science for a New Generation of AI

Jingan Yang, Yang Peng

专题命中 多模态生成 :multi-modal(abstract);分类 cs.AI

Comments 13 pages; 7 figures; 23 references

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.10436 2020-08-25 cs.CV 57%

Cross-Modality 3D Object Detection

Ming Zhu, Chao Ma, Pan Ji, Xiaokang Yang

专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV

Comments Accepted by WACV 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.04557 2020-07-10 cs.CV cs.RO 57%

Alleviating the Burden of Labeling: Sentence Generation by Attention Branch Encoder-Decoder Network

Tadashi Ogura, Aly Magassouba, Komei Sugiura, Tsubasa Hirakawa, Takayoshi Yamashita, Hironobu Fujiyoshi, Hisashi Kawai

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

Comments 9 pages, 8 figures. accepted for IEEE Robotics and Automation Letters (RA-L) with presentation at IROS 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.03199 2020-07-08 eess.IV cs.CV 57%

Automatic lesion detection, segmentation and characterization via 3D multiscale morphological sifting in breast MRI

Hang Min, Darryl McClymont, Shekhar S. Chandra, Stuart Crozier, Andrew P. Bradley

专题命中 多模态生成 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏