arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 3475 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 跨模态检索 3475 篇

1905.12884 2019-05-31 cs.CL cs.HC 79%

M-GWAP: An Online and Multimodal Game With A Purpose in WordPress for Mental States Annotation

Fabio Paolizzo

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CL

Comments 2 figures, 4 tables. The research is supported by the EU through the MUSICAL-MOODS project funded by the Marie Sklodowska-Curie Actions Individual Fellowships Global Fellowships (MSCA-IF-GF) of the Horizon 2020 Programme H2020/2014-2020, REA grant agreement n.659434

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.11139 2019-05-28 cs.CV cs.IR 79%

Label Prediction Framework for Semi-Supervised Cross-Modal Retrieval

Devraj Mandal, Pramod Rao, Soma Biswas

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments 12 pages, 3 tables, 2 figures, 1 algorithm flowchart

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.00357 2019-05-17 cs.CL 79%

Latent Variable Model for Multi-modal Translation

Iacer Calixto, Miguel Rios, Wilker Aziz

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CL

Comments Paper accepted at ACL 2019. Contains 8 pages (11 including references, 13 including appendix), 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.13325 2019-05-16 cs.IR cs.MM 79%

Effective and Efficient Indexing in Cross-Modal Hashing-Based Datasets

Sarawut Markchit, Chih-Yi Chiu

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.02430 2019-05-08 cs.IR cs.CL cs.SI 79%

Interactive Search and Exploration in Online Discussion Forums Using Multimodal Embeddings

Iva Gornishka, Stevan Rudinac, Marcel Worring

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.01304 2019-05-07 cs.LG cs.MM stat.ML 79%

Efficient Discrete Supervised Hashing for Large-scale Cross-modal Retrieval

Tao Yao, Xiangwei Kong, Lianshan Yan, Wenjing Tang, Qi Tian

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.10615 2019-04-25 cs.CV 79%

Understanding Art through Multi-Modal Retrieval in Paintings

Noa Garcia, Benjamin Renoust, Yuta Nakashima

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.08042 2019-04-18 cs.MM cs.IR 79%

Adversarial Cross-Modal Retrieval via Learning and Transferring Single-Modal Similarities

Xin Wen, Zhizhong Han, Xinyu Yin, Yu-Shen Liu

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.07488 2019-04-17 cs.CV 79%

Shared Predictive Cross-Modal Deep Quantization

Erkun Yang, Cheng Deng, Chao Li, Wei Liu, Jie Li, Dacheng Tao

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.00639 2019-04-02 cs.CL 79%

Multimodal Machine Translation with Embedding Prediction

Tosho Hirasawa, Hayahide Yamagishi, Yukio Matsumura, Mamoru Komachi

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CL

Comments 6 pages; NAACL 2019 Student Research Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.08888 2019-02-26 cs.LG cs.CV stat.ML 79%

Medical Multimodal Classifiers Under Scarce Data Condition

Faik Aydin, Maggie Zhang, Michelle Ananda-Rajah, Gholamreza Haffari

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1801.03002 2019-02-21 cs.CV 79%

DeepStyle: Multimodal Search Engine for Fashion and Interior Design

Ivona Tautkute, Tomasz Trzcinski, Aleksander Skorupa, Lukasz Brocki, Krzysztof Marasek

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV

Comments Copyright held by IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

Journal ref IEEE Access 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.00378 2019-02-04 cs.CV 79%

Self-Supervised Visual Representations for Cross-Modal Retrieval

Yash Patel, Lluis Gomez, Marçal Rusiñol, Dimosthenis Karatzas, C. V. Jawahar

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments arXiv admin note: text overlap with arXiv:1807.02110

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.09854 2019-01-31 cs.CV 79%

Multi-modal dialog for browsing large visual catalogs using exploration-exploitation paradigm in a joint embedding space

Indrani Bhattacharya, Arkabandhu Chowdhury, Vikas Raykar

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV

Comments 10 pages including reference, 8 figures. First two authors are equal contributors

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.07702 2019-01-24 cs.CV 79%

Exploring Uncertainty in Conditional Multi-Modal Retrieval Systems

Ahmed Taha, Yi-Ting Chen, Xitong Yang, Teruhisa Misu, Larry Davis

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.11013 2018-12-26 cs.CV 79%

Cycle-Consistent Deep Generative Hashing for Cross-Modal Retrieval

Lin Wu, Yang Wang, Ling Shao

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments To appeared on IEEE Trans. Image Processing. arXiv admin note: text overlap with arXiv:1703.10593 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.02534 2018-11-15 cs.CL 79%

Using Sparse Semantic Embeddings Learned from Multimodal Text and Image Data to Model Human Conceptual Knowledge

Steven Derby, Paul Miller, Brian Murphy, Barry Devereux

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CL

Comments Proceedings of the 22nd Conference on Computational Natural Language Learning (CoNLL 2018), pages 260-270. Brussels, Belgium, October 31 - November 1, 2018. Association for Computational Linguistics

Journal ref Proceedings of the 22nd Conference on Computational Natural Language Learning (CoNLL 2018), pages 260-270. Brussels, Belgium, October 31 - November 1, 2018. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.09617 2018-10-24 cs.CV 79%

How to Read Paintings: Semantic Art Understanding with Multi-Modal Retrieval

Noa Garcia, George Vogiatzis

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1808.08266 2018-08-29 cs.CL 79%

A Visual Attention Grounding Neural Model for Multimodal Machine Translation

Mingyang Zhou, Runxiang Cheng, Yong Jae Lee, Zhou Yu

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
1808.06468 2018-08-21 cs.CY cs.MM 79%

Endogenous and Exogenous Multi-Modal Layers in Context Aware Recommendation Systems for Health

Nitish Nag, Vaibhav Pandey, Ramesh C. Jain

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.MM

Comments 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1808.06277 2018-08-21 cs.MM 79%

An Efficient Approach for Geo-Multimedia Cross-Modal Retrieval

Lei Zhu, Jun Long, Chengyuan Zhang, Ruipeng Chen, Xinpan Yuan, Zhan Yang

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.MM

Comments 27 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.00833 2018-07-27 cs.CV 79%

Learnable PINs: Cross-Modal Embeddings for Person Identity

Arsha Nagrani, Samuel Albanie, Andrew Zisserman

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments To appear in ECCV 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.07998 2018-06-14 cs.CV 79%

Deep Sparse Coding for Invariant Multimodal Halle Berry Neurons

Edward Kim, Darryl Hannan, Garrett Kenyon

专题命中 跨模态检索 :multimodal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.10819 2018-05-01 cs.CV 79%

Learning Cross-Modal Deep Embeddings for Multi-Object Image Retrieval using Text and Sketch

Sounak Dey, Anjan Dutta, Suman K. Ghosh, Ernest Valveny, Josep Lladós, Umapada Pal

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments Accepted at ICPR 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.01223 2018-04-05 cs.CV 79%

Self-Supervised Adversarial Hashing Networks for Cross-Modal Retrieval

Chao Li, Cheng Deng, Ning Li, Wei Liu, Xinbo Gao, Dacheng Tao

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.02488 2018-02-08 cs.CV 79%

SCH-GAN: Semi-supervised Cross-modal Hashing by Generative Adversarial Network

Jian Zhang, Yuxin Peng, Mingkuan Yuan

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments 12 pages, submitted to IEEE Transactions on Cybernetics

详情

展开后加载摘要…

URL PDF HTML 收藏
1802.01943 2018-02-07 cs.CV 79%

Attribute-Guided Network for Cross-Modal Zero-Shot Hashing

Zhong Ji, Yuxin Sun, Yunlong Yu, Yanwei Pang, Jungong Han

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments 9 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.00358 2017-12-04 cs.CV 79%

Unsupervised Generative Adversarial Cross-modal Hashing

Jian Zhang, Yuxin Peng, Mingkuan Yuan

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

Comments 8 pages, accepted by 32th AAAI Conference on Artificial Intelligence (AAAI), 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1605.09696 2017-09-01 cs.CV cs.LG 79%

Generalized Multi-view Embedding for Visual Recognition and Cross-modal Retrieval

Guanqun Cao, Alexandros Iosifidis, Ke Chen, Moncef Gabbouj

专题命中 跨模态检索 :cross-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.04047 2017-07-14 cs.CV 79%

Discrete Multi-modal Hashing with Canonical Views for Robust Mobile Landmark Search

Lei Zhu, Zi Huang, Xiaobai Liu, Xiangnan He, Jingkuan Song, Xiaofang Zhou

专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏