arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

多模态大模型

跨文本、图像、视频、音频等模态的大模型与学习方法。

共收录 4729 信号源:cs.CV, cs.CL, cs.AI, cs.MM, eess.AS

1. 视频多模态 4729 篇

1704.02163 2017-11-10 cs.CV 57%

Egocentric Video Description based on Temporally-Linked Sequences

Marc Bolaños, Álvaro Peris, Francisco Casacuberta, Sergi Soler, Petia Radeva

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

Comments 19 pages, 10 figures, 3 tables. Submitted to Journal of Visual Communication and Image Representation

详情

展开后加载摘要…

URL PDF HTML 收藏
1710.02566 2017-10-10 cs.CV 57%

CAMREP- Concordia Action and Motion Repository

Kaustubha Mendhurwar, Qing Gu, Vladimir de la Cruz, Sudhir Mudur, Tiberiu Popa

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.04923 2017-08-17 cs.CL cs.LG 57%

mAnI: Movie Amalgamation using Neural Imitation

Naveen Panwar, Shreya Khare, Neelamadhav Gantayat, Rahul Aralikatte, Senthil Mani, Anush Sankaran

专题命中 视频多模态 :cross-modal(abstract);分类 cs.CL

Comments Accepted in ML4Creativity workshop in KDD 2017. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.03958 2017-08-15 cs.CV 57%

Lattice Long Short-Term Memory for Human Action Recognition

Lin Sun, Kui Jia, Kevin Chen, Dit Yan Yeung, Bertram E. Shi, Silvio Savarese

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

Comments ICCV2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1703.10106 2017-08-08 cs.CV 57%

Pose-conditioned Spatio-Temporal Attention for Human Action Recognition

Fabien Baradel, Christian Wolf, Julien Mille

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

Comments 10 pages, project page: https://fabienbaradel.github.io/pose_rgb_attention_human_action

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.02101 2017-08-07 cs.CV 57%

TALL: Temporal Activity Localization via Language Query

Jiyang Gao, Chen Sun, Zhenheng Yang, Ram Nevatia

专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV

Comments ICCV 2017 camera ready (with supplemental material)

详情

展开后加载摘要…

URL PDF HTML 收藏
1708.01034 2017-08-04 cs.CV 57%

What Will I Do Next? The Intention from Motion Experiment

Andrea Zunino, Jacopo Cavazza, Atesh Koul, Andrea Cavallo, Cristina Becchio, Vittorio Murino

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

Comments 2017 IEEE Conference on Computer Vision and Pattern Recognition Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.04408 2017-07-25 cs.AI cs.RO 57%

Incremental learning of high-level concepts by imitation

Mina Alibeigi, Majid Nili Ahmadabadi, Babak Nadjar Araabi

专题命中 视频多模态 :multimodal(abstract);分类 cs.AI

Comments 6 pages, 5 figures, 2 tables, supplementary material, conference

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.06307 2017-07-06 cs.CV 57%

Multi-Scale Saliency Detection using Dictionary Learning

Shubham Pachori

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.08276 2017-06-27 cs.CV 57%

Skeleton-Based Action Recognition Using Spatio-Temporal LSTM Network with Trust Gates

Jun Liu, Amir Shahroudy, Dong Xu, Alex C. Kot, Gang Wang

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.00493 2017-06-05 cs.CV cs.LG 57%

Personalized Pancreatic Tumor Growth Prediction via Group Learning

Ling Zhang, Le Lu, Ronald M. Summers, Electron Kebebew, Jianhua Yao

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1704.00570 2017-04-13 cs.CV 57%

Spatiotemporal Networks for Video Emotion Recognition

Lijie Fan, Yunjie Ke

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

Comments The reason is that the article is just an experimental report without being reviewed carefully. There are some fatal drawbacks in this article and it may not be suitable for being published

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.09401 2017-01-02 cs.CV 57%

Action Recognition Based on Joint Trajectory Maps with Convolutional Neural Networks

Pichao Wang, Wanqing Li, Chuankun Li, Yonghong Hou

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.02447 2016-11-15 cs.CV 57%

Action Recognition Based on Joint Trajectory Maps Using Convolutional Neural Networks

Pichao Wang, Zhaoyang Li, Yonghong Hou, Wanqing Li

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1610.05613 2016-10-19 cs.CV 57%

From Traditional to Modern : Domain Adaptation for Action Classification in Short Social Video Clips

Aditya Singh, Saurabh Saini, Rajvi Shah, P J Narayanan

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

Comments 9 pages, GCPR, 2016

Journal ref Pattern Recognition,38th German Conference, GCPR 2016, Hannover, Germany, September 12-15, 2016, Proceedings,pp 245-257

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.01819 2016-09-08 cs.LG cs.CV 57%

Semantic Video Trailers

Harrie Oosterhuis, Sujith Ravi, Michael Bendersky

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1609.01344 2016-09-07 cs.CV 57%

Vision-based Engagement Detection in Virtual Reality

Ghassem Tofighi, Kaamraan Raahemifar, Maria Frank, Haisong Gu

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

Comments Paper has been published in Digital Media Industry and Academic Forum 2016 (DMIAF 2016) http://ieee-digitalmediaforum.org/

详情

展开后加载摘要…

URL PDF HTML 收藏
1608.02183 2016-09-06 cs.CV 57%

Multiview Cauchy Estimator Feature Embedding for Depth and Inertial Sensor-Based Human Action Recognition

Yanan Guo, Lei Li, Weifeng Liu, Jun Cheng, Dapeng Tao

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

Comments This paper has been withdrawn by the author due to a crucial error

详情

展开后加载摘要…

URL PDF HTML 收藏
cmp-lg/9503016 2016-08-30 cmp-lg cs.CL 57%

Natural Language Interfaces to Databases - An Introduction

I. Androutsopoulos, G. D. Ritchie, P. Thanisch

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CL

Comments 50 pages, uuencoded compressed tar file, containing LaTeX code and .eps figures. Uses a4wide.sty. (No changes in the text. Fixed problem with epsf macro. Use the epsf.sty included in the tar file, not the epsf.sty of the cmp-lg server.)

Journal ref Natural Language Engineering 1:1, 29-81

详情

展开后加载摘要…

URL PDF HTML 收藏
cs/0611104 2016-08-16 cs.NE cs.AI 57%

Learning and discrimination through STDP in a top-down modulated associative memory

Anthony Mouraud, Hélène Paugam-Moisy

专题命中 视频多模态 :multimodal(abstract);分类 cs.AI

Journal ref Proceedings of 14 European Symposium on Artificial Neural Networks (ESANN 2006) (03/2006) 611-616

详情

展开后加载摘要…

URL PDF HTML 收藏
1511.03908 2016-04-22 cs.LG cs.CV cs.NE 57%

Learning Human Identity from Motion Patterns

Natalia Neverova, Christian Wolf, Griffin Lacey, Lex Fridman, Deepak Chandra, Brandon Barbello, Graham Taylor

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

Comments 10 pages, 6 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
1604.05799 2016-04-21 cs.MM cs.CY 57%

Mainstreaming video annotation software for critical video analysis

Matthew Martin, James Charlton, Andy M. Connor

专题命中 视频多模态 :multimodal(abstract);分类 cs.MM

Journal ref Journal of Technologies and Human Usability, 11(3), 1-13 (2015)

详情

展开后加载摘要…

URL PDF HTML 收藏
1604.04784 2016-04-19 cs.CV 57%

ACD: Action Concept Discovery from Image-Sentence Corpora

Jiyang Gao, Chen Sun, Ram Nevatia

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

Comments 8 pages, accepted by ICMR 2016

详情

展开后加载摘要…

URL PDF HTML 收藏
1501.07738 2015-02-02 cs.CV 57%

Co-Regularized Deep Representations for Video Summarization

Olivier Morère, Hanlin Goh, Antoine Veillard, Vijay Chandrasekhar, Jie Lin

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

Comments Video summarization, deep convolutional neural networks, co-regularized restricted Boltzmann machines

详情

展开后加载摘要…

URL PDF HTML 收藏
1403.4232 2014-03-18 cs.CV 57%

Automatic Image Registration in Infrared-Visible Videos using Polygon Vertices

Tanushri Chakravorty, Guillaume-Alexandre Bilodeau, Eric Granger

专题命中 视频多模态 :multimodal(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1403.0829 2014-03-05 cs.CV cs.LG stat.ML 57%

Multiview Hessian regularized logistic regression for action recognition

W. Liu, H. Liu, D. Tao, Y. Wang, Ke Lu

专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV

Comments 13 pages,2 figures, submitted to signal processing

详情

展开后加载摘要…

URL PDF HTML 收藏
1401.3835 2014-01-17 cs.AI 57%

On Action Theory Change

Ivan José Varzinczak

专题命中 视频多模态 :multimodal(abstract);分类 cs.AI

Journal ref Journal Of Artificial Intelligence Research, Volume 37, pages 189-246, 2010

详情

展开后加载摘要…

URL PDF HTML 收藏
1312.4640 2013-12-18 cs.HC cs.AI 57%

A Review of Temporal Aspects of Hand Gesture Analysis Applied to Discourse Analysis and Natural Conversation

Renata Cristina Barros Madeo, Priscilla Koch Wagner, Sarajane Marques Peres

专题命中 视频多模态 :multimodal(abstract);分类 cs.AI

Comments 20 pages, International Journal of Computer Science & Information Technology (IJCSIT) Vol 5, No 4, August 2013

详情

展开后加载摘要…

URL PDF HTML 收藏
1301.3883 2013-01-18 cs.AI 57%

Conversation as Action Under Uncertainty

Tim Paek, Eric J. Horvitz

专题命中 视频多模态 :multimodal(abstract);分类 cs.AI

Comments Appears in Proceedings of the Sixteenth Conference on Uncertainty in Artificial Intelligence (UAI2000)

详情

展开后加载摘要…

URL PDF HTML 收藏
0705.1999 2009-12-01 cs.AI cs.LO 57%

A first-order Temporal Logic for Actions

Camilla Schwind

专题命中 视频多模态 :multi-modal(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏