arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

视觉大模型 / VLM

视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。

共收录 7464 信号源:cs.CV, cs.AI, cs.LG

1. 视觉定位与Grounding 7464 篇

1808.00417 2020-02-19 cs.AI 57%

Debugging Non-Ground ASP Programs: Technique and Graphical Tools

Carmine Dodaro, Philip Gasteiger, Kristian Reale, Francesco Ricca, Konstantin Schekotihin

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments 27 pages, 6 figures, Under consideration in Theory and Practice of Logic Programming (TPLP)

Journal ref Theory and Practice of Logic Programming 19 (2019) 290-316

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.09602 2020-02-17 cs.CL cs.LG cs.SD eess.AS 57%

Learning Hierarchical Discrete Linguistic Units from Visually-Grounded Speech

David Harwath, Wei-Ning Hsu, James Glass

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments Camera-ready version for ICLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.04186 2020-01-14 cs.AI 57%

Towards Evaluating Plan Generation Approaches with Instructional Texts

Debajyoti Paul Chowdhury, Arghya Biswas, Tomasz Sosnowski, Kristina Yordanova

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.10005 2019-12-23 cs.AI 57%

Does AlphaGo actually play Go? Concerning the State Space of Artificial Intelligence

Holger Lyre

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.10524 2019-11-26 cs.CL cs.LG 57%

Causally Denoise Word Embeddings Using Half-Sibling Regression

Zekun Yang, Tianlin Liu

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments Accepted by AAAI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09944 2019-10-28 cs.CV 57%

Watch, Listen and Tell: Multi-modal Weakly Supervised Dense Event Captioning

Tanzila Rahman, Bicheng Xu, Leonid Sigal

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Journal ref ICCV2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.11287 2019-09-26 cs.CL cs.AI 57%

Task-Oriented Conversation Generation Using Heterogeneous Memory Networks

Zehao Lin, Xinjing Huang, Feng Ji, Haiqing Chen, Ying Zhang

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments Accepted as a long paper at EMNLP-IJCNLP 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.02890 2019-09-26 cs.CL cs.CV 57%

Visually Grounded Neural Syntax Acquisition

Haoyue Shi, Jiayuan Mao, Kevin Gimpel, Karen Livescu

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments ACL 2019. Project page: https://ttic.uchicago.edu/~freda/project/vgnsl/

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09788 2019-09-24 cs.CL cs.AI cs.NE 57%

Visuallly Grounded Generation of Entailments from Premises

Somaye Jafaritazehjani, Albert Gatt, Marc Tanti

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments Proceedings of the 12th International Conference on Natural Language Generation (INLG 2019), 11 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.08927 2019-09-20 cs.CL cs.AI 57%

Extracting Conceptual Knowledge from Natural Language Text Using Maximum Likelihood Principle

Shipra Sharma, Balwinder Sodhi

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments 12 pages, Under review in IEEE TKDE

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.08263 2019-09-19 cs.DC cs.AI cs.DS cs.LO 57%

Distributed Answer Set Coloring: Stable Models Computation via Graph Coloring

Marco De Bortoli

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments In Proceedings ICLP 2019, arXiv:1909.07646

Journal ref EPTCS 306, 2019, pp. 441-451

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.08231 2019-09-19 cs.AI cs.LO 57%

Exploiting Partial Knowledge in Declarative Domain-Specific Heuristics for ASP

Richard Taupe, Konstantin Schekotihin, Peter Schüller, Antonius Weinzierl, Gerhard Friedrich

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments In Proceedings ICLP 2019, arXiv:1909.07646

Journal ref EPTCS 306, 2019, pp. 22-35

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.07141 2019-08-21 cs.AI cs.CL 57%

LogicENN: A Neural Based Knowledge Graphs Embedding Model with Logical Rules

Mojtaba Nayyeri, Chengjin Xu, Jens Lehmann, Hamed Shariat Yazdi

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.02766 2019-08-14 eess.IV cs.CV 57%

Data Efficient Unsupervised Domain Adaptation for Cross-Modality Image Segmentation

Cheng Ouyang, Konstantinos Kamnitsas, Carlo Biffi, Jinming Duan, Daniel Rueckert

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments Accepted by MICCAI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.03846 2019-08-13 cs.MM cs.CV cs.IR 57%

Exploiting Temporal Relationships in Video Moment Localization with Natural Language

Songyang Zhang, Jinsong Su, Jiebo Luo

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.05084 2019-07-12 cs.CL cs.CV 57%

MeetUp! A Corpus of Joint Activity Dialogues in a Visual Environment

Nikolai Ilinykh, Sina Zarrieß, David Schlangen

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments In Proceedings of the 23rd Workshop on the Semantics and Pragmatics of Dialogue (semdial / LondonLogue), London, September 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.01963 2019-06-06 cs.CV 57%

Grounded Human-Object Interaction Hotspots from Video (Extended Abstract)

Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments arXiv admin note: substantial text overlap with arXiv:1812.04558

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.06184 2019-05-16 cs.LO cs.AI 57%

Extensions to Justification Theory

Simon Marynissen

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments 7 pages, extended abstract for LPNMR 2019 doctoral consortium (15th International Conference on Logic Programming and Non-monotonic Reasoning, Saint Joseph's University, Philadelphia, PA (USA))

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.10652 2019-05-10 cs.CV cs.CL 57%

Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions

Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments CVPR 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.02925 2019-05-09 cs.CL cs.CV 57%

ShapeGlot: Learning Language for Shape Differentiation

Panos Achlioptas, Judy Fan, Robert X. D. Hawkins, Noah D. Goodman, Leonidas J. Guibas

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.06587 2019-05-07 cs.CV 57%

Grounded Video Description

Luowei Zhou, Yannis Kalantidis, Xinlei Chen, Jason J. Corso, Marcus Rohrbach

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments CVPR 2019 oral, camera-ready version including appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.13324 2019-05-01 cs.AI cs.RO 57%

Learning from Implicit Information in Natural Language Instructions for Robotic Manipulations

Ozan Arkan Can, Pedro Zuidberg Dos Martires, Andreas Persson, Julian Gaal, Amy Loutfi, Luc De Raedt, Deniz Yuret, Alessandro Saffiotti

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.02794 2019-04-15 cs.CV 57%

VQD: Visual Query Detection in Natural Scenes

Manoj Acharya, Karan Jariwala, Christopher Kanan

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments To appear in NAACL 2019 ( To download the dataset please go to http://www.manojacharya.com/ )

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.04558 2019-04-04 cs.CV 57%

Grounded Human-Object Interaction Hotspots from Video

Tushar Nagarajan, Christoph Feichtenhofer, Kristen Grauman

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.01399 2019-04-03 cs.LG stat.ML 57%

On Geometric Structure of Activation Spaces in Neural Networks

Yuting Jia, Haiwen Wang, Shuo Shao, Huan Long, Yunsong Zhou, Xinbing Wang

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.03408 2019-03-18 cs.CL cs.CV 57%

Beyond task success: A closer look at jointly learning to see, ask, and GuessWhat

Ravi Shekhar, Aashish Venkatesh, Tim Baumgärtner, Elia Bruni, Barbara Plank, Raffaella Bernardi, Raquel Fernández

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments Accepted to NAACL 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.03094 2019-03-08 cs.CL cs.AI 57%

Learning to Speak and Act in a Fantasy Text Adventure Game

Jack Urbanek, Angela Fan, Siddharth Karamcheti, Saachi Jain, Samuel Humeau, Emily Dinan, Tim Rocktäschel, Douwe Kiela, Arthur Szlam, Jason Weston

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.07742 2019-02-22 cs.LG stat.ML 57%

From Language to Goals: Inverse Reinforcement Learning for Vision-Based Instruction Following

Justin Fu, Anoop Korattikara, Sergey Levine, Sergio Guadarrama

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.08006 2019-02-06 cs.CV 57%

Video Object Segmentation with Language Referring Expressions

Anna Khoreva, Anna Rohrbach, Bernt Schiele

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments ACCV 2018: 14th Asian Conference on Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.06737 2018-12-18 cs.LG astro-ph.IM eess.SP stat.ML 57%

Heuristics for Efficient Sparse Blind Source Separation

Christophe Kervazo, Jerome Bobin, Cecile Chenot

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments in Proceedings of iTWIST'18, Paper-ID: 11, Marseille, France, November, 21-23, 2018

详情

展开后加载摘要…

URL PDF HTML 收藏