arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

视觉大模型 / VLM

视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。

共收录 7464 信号源:cs.CV, cs.AI, cs.LG

1. 视觉定位与Grounding 7464 篇

2108.07732 2021-08-18 cs.PL cs.LG 57%

Program Synthesis with Large Language Models

Jacob Austin, Augustus Odena, Maxwell Nye, Maarten Bosma, Henryk Michalewski, David Dohan, Ellen Jiang, Carrie Cai, Michael Terry, Quoc Le, Charles Sutton

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments Jacob and Augustus contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.09285 2021-07-21 cs.CL cs.AI 57%

Neural Abstructions: Abstractions that Support Construction for Grounded Language Learning

Kaylee Burns, Christopher D. Manning, Li Fei-Fei

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments 17 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.07314 2021-07-13 cs.LG stat.ML 57%

Long-tail learning via logit adjustment

Aditya Krishna Menon, Sadeep Jayasumana, Ankit Singh Rawat, Himanshu Jain, Andreas Veit, Sanjiv Kumar

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments Published as a conference paper in ICLR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.02951 2021-07-08 cs.LG stat.ML 57%

Universal Approximation for Log-concave Distributions using Well-conditioned Normalizing Flows

Holden Lee, Chirag Pabbaraju, Anish Sevekari, Andrej Risteski

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments 40 pages, 0 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.11147 2021-07-01 cs.DB cs.AI 57%

Harmless but Useful: Beyond Separable Equality Constraints in Datalog+/-

Luigi Bellomarini, Emanuel Sallinger

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.12352 2021-06-18 cs.CV cs.CL 57%

Seeing past words: Testing the cross-modal capabilities of pretrained V&L models on counting tasks

Letitia Parcalabescu, Albert Gatt, Anette Frank, Iacer Calixto

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments Paper accepted for publication at MMSR 2021; 13 pages, 3 figures, 7 Tables

Journal ref Proceedings of the 1st Workshop on Multimodal Semantic Representations (MMSR), 2021, Groningen, Netherlands (Online), Association for Computational Linguistics, p. 32--44

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.10306 2021-06-16 stat.ML cs.CY cs.LG stat.AP 57%

An Empirical Characterization of Fair Machine Learning For Clinical Risk Prediction

Stephen R. Pfohl, Agata Foryciarz, Nigam H. Shah

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments Published in the Journal of Biomedical Informatics (https://doi.org/10.1016/j.jbi.2020.103621). Version 3 updates acknowledgements and fixes typos

Journal ref Journal of Biomedical Informatics, Volume 113, January 2021, 103621

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.06939 2021-06-15 cs.CV 57%

Cross-Modal Attention Consistency for Video-Audio Unsupervised Learning

Shaobo Min, Qi Dai, Hongtao Xie, Chuang Gan, Yongdong Zhang, Jingdong Wang

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.06138 2021-06-14 cs.CV 57%

Team RUC_AIM3 Technical Report at ActivityNet 2021: Entities Object Localization

Ludan Ruan, Jieting Chen, Yuqing Song, Shizhe Chen, Qin Jin

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments 6 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.04403 2021-06-10 cs.CV cs.CL cs.MM 57%

SynthRef: Generation of Synthetic Referring Expressions for Object Segmentation

Ioannis Kazakos, Carles Ventura, Miriam Bellver, Carina Silberer, Xavier Giro-i-Nieto

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments Accepted as poster at the NAACL 2021 Visually Grounded Interaction and Language (ViGIL) Workshop. 4 pages. Project website: https://imatge-upc.github.io/synthref/

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.14207 2021-06-01 cs.CL cs.AI 57%

Maintaining Common Ground in Dynamic Environments

Takuma Udagawa, Akiko Aizawa

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments Accepted at TACL; pre-MIT Press publication version

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.11541 2021-05-26 cs.CV 57%

Learning Better Visual Dialog Agents with Pretrained Visual-Linguistic Representation

Tao Tu, Qing Ping, Govind Thattai, Gokhan Tur, Prem Natarajan

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.05964 2021-05-14 cs.CV 57%

Connecting What to Say With Where to Look by Modeling Human Attention Traces

Zihang Meng, Licheng Yu, Ning Zhang, Tamara Berg, Babak Damavandi, Vikas Singh, Amy Bearman

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.13367 2021-04-20 cs.CV 57%

SoccerNet-v2: A Dataset and Benchmarks for Holistic Understanding of Broadcast Soccer Videos

Adrien Deliège, Anthony Cioppa, Silvio Giancola, Meisam J. Seikavandi, Jacob V. Dueholm, Kamal Nasrollahi, Bernard Ghanem, Thomas B. Moeslund, Marc Van Droogenbroeck

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments Paper accepted for the CVsports workshop at CVPR2021. This document contains 8 pages + references + supplementary material

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.03135 2021-04-09 cs.CV 57%

Seeing Out of tHe bOx: End-to-End Pre-training for Vision-Language Representation Learning

Zhicheng Huang, Zhaoyang Zeng, Yupan Huang, Bei Liu, Dongmei Fu, Jianlong Fu

专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV

Comments Accepted by CVPR2021 oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.01190 2021-04-06 cs.AI 57%

grASP: A Graph Based ASP-Solver and Justification System

Fang Li, Huaduo Wang, Gopal Gupta

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.09977 2021-03-19 cs.AI cs.CL 57%

Situated Language Learning via Interactive Narratives

Prithviraj Ammanabrolu, Mark O. Riedl

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments Preprint. Under journal review

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.07435 2021-03-04 stat.ML cs.LG 57%

A Unified Framework for Random Forest Prediction Error Estimation

Benjamin Lu, Johanna Hardin

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments 41 pages, 8 figures, 6 tables

Journal ref Journal of Machine Learning Research, 22(8):1-41, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.11073 2021-02-23 eess.SP cs.LG eess.IV 57%

Determination of Fault Location in Transmission Lines with Image Processing and Artificial Neural Networks

Serkan Budak, Bahadir Akbal

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments in Turkish language

Journal ref Konya Journal of Engineering Sciences (2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.06112 2021-02-12 cs.AI 57%

A Metamodel and Framework for Artificial General Intelligence From Theory to Practice

Hugo Latapie, Ozkan Kilic, Gaowen Liu, Yan Yan, Ramana Kompella, Pei Wang, Kristinn R. Thorisson, Adam Lawrence, Yuhong Sun, Jayanth Srinivasa

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments arXiv admin note: text overlap with arXiv:2008.12879

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.14285 2020-12-29 cs.CY cs.LG 57%

Affirmative Algorithms: The Legal Grounds for Fairness as Awareness

Daniel E. Ho, Alice Xiang

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG

Comments 12 pages, 3 figures

Journal ref 10/30/20 U. Chi. L. Rev. Online 143, https://lawreviewblog.uchicago.edu/2020/10/30/aa-ho-xiang/

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.12146 2020-12-24 cs.CV 57%

Spatially Aware Multimodal Transformers for TextVQA

Yash Kant, Dhruv Batra, Peter Anderson, Alex Schwing, Devi Parikh, Jiasen Lu, Harsh Agrawal

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments Accepted at European Conference on Computer Vision, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.09216 2020-12-18 cs.CL cs.CV 57%

MELINDA: A Multimodal Dataset for Biomedical Experiment Method Classification

Te-Lin Wu, Shikhar Singh, Sayan Paul, Gully Burns, Nanyun Peng

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments In The Thirty-Fifth AAAI Conference on Artificial Intelligence (AAAI-21), 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.05684 2020-12-14 cs.CV 57%

AttnGrounder: Talking to Cars with Attention

Vivek Mittal

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.02947 2020-12-08 cs.AI 57%

Neurosymbolic AI for Situated Language Understanding

Nikhil Krishnaswamy, James Pustejovsky

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments 18 pages + refs, 16 figures, presented at the 8th Annual Conference on Advances in Cognitive Systems (ACS), 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.13297 2020-11-30 cs.AI 57%

Totally and Partially Ordered Hierarchical Planners in PDDL4J Library

Damien Pellier, Humbert Fiorino

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments 2 pages

Journal ref Proceedings of the International Planning Competition, ICAPS, Nancy, France, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.09663 2020-11-20 cs.CV cs.MM 57%

Modeling Fashion Influence from Photos

Ziad Al-Halah, Kristen Grauman

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments To appear in the IEEE Transactions on Multimedia, 2020. Project page: https://www.cs.utexas.edu/~ziad/influence_from_photos.html. arXiv admin note: substantial text overlap with arXiv:2004.01316

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.09413 2020-10-21 cs.CV cs.CL 57%

Image Captioning with Visual Object Representations Grounded in the Textual Modality

Dušan Variš, Katsuhito Sudoh, Satoshi Nakamura

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.07212 2020-10-14 cs.CV 57%

Revisiting Image-Language Networks for Open-ended Phrase Detection

Bryan A. Plummer, Kevin J. Shih, Yichen Li, Ke Xu, Svetlana Lazebnik, Stan Sclaroff, Kate Saenko

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV

Comments Accepted to TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.03205 2020-10-08 cs.CL cs.AI 57%

Like hiking? You probably enjoy nature: Persona-grounded Dialog with Commonsense Expansions

Bodhisattwa Prasad Majumder, Harsh Jhamtani, Taylor Berg-Kirkpatrick, Julian McAuley

专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI

Comments Accepted in EMNLP 2020

详情

展开后加载摘要…

URL PDF HTML 收藏