arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

VLA / 视觉-语言-动作模型

视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。

共收录 9832 信号源:cs.RO, cs.CV, cs.AI, cs.LG

1. VLA模型 9157 篇

2305.09212 2023-05-17 eess.AS cs.CV cs.MM cs.SD 57%

Cross-Modal Global Interaction and Local Alignment for Audio-Visual Speech Recognition

Yuchen Hu, Ruizhe Li, Chen Chen, Heqing Zou, Qiushi Zhu, Eng Siong Chng

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments 12 pages, 5 figures, Accepted by IJCAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.07138 2023-04-14 cs.AI cs.CL 57%

Integrating AI Planning with Natural Language Processing: A Combination of Explicit and Tacit Knowledge

Kebing Jin, Hankz Hankui Zhuo

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.02714 2023-04-07 cs.CV eess.SP 57%

Learning Stage-wise GANs for Whistle Extraction in Time-Frequency Spectrograms

Pu Li, Marie Roch, Holger Klinck, Erica Fleishman, Douglas Gillespie, Eva-Marie Nosal, Yu Shiu, Xiaobai Liu

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted by IEEE Transactions of Multimedia (2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.10434 2023-04-06 cs.CV cs.MM 57%

Frame-wise Cross-modal Matching for Video Moment Retrieval

Haoyu Tang, Jihua Zhu, Meng Liu, Zan Gao, Zhiyong Cheng

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments 12 pages; accepted by IEEE TMM

Journal ref IEEE Transactions on Multimedia 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.00150 2023-04-04 cs.LG physics.flu-dyn 57%

E($3$) Equivariant Graph Neural Networks for Particle-Based Fluid Mechanics

Artur P. Toshev, Gianluca Galletti, Johannes Brandstetter, Stefan Adami, Nikolaus A. Adams

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments ICLR 2023 Workshop on Physics for Machine Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10876 2023-03-28 cs.CV cs.MA 57%

EqMotion: Equivariant Multi-agent Motion Prediction with Invariant Interaction Reasoning

Chenxin Xu, Robby T. Tan, Yuhong Tan, Siheng Chen, Yu Guang Wang, Xinchao Wang, Yanfeng Wang

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted to CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10904 2023-03-22 cs.CV 57%

Actionlet-Dependent Contrastive Learning for Unsupervised Skeleton-Based Action Recognition

Lilang Lin, Jiahang Zhang, Jiaying Liu

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted by CVPR2023 (Highlight). The project page is at https://langlandslin.github.io/projects/ActCLR/

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02078 2023-03-21 cs.CL cs.AI 57%

FGSI: Distant Supervision for Relation Extraction method based on Fine-Grained Semantic Information

Chenghong Sun, Weidong Ji, Guohui Zhou, Hui Guo, Zengxiang Yin, Yuqi Yue

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.07202 2023-03-14 cs.AI math.OC 57%

Optimization of the location and design of urban green spaces

Caroline Leboeuf, Margarida Carvalho, Yan Kestens, Benoît Thierry

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.02073 2023-03-06 cs.LG 57%

How To Guide Your Learner: Imitation Learning with Active Adaptive Expert Involvement

Xu-Hui Liu, Feng Xu, Xinyu Zhang, Tianyuan Liu, Shengyi Jiang, Ruifeng Chen, Zongzhang Zhang, Yang Yu

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.14506 2023-02-27 cs.AI 57%

An extension of process calculus for asynchronous communications between agents with epistemic states

Huili Xing

专题命中 VLA模型 :action model(abstract);分类 cs.AI

Comments 22 pages and 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.01367 2023-02-06 stat.ML cs.LG stat.ME 57%

Augmented Learning of Heterogeneous Treatment Effects via Gradient Boosting Trees

Heng Chen, Michael L. LeBlanc, James Y. Dai

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.11159 2022-12-23 cs.IR cs.LG 57%

Directed Acyclic Graph Factorization Machines for CTR Prediction via Knowledge Distillation

Zhen Tian, Ting Bai, Zibin Zhang, Zhiyuan Xu, Kangyi Lin, Ji-Rong Wen, Wayne Xin Zhao

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.06271 2022-12-23 cs.IT cs.LG math.IT 57%

Bit-Metric Decoding Rate in Multi-User MIMO Systems: Theory

K. Pavan Srinath, Jakob Hoydis

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments This is the first part of a two-part paper and has 30 pages and 6 figures. The second part is titled "Bit-Metric Decoding Rate in Multi-User MIMO Systems: Applications". This part has been significantly revised. In particular, we relate BMDR to the mismatched decoding framework that exists in the literature, and have rewritten the claims and proof of the main theorem

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.05286 2022-12-13 cs.RO 57%

A Survey of Multi-Agent Human-Robot Interaction Systems

Abhinav Dahiya, Alexander M. Aroyo, Kerstin Dautenhahn, Stephen L. Smith

专题命中 VLA模型 :action model(abstract);分类 cs.RO

Comments 23 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.00843 2022-12-05 cs.CV cs.CL 57%

Focus! Relevant and Sufficient Context Selection for News Image Captioning

Mingyang Zhou, Grace Luo, Anna Rohrbach, Zhou Yu

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Findings of EMNLP 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.09716 2022-11-18 cs.RO 57%

A Flexible MATLAB/Simulink Simulator for Robotic Floating-base Systems in Contact with the Ground

Nuno Guedelha, Venus Pasandi, Giuseppe L'Erario, Silvio Traversaro, Daniele Pucci

专题命中 VLA模型 :action model(abstract);分类 cs.RO

Comments To be published in IEEE-IRC 2022 proceedings, 5 pages with 6 figures, equal contribution by authors Nuno Guedelha and Venus Pasandi

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.10367 2022-11-01 cs.SD cs.LG eess.AS 57%

Multi-View Attention Transfer for Efficient Speech Enhancement

Wooseok Shin, Hyun Joon Park, Jin Sob Kim, Byung Hoon Lee, Sung Won Han

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments Proceedings of Interspeech 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15219 2022-10-25 cs.RO 57%

AR Point&Click: An Interface for Setting Robot Navigation Goals

Morris Gu, Elizabeth Croft, Akansel Cosgun

专题命中 VLA模型 :action model(abstract);分类 cs.RO

Comments Accepted at ICSR 2022 "14th International Conference on Social Robotics", 6 Pages, 5 Figures, 4 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.11020 2022-10-21 cs.LG cs.IR 57%

Maximum Common Subgraph Guided Graph Retrieval: Late and Early Interaction Networks

Indradyumna Roy, Soumen Chakrabarti, Abir De

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Journal ref NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.05863 2022-10-21 cs.LG physics.chem-ph q-bio.MN q-bio.QM 57%

GEM-2: Next Generation Molecular Property Prediction Network by Modeling Full-range Many-body Interactions

Lihang Liu, Donglong He, Xiaomin Fang, Shanzhuo Zhang, Fan Wang, Jingzhou He, Hua Wu

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.01443 2022-10-19 cond-mat.mes-hall cs.LG 57%

Learning Coulomb Diamonds in Large Quantum Dot Arrays

Oswin Krause, Anasua Chatterjee, Ferdinand Kuemmeth, Evert van Nieuwenburg

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Journal ref SciPost Phys. 13, 084 (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.06829 2022-10-14 cs.CL cs.LG 57%

Ensemble Creation via Anchored Regularization for Unsupervised Aspect Extraction

Pulah Dhandekar, Manu Joseph

专题命中 VLA模型 :action model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.05391 2022-10-14 cs.CV 57%

PP-StructureV2: A Stronger Document Analysis System

Chenxia Li, Ruoyu Guo, Jun Zhou, Mengtao An, Yuning Du, Lingfeng Zhu, Yi Liu, Xiaoguang Hu, Dianhai Yu

专题命中 VLA模型 :action model(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04992 2022-10-13 cs.CL cs.AI 57%

Extracting or Guessing? Improving Faithfulness of Event Temporal Relation Extraction

Haoyu Wang, Hongming Zhang, Yuqian Deng, Jacob R. Gardner, Dan Roth, Muhao Chen

专题命中 VLA模型 :action model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.07824 2022-10-11 stat.ML cs.LG q-bio.GN 57%

Large-Scale Differentiable Causal Discovery of Factor Graphs

Romain Lopez, Jan-Christian Hütter, Jonathan K. Pritchard, Aviv Regev

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments 33 pages, 12 figures

Journal ref Advances in Neural Information Processing Systems 35 (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.10007 2022-10-11 cs.AI 57%

Gradient-Based Mixed Planning with Symbolic and Numeric Action Parameters

Kebing Jin, Hankz Hankui Zhuo, Zhanhao Xiao, Hai Wan, Subbarao Kambhampati

专题命中 VLA模型 :action model(abstract);分类 cs.AI

Comments 41 pages, 22 figures. Accepted by Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.14690 2022-09-30 cs.CV 57%

Prompt-guided Scene Generation for 3D Zero-Shot Learning

Majid Nasiri, Ali Cheraghian, Townim Faisal Chowdhury, Sahar Ahmadi, Morteza Saberi, Shafin Rahman

专题命中 VLA模型 :action model(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.05054 2022-09-21 eess.IV cs.CV 57%

A No-reference Quality Assessment Metric for Point Cloud Based on Captured Video Sequences

Yu Fan, Zicheng Zhang, Wei Sun, Xiongkuo Min, Wei Lu, Tao Wang, Ning Liu, Guangtao Zhai

专题命中 VLA模型 :action model(abstract);分类 cs.CV

Comments Accepted to IEEE 24th International Workshop on Multimedia Signal Processing, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.08026 2022-09-19 cs.SI cs.DM cs.LG math.PR physics.soc-ph 57%

Interactions in Information Spread

Gaël Poux-Médard

专题命中 VLA模型 :action model(abstract);分类 cs.LG

Comments PhD thesis defended on 2022/09/13

详情

展开后加载摘要…

URL PDF HTML 收藏