Language-Guided Grasping under Partial Observation for Mobile Manipulation in Field Inspection and Maintenance
用于现场检查和维护中移动操作的部分观察下语言引导抓取
Dilermando Almeida, Juliano Negri, Guilherme Lazzarini, Thiago H. Segreto, Ranulfo Bezerra, Gustavo J. G. Lahr, Ricardo V. Godoy, Marcelo Becker
机构
*
Department of Mechanical Engineering, Federal University of Uberlândia(联邦大学伯南迪利亚机械工程系)
;
Department of Mechanical Engineering, University of São Paulo(圣保罗大学机械工程系)
;
Graduate School of Information Sciences, Tohoku University(东北大学信息科学研究生院)
;
Faculdade Israelita de Ensino e Pesquisa Albert Einstein, Hospital Israelita Albert Einstein(艾伯特·爱因斯坦以色列教学与研究学院,艾伯特·爱因斯坦医院)
Dynamic Object Masks as Goal Representations for Visual Goal-Conditioned Reinforcement Learning
动态目标掩码作为视觉目标条件强化学习的目标表示
Fahim Shahriar, Cheryl Wang, Alireza Azimi, Gautham Vasan, Hany Hamed, Abhishek Naik, A. Rupam Mahmood, Colin Bellinger
机构
*
University of Alberta(阿尔伯塔大学)
;
McGill University(麦吉尔大学)
;
University of Ottawa(渥太华大学)
;
AMII
;
CIFAR Canada AI Chair(CIFAR加拿大人工智能主席)
;
Vector Institute(向量研究所)
机构
*
Shanghai Jiao Tong University China
;
Hohai University China
;
Singapore Management University Singapore
;
Imperial College London United Kingdom
;
East China Normal University \& Shanghai Innovation Institute China
;
Chongqing University China
;
Shanghai Jiao Tong University
;
Hohai University
;
Singapore Management University
;
Imperial College London
;
East China Normal University \& Shanghai Innovation Institute
;
Chongqing University
NBA_Streaming: A Large-Scale Benchmark for Fine-Grained Basketball Commentary Generation in Continuous Streams
NBA_Streaming:用于连续流中细粒度篮球评论生成的大规模基准测试
Lifang Wu, Yuyang Wu, Yangdong Gao, Fengyu Liu, Ya Jing, Liang Wang
机构
*
School of Information Science and Technology, Beijing University of Technology(北京工业大学信息科学与技术学院)
;
Fudan University(复旦大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
Omni-Persona:系统性基准测试与改进多模态个性化
Yeongtak Oh, Dongwook Lee, Sangkwon Park, Heeseung Kim, Sungroh Yoon
机构
*
Department of Electrical and Computer Engineering, Seoul National University(首尔国立大学电气与计算机工程系)
;
Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学人工智能跨学科项目)
;
Department of Artificial Intelligence, University of Seoul(首尔大学人工智能系)
专题命中
视觉定位与Grounding
:grounding(abstract);multimodal large language model(abstract);分类 cs.CV
机构
*
Tsinghua University(清华大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Sun Yat-sen University(中山大学)
;
The University of Hong Kong(香港大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Nanyang Technological University(南洋理工大学)
机构
*
School of Information Science and Technology, Yunnan Normal University(云南师范大学信息科学与技术学院)
;
School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
;
College of Computer Science, Beijing University of Technology(北京工业大学计算机学院)