CommentsThe 1st place report of 7th LSVOS challenge RVOS track in ICCV 2025. The code is released in Sa2VA repository: https://github.com/bytedance/Sa2VA
CommentsThis submission (arXiv:2510.15349) was mistakenly uploaded as a new article. It was intended to replace our previous work arXiv:2506.03197. All subsequent updates will be made to arXiv:2506.03197
DINO-CVA: A Multimodal Goal-Conditioned Vision-to-Action Model for Autonomous Catheter Navigation
Pedram Fekri, Majid Roshanfar, Samuel Barbeau, Seyedfarzad Famouri, Thomas Looi, Dale Podolsky, Mehrdad Zadeh, Javad Dargahi
机构
*
Gina Cody School of Engineering and Computer Science, Concordia University(甘娜·柯迪工程与计算机科学学院,康科迪亚大学)
;
The Wilfred and Joyce Posluns Centre for Image Guided Innovation & Therapeutic Intervention (PCIGITI) at the Hospital for Sick Children (SickKids)(威廉与乔伊斯·波斯卢斯影像引导创新与治疗干预中心(PCIGITI)(SickKids医院))
;
Electrical and Computer Engineering Department, Kettering University(电气与计算机工程系,凯特林大学)
机构
*
Field Robotics Engineering and Science Hub (FRESH), Illinois Autonomous Farm, University of Illinois at Urbana-Champaign (UIUC), IL(伊利诺伊大学厄巴纳-香槟分校)
;
Mobile Robotics Group, São Carlos School of Engineering, University of São Paulo (EESC-USP), São Carlos, SP, Brazil(圣保罗大学)
AGENTSAFE: Benchmarking the Safety of Embodied Agents on Hazardous Instructions
Zonghao Ying, Le Wang, Yisong Xiao, Jiakai Wang, Yuqing Ma, Jinyang Guo, Zhenfei Yin, Mingchuan Zhang, Aishan Liu, Xianglong Liu
机构
*
SKLCCSE, Beihang University(北京航空航天大学智能科学与技术研究中心)
;
Zhongguancun Laboratory(中关村实验室)
;
The University of Sydney(悉尼大学)
;
Henan University of Science(河南科技大学)
Market-Driven Subset Selection for Budgeted Training
Ashish Jha, Valentin Leplat, AH Phan
机构
*
Skolkovo Institute of Science and Technology(斯克洛夫诺科学与技术研究所)
;
Innopolis University(因诺波利斯大学)
专题命中
幻觉与鲁棒性
:grounding(abstract);分类 cs.AI、cs.LG
CommentsRetitled major revision of the same work (formerly "Market-Based Data Subset Selection -- Principled Aggregation of Multi-Criteria Example Utility"). Abstract and exposition revised; ablations added; theory clarified. Core results unchanged. Supersedes v1; please process as a replacement
机构
*
National University of Singapore(新加坡国立大学)
;
University of Toronto(多伦多大学)
;
Peking University(北京大学)
;
Sichuan University(四川大学)
;
Zhejiang University(浙江大学)
CommentsWithdrawn due to an accidental duplicate submission. This paper (arXiv:2510.15430) was unintentionally submitted as a new entry instead of a new version of our previous work (arXiv:2508.09201)
机构
*
Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative, Institute of Digital Twin, EIT, Ningbo(宁波空间智能与数字衍生关键实验室,数字孪生研究院,EIT,宁波)
;
Shanghai Jiao Tong University(上海交通大学)
;
Hong Kong Polytechnic University(香港理工大学)
;
Meituan Inc.(美团公司)
;
National University of Singapore(新加坡国立大学)
专题命中
VLM训练与架构
:LLaVA(abstract);multimodal large language model(abstract);分类 cs.CV
机构
*
Institute of Software Chinese Academy of Sciences(中国科学院软件研究所)
;
University of the Chinese Academy of Sciences(中国科学院大学)
;
Beijing University of Technology(北京理工大学)