CommentsThis paper is a development of the visual riddle game with Human-AI interaction, entitled "GuessWhat - Riddle Eye with AI", developed by Ciprian Constantinescu (POLItEHNICA Bucharest), Serena Stan (Instituto Cervantes Bucarest) and Marius Leordeanu (POLITEHNICA Bucharest), which was the winner (1st place) of the NeoArt Connect NAC 2025 Scholarship Program
CommentsThis paper is accepted by the IEEE Internet of Things Journal (IoT-J) for publication in the Special Issue on "Augmented Edge Sensing Intelligence for Low-Altitude IoT Systems"
Embodied4C: Measuring What Matters for Embodied Vision-Language Navigation
Embodied4C: 评估具身视觉-语言导航中至关重要的因素
Tin Stribor Sohn, Maximilian Dillitzer, Jason J. Corso, Eric Sax
机构
*
Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)
;
UAS Esslingen(埃森嫩大学)
;
TU Wien(维也纳技术大学)
;
Dr. Ing. h.c. F. Porsche AG(保时捷股份有限公司)
;
University of Michigan(密歇根大学)
;
Voxel51 Inc.(Voxel51公司)
Deadlock-Free Hybrid RL-MAPF Framework for Zero-Shot Multi-Robot Navigation
无死锁的混合式RL-MAPF框架用于零样本多机器人导航
Haoyi Wang, Licheng Luo, Yiannis Kantaros, Bruno Sinopoli, Mingyu Cai
机构
*
Department of Mechanical Engineering University of California Riverside CA USA(加州大学河滨分校机械工程系)
;
Department of Electrical and Systems Engineering Washington University in St. Louis MO USA(华盛顿大学圣路易斯分校电气与系统工程系)
VERM: Leveraging Foundation Models to Create a Virtual Eye for Efficient 3D Robotic Manipulation
VERM:利用基础模型创建虚拟眼睛以实现高效的3D机器人操作
Yixiang Chen, Yan Huang, Keji He, Peiyan Li, Liang Wang
机构
*
New Laboratory of Pattern Recognition (NLPR), State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别新技术实验室,多模态人工智能系统国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
FiveAges
;
Shandong University(山东大学)
UnReflectAnything: RGB-Only Highlight Removal by Rendering Synthetic Specular Supervision
UnReflectAnything:通过渲染合成镜面监督实现仅RGB的高光去除
Alberto Rota, Mert Kiray, Mert Asim Karaoglu, Patrick Ruhkamp, Elena De Momi, Nassir Navab, Benjamin Busam
机构
*
Politecnico di Milano(米兰理工学院)
;
Technical University of Munich(慕尼黑技术大学)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
;
ImFusion(ImFusion公司)
Comments15 pages, 6 figures, 1 table; accepted for AI-2025 Forty-fifth SGAI International Conference on Artificial Intelligence CAMBRIDGE, ENGLAND 16-18 DECEMBER 2025
VFM-ISRefiner: Towards Better Adapting Vision Foundation Models for Interactive Segmentation of Remote Sensing Images
VFM-ISRefiner: 向更适应交互分割遥感图像的视觉基础模型迈进
Deliang Wang, Peng Liu, Yan Ma, Rongkai Zhuang, Lajiao Chen, Bing Li, Yi Zeng
机构
*
Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航天信息研究所)
;
School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences(中国科学院大学电子电气与通信工程学院)
;
School of Information Science and Technology, Beijing Forestry University(北京林业大学信息科学与技术学院)
VOST-SGG: VLM-Aided One-Stage Spatio-Temporal Scene Graph Generation
VOST-SGG:基于视觉语言模型的一阶段时空场景图生成
Chinthani Sugandhika, Chen Li, Deepu Rajan, Basura Fernando
机构
*
College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算与数据科学学院)
;
Institute of High-Performance Computing, Agency for Science, Technology and Research, Singapore(科学、技术与研究局高性能计算研究所)
;
Centre for Frontier AI Research, Agency for Science, Technology and Research, Singapore(科学、技术与研究局前沿人工智能研究中心)