机构
*
Department of Computer Science, Durham University(杜伦大学计算机科学系)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
College of Computer Science and Engineering, Dalian Minzu University(大连民族大学计算机科学与工程学院)
;
Department of Engineering Science, University of Oxford(牛津大学工程科学系)
;
Jarvis Research Center, Tencent YouTu Lab(腾讯YouTu实验室 Jarvis 研究中心)
;
School of Engineering Mathematics and Technology, University of Bristol(布里斯托大学工程数学与技术学院)
;
Medical Artificial Intelligence Laboratory, School of Engineering, Westlake University(西湖大学工程学院医学人工智能实验室)
机构
*
Faculty of Dentistry, The University of Hong Kong(香港大学牙科学院)
;
College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院)
;
The Hong Kong University of Science and Technology (GZ)(香港科学与技术大学)
;
School of Biomedical Engineering, Southern Medical University(南方医科大学生物医学工程学院)
;
Singapore University of Technology and Design(新加坡科技与设计大学)
;
University of Auckland(奥克兰大学)
;
University of Science and Technology of China(中国科学技术大学)
;
School of Computer Science, Peking University(北京大学计算机学院)
;
College of Artificial Intelligence, Shenzhen University(深圳大学人工智能学院)
SAMChat: Introducing Chain of Thought Reasoning and GRPO to a Multimodal Small Language Model for Small Scale Remote Sensing
SAMChat:引入链式推理和GRPO以增强小规模遥感遥感小语言模型
Aybora Koksal, A. Aydin Alatan
机构
*
Center for the Image Analysis (OGAM) and Department of Electrical and Electronics Engineering of Middle East Technical University (METU)(图像分析中心(OGAM)和中东部技术大学(METU)电子与电气工程系)
CommentsAccepted to Journal of Selected Topics in Applied Earth Observations and Remote Sensing (JSTARS) Special Issue on Foundation and Large Vision Models for Remote Sensing. Code and dataset are available at https://github.com/aybora/SAMChat
Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
通过多模态视觉序列变压器推进语义未来预测
Efstathios Karypidis, Ioannis Kakogeorgiou, Spyros Gidaris, Nikos Komodakis
机构
*
Archimedes, Athena Research Center(阿基米德研究中心)
;
National Technical University of Athens(希腊国家技术大学)
;
University of Crete(克里特大学)
;
IACM-Forth(第四研究机构(IACM-Forth))
Automated segmentation of pediatric neuroblastoma on multi-modal MRI: Results of the SPPIN challenge at MICCAI 2023
多模态MRI上儿科神经母细胞瘤自动分割:2023年MICCAI SPPIN挑战赛结果
M. A. D. Buser, D. C. Simons, M. Fitski, M. H. W. A. Wijnen, A. S. Littooij, A. H. ter Brugge, I. N. Vos, M. H. A. Janse, M. de Boer, R. ter Maat, J. Sato, S. Kido, S. Kondo, S. Kasai, M. Wodzinski, H. Muller, J. Ye, J. He, Y. Kirchhoff, M. R. Rokkus, G. Haokai, S. Zitong, M. Fernández Patón, D. Veiga-Canuto, D. G. Ellis, M. R. Aizenberg, B. H. M. van der Velden, H. Kuijf, A. De Luca, A. F. W. van der Steeg
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究所)
;
Sofia University(索菲亚大学)
;
University of Macau(澳门大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shenzhen University of Advanced Technology(深圳先进技术大学)
机构
*
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院)
;
Multimedia Laboratory, The Chinese University of Hong Kong(香港中文大学多媒体实验室)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Shenzhen University of Advanced Technology(深圳先进技术大学)
;
CPII under InnoHK(创新香港下的CPII)
Closing the Performance Gap Between AI and Radiologists in Chest X-Ray Reporting
弥合AI与放射科医生在胸部X光报告中的性能差距
Harshita Sharma, Maxwell C. Reynolds, Valentina Salvatelli, Anne-Marie G. Sykes, Kelly K. Horst, Anton Schwaighofer, Maximilian Ilse, Olesya Melnichenko, Sam Bond-Taylor, Fernando Pérez-García, Vamshi K. Mugu, Alex Chan, Ceylan Colak, Shelby A. Swartz, Motassem B. Nashawaty, Austin J. Gonzalez, Heather A. Ouellette, Selnur B. Erdal, Beth A. Schueler, Maria T. Wetscherek, Noel Codella, Mohit Jain, Shruthi Bannur, Kenza Bouzid, Daniel C. Castro, Stephanie Hyland, Panos Korfiatis, Ashish Khandelwal, Javier Alvarez-Valle
LiHRA: A LiDAR-Based HRI Dataset for Automated Risk Monitoring Methods
LiHRA:基于LiDAR的人机交互风险监测数据集
Frederik Plahl, Georgios Katranis, Ilshat Mamaev, Andrey Morozov
机构
*
Proximity Robotics & Automation GmbH(近距机器人与自动化有限公司)
;
Institute of Industrial Automation and Software Engineering, University of Stuttgart(工业自动化与软件工程学院,斯图加特大学)
CommentsPreprint of final paper that will appear in the Proceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025)
MAKO: Meta-Adaptive Koopman Operators for Learning-based Model Predictive Control of Parametrically Uncertain Nonlinear Systems
MAKO:基于元学习的Koopman算子用于参数不确定非线性系统的基于学习的模型预测控制
Minghao Han, Kiwan Wong, Adrian Wing-Keung Law, Xunyuan Yin
机构
*
Water Research Institute (NEWRI), Nanyang Technological University, Singapore(新跃大学水研究 institute(NEWRI))
;
School of Chemistry, Chemical Engineering and Biotechnology, Nanyang Technological University, Singapore(新跃大学化学、化工与生物技术学院)
;
Soft Robotics Lab, ETH Zurich, Switzerland(苏黎世联邦理工学院软机器人实验室)
;
Department of Civil and Environmental Engineering, National University of Singapore, Singapore(新加坡国立大学土木与环境工程系)
MTR-VP: Towards End-to-End Trajectory Planning through Context-Driven Image Encoding and Multiple Trajectory Prediction
MTR-VP: 通过基于上下文的图像编码和多轨迹预测实现端到端轨迹规划
Maitrayee Keskar, Mohan Trivedi, Ross Greer
机构
*
Machine Intelligence, Interaction, and Imagination (Mi 3 ) Laboratory(机器智能、交互与想象实验室)
;
University of California, Merced(加州大学默塞德分校)
;
Laboratory for Intelligent & Safe Automobiles (LISA)(智能与安全汽车实验室)
;
University of California, San Diego(加州大学圣地亚哥分校)
Agentic AI Framework for Individuals with Disabilities and Neurodivergence: A Multi-Agent System for Healthy Eating, Daily Routines, and Inclusive Well-Being
具有残疾和神经多样性个体的代理AI框架:一个用于健康饮食、日常习惯和包容性福祉的多代理系统
Salman Jan, Toqeer Ali Syed, Gohar Ali, Ali Akarma, Mohammad Riyaz Belgaum, Ahmad Ali
机构
*
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知国家重点实验室,中国科学技术大学)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,北京大学计算机学院)
;
CUHK(香港大学)
Unlabeled Data Improves Fine-Grained Image Zero-shot Classification with Multimodal LLMs
未标记数据提升多模态大语言模型在细粒度图像零样本分类中的性能
Yunqi Hong, Sohyun An, Andrew Bai, Neil Y. C. Lin, Cho-Jui Hsieh
机构
*
Computer Science Department, University of California, Los Angeles(加州大学洛杉矶分校计算机科学系)
;
Mechanical and Aerospace Engineering Department, University of California, Los Angeles(加州大学洛杉矶分校机械与航空航天工程系)