Delineate Anything v2: A Global Foundation Model for Field Delineation
描绘一切 v2:用于田地描绘的全球基础模型
Mykola Lavreniuk, Nataliia Kussul, Andrii Shelestov, Yevhenii Salii, Volodymyr Kuzin, Charlotte Julia Li-Xing Wang, Zoltan Szantoi
机构
*
European Space Agency(欧洲航天局)
;
Space Research Institute NASU-SSAU(乌克兰国家科学院-乌克兰国家空间局空间研究所)
;
University of Maryland(马里兰大学)
;
National Technical University of Ukraine “Igor Sikorsky Kyiv Polytechnic Institute”(乌克兰国立技术大学“ Igor Sikorsky基辅理工学院”)
Hebaixu Wang, Jing Zhang, Haonan Guo, Di Wang, Jiayi Ma, Bo Du, Liangpei Zhang
机构
*
School of Electronic Information, Wuhan University(武汉大学电子信息学院)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
;
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing, Wuhan University(武汉大学测绘遥感信息工程国家重点实验室)
;
Zhongguancun Academy(中关村学院)
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation
迈向医学数据的视觉-语言基础模型:越南语PET/CT报告生成的多模态数据集和基准
Huu Tien Nguyen, Dac Thai Nguyen, The Minh Duc Nguyen, Trung Thanh Nguyen, Thao Nguyen Truong, Huy Hieu Pham, Johan Barthelemy, Minh Quan Tran, Thanh Tam Nguyen, Quoc Viet Hung Nguyen, Quynh Anh Chau, Hong Son Mai, Thanh Trung Nguyen, Phi Le Nguyen
机构
*
AI4LIFE, Hanoi University of Science and Technology, Vietnam(AI4LIFE,河内科学技术大学,越南)
;
Nagoya University, Japan(名古屋大学,日本)
;
AIST, Japan(日本国家先进工业技术研究院)
;
VinUniversity, Vietnam(文园大学,越南)
;
NVIDIA, USA(NVIDIA,美国)
;
Griffith University, Australia(格里菲斯大学,澳大利亚)
;
Hanoi Medical University, Vietnam(河内医学院,越南)
;
Military Central Hospital, Vietnam(越南108中央军医院)
Prompt-Guided Foundation Model Tuning for Pathology Image Classification
用于病理图像分类的提示引导基础模型调优
Yi Lin, Zhengjie Zhu, Kwang-Ting Cheng, Hao Chen
机构
*
Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系)
;
Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(香港科技大学电子与计算机工程系)
;
Department of Chemical and Biological Engineering, The Hong Kong University of Science and Technology(香港科技大学化学与生物工程系)
;
HKUST Shenzhen-Hong Kong Collaborative Innovation Research Institute(香港科技大学深圳-香港协同创新研究院)
;
State Key Laboratory of Nervous System Disorders, The Hong Kong University of Science and Technology(香港科技大学神经系统疾病国家重点实验室)
CommentsTo appear at ICSE 2026. 13 pages. The L-AVRBench benchmark, Docker images, and evaluation scripts are available at https://github.com/rimwoohan/L-AVRBench
机构
*
Rutgers University(罗格斯大学)
;
Michigan State University(密歇根州立大学)
;
JD Logistics(京东物流)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
Audio-Language Models for Audio-Centric Tasks: A Systematic Survey
用于以音频为中心任务的音频-语言模型:系统综述
Yi Su, Jisheng Bai, Qisheng Xu, Kele Xu, Yong Dou
机构
*
College of Computer Science and Technology, National University of Defense Technology(计算机科学与技术学院,国防科技大学)
;
School of Communications and Information Engineering, Xi’an University of Posts and Telecommunications(通信与信息工程学院,西安邮电大学)
机构
*
Department of Breast Pathology and Laboratory, Tianjin Medical University Cancer Institute & Hospital, National Clinical Research Center for Cancer, Key Laboratory of Breast Cancer Prevention and Therapy, Tianjin Medical University, Ministry of Education, Tianjin’s Clinical Research Center for Cancer, West Huanhu Road, Tianjin, China(天津医科大学肿瘤医院乳腺病理科及实验室,国家癌症临床研究中心,天津医科大学乳腺癌预防与治疗重点实验室,天津医科大学,教育部,天津癌症临床研究中心,西湖南路,天津,中国)
;
Guangdong Provincial Key Laboratory of Artificial Intelligence in Medical Image Analysis and Application, Guangdong Provincial People’s Hospital (Guangdong Academy of Medical Sciences), Southern Medical University, Guangzhou, China(广东省人工智能在医学影像分析与应用重点实验室,广东省人民医院(广东省医学科学院),南方医科大学,广州,中国)