Felix B. Mueller, Timo Lueddecke, Richard Vogg, Alexander S. Ecker
机构
*
Institute of Computer Science and Campus Institute Data Science, University of Göttingen(计算机科学研究所和校园数据科学研究所,哥廷根大学)
;
Max Planck Institute for Dynamics and Self-Organization(动态与自组织Max Planck研究所)
Physical Autoregressive Model for Robotic Manipulation without Action Pretraining
Zijian Song, Sihan Qin, Tianshui Chen, Liang Lin, Guangrun Wang
机构
*
Sun Yat-sen University(中山大学)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东大数据分析与处理重点实验室)
;
X-Era AI Lab(X-Era人工智能实验室)
;
Guangdong University of Technology(广东工业大学)
MM-Retinal V2: Transfer an Elite Knowledge Spark into Fundus Vision-Language Pretraining
Ruiqi Wu, Na Su, Chenran Zhang, Tengfei Ma, Tao Zhou, Zhiting Cui, Nianfeng Tang, Tianyu Mao, Yi Zhou, Wen Fan, Tianxing Wu, Shenqi Jing, Huazhu Fu
机构
*
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Department of Ophthalmology, The First Affiliated Hospital of Nanjing Medical University(南京医科大学第一附属医院眼科部)
;
School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院)
;
Institute of High-Performance Computing, Agency for Science, Technology and Research(科技研究局高性能计算研究所)
Emerging Properties in Unified Multimodal Pretraining
Chaorui Deng, Deyao Zhu, Kunchang Li, Chenhui Gou, Feng Li, Zeyu Wang, Shu Zhong, Weihao Yu, Xiaonan Nie, Ziang Song, Guang Shi, Haoqi Fan
机构
*
ByteDance Seed(字节跳动种子)
;
Shenzhen Institutes of Advanced Technology(深圳先进技术研究院)
;
Monash University(墨尔本大学)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
UC Santa Cruz(加州大学圣克ruz分校)
机构
*
Beijing Institute of Technology(北京理工大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Nanjing University of Science and Technology(南京理工大学)
;
National University of Singapore(新加坡国立大学)
;
Systems Laboratory of Anhui Province, Hefei University of Technology(安徽省系统实验室,合肥工业大学)
;
National Innovative Institute of Defense Technology(国家创新防御技术研究院)
;
Intelligent Science & Technology Academy of CASIC(中国航天科技集团智能科学与技术学院)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
Finetuning Vision-Language Models as OCR Systems for Low-Resource Languages: A Case Study of Manchu
Yan Hon Michael Chung, Donghyeok Choi
机构
*
Division of Humanities(人文学院)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
History, Religious and Philosophy Academy(历史、宗教与哲学学院)
;
Hong Kong Baptist University(香港 Baptist 大学)
Revisiting Automatic Data Curation for Vision Foundation Models in Digital Pathology
Boqi Chen, Cédric Vincent-Cuaz, Lydia A. Schoenpflug, Manuel Madeira, Lisa Fournier, Vaishnavi Subramanian, Sonali Andani, Samuel Ruiperez-Campillo, Julia E. Vogt, Raphaëlle Luisier, Dorina Thanou, Viktor H. Koelzer, Pascal Frossard, Gabriele Campanella, Gunnar Rätsch
机构
*
Dept. of Computer Science, ETH Zurich, Zurich, Switzerland(苏黎世联邦理工学院计算机科学系)
;
AI Center, ETH Zurich, Zurich, Switzerland(苏黎世联邦理工学院人工智能中心)
;
Signal Processing Laboratory (LTS4), EPFL, Lausanne, Switzerland(日内瓦联邦理工学院信号处理实验室)
;
University of Basel, Basel, Switzerland(巴塞尔大学)
;
University Hospital of Basel, Basel, Switzerland(巴塞尔大学医院)
;
Idiap Research Institute, Martigny, Switzerland(伊迪普研究 institute)
;
Icahn School of Medicine at Mount Sinai, New York, United States(辛辛那提医学中心伊坎医学院)
LLäMmlein: Transparent, Compact and Competitive German-Only Language Models from Scratch
Jan Pfister, Julia Wunderle, Andreas Hotho
机构
*
Data Science Chair Center for Artificial Intelligence and Data Science (CAIDAS)(数据科学主席中心人工智能与数据科学中心(CAIDAS))
;
Julius-Maximilians-Universität Würzburg (JMU)(乌尔姆-马克斯-普朗克大学(JMU))
机构
*
Center on Frontiers of Computing Studies, School of Computer Science, Peking University(前沿计算研究中心,计算机科学学院,北京大学)
;
PKU-Agibot Lab, School of Computer Science, Peking University(北京大学计算机科学学院PKU-Agibot实验室)
;
National Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(国家多媒体信息处理重点实验室,计算机科学学院,北京大学)