A Survey of Multimodal Ophthalmic Diagnostics: From Task-Specific Approaches to Foundational Models
机构 * College of Computer Science and Software Engineering, Shenzhen University, Shenzhen, China(深圳大学计算机科学与软件工程学院) ; Shenzhen Key Laboratory of Visual Object Detection and Recognition, Harbin Institute of Technology, Shenzhen, 518055, China(视觉对象检测与识别深圳重点实验室) ; Laboratory for Artificial Intelligence in Design, Hong Kong(人工智能设计实验室) ; School of Artificial Intelligence, Shenzhen University, Shenzhen, China(深圳大学人工智能学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);multimodal foundation model(abstract);分类 cs.CV、cs.AI