Any-to-Any Vision-Language Model for Multimodal X-ray Imaging and Radiological Report Generation
任意到任意的视觉-语言模型用于多模态X射线成像与放射学报告生成
机构 * Department of Diagnostics and Intervention, Biomedical Engineering and Radiation Physics, Umeå University(诊断与介入部门,生物医学工程与放射物理,乌梅大学) ; College of Computer Science and Software Engineering, Shenzhen University(计算机科学与软件工程学院,深圳大学)
专题命中 多模态生成 :multimodal(title,abstract);any-to-any(title);分类 cs.CV、cs.AI
AI总结 本文提出了一种任意到任意的视觉-语言模型,用于多模态X射线成像和放射学报告生成,通过生成高质量图像和语义连贯的报告,提升了医疗领域生成模型的临床应用价值。
Comments arXiv admin note: substantial text overlap with arXiv:2501.04614