Wasm: A Pipeline for Constructing Structured Arabic Interleaved Multimodal Corpora
专题命中 图文多模态 :multimodal(title,abstract);image-text(abstract);分类 cs.CL、cs.AI
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 图文多模态 :multimodal(title,abstract);image-text(abstract);分类 cs.CL、cs.AI
机构 * Lingran Song, Yucheng Zhou, Jianbing Shen(作者)
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CV、cs.AI
Comments AAAI 2026
机构 * Department of EECS University of Arkansas(电子工程与科学系 亚拉荷加大学) ; Department of CS Baylor University(计算机科学系 基尔默大学)
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.CL、cs.AI
专题命中 图文多模态 :image-text(title,abstract);分类 cs.CV
Comments This version is incomplete and requires substantial revisions and extensions. We withdraw the paper and plan to submit a thoroughly revised version as a new submission
机构 * Nova School of Business and Economics(诺瓦商业与经济学院) ; Nova School of Business(诺瓦商业学院)
专题命中 图文多模态 :multimodal(title,abstract);分类 cs.AI
Comments Accetped at IEEE BigData 2025, 10 pages, 5 figures, 3 tables
专题命中 图文多模态 :multimodal(title,abstract)
机构 * Technical University of Munich (TUM)(慕尼黑技术大学) ; TUM University Hospital(慕尼黑技术大学医院) ; German Heart Center TUM University Hospital(慕尼黑技术大学医院德国心脏中心) ; Department of Radiology(放射科) ; Klinikum rechts der Isar TUM University Hospital(慕尼黑技术大学医院右岸诊所) ; Uniklinik RWTH Aachen(亚琛工业大学医院) ; HOPPR IL USA(HOPPR美国) ; University of Oxford(牛津大学) ; Imperial College London(伦敦帝国学院)
专题命中 图文多模态 :multimodal(title);分类 cs.CV、cs.CL
Comments Accepted to ML4H 2025 Proceedings
机构 * Paul C. Lauterbur Research Center for Biomedical Imaging(生物医学成像研究中心) ; Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院) ; Pengcheng Laboratory(鹏城实验室) ; University of Chinese Academy of Sciences(中国科学院大学) ; Chinese Medicine Guangdong Laboratory(广东中医药实验室) ; Beijing Chaoyang Hospital, Capital Medical University(首都医科大学北京朝阳医院)
专题命中 图文多模态 :multi-modal(title);分类 cs.CV
机构 * Laboratory for Big Data and Decision, National University of Defense Technology, Changsha 410073, China(大数据与决策实验室,国防科技大学,长沙410073,中国)
专题命中 图文多模态 :multimodal(abstract);image-text(abstract);分类 cs.CV、cs.AI
Comments 15 page, 9 figures, published to PRCV
专题命中 图文多模态 :cross-modal(abstract);image-text(abstract);分类 cs.CV、cs.CL
Comments Accepted by AAAI 2026 as an Oral Presentation (13 pages, 7 figures, 7 tables)
Journal ref AAAI2026
机构 * esieaLab, ETIS Laboratory(esiea实验室,ETIS实验室) ; ESIEA, University of CY Cergy(ESIEA,CY塞克大学) ; ETIS Laboratory, CNR1S, UMR8051(ETIS实验室,CNR1S,UMR8051)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV、cs.AI
Comments 5 pages, 3 figures, ICTAI 2025
专题命中 图文多模态 :cross-modal(abstract);分类 cs.CV、cs.MM
Comments 10 pages
专题命中 图文多模态 :multi-modal(abstract);分类 cs.CL、cs.AI
Comments Accepted at the Ninth Conference on Machine Translation (WMT24), co-located with EMNLP 2024
Journal ref https://aclanthology.org/2024.wmt-1.81/
机构 * Heinrich Heine University of Düsseldorf(海因里希-海涅大学杜塞尔多夫分校)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments Accepted in WACV 2026. Code in https://github.com/HHU-MMBS/clustermine_wacv_official 9 Tables, 11 Figures
机构 * Department of Cyber-Physical Systems, Clark Atlanta University(克雷克阿特拉大学计算机物理系统系) ; Siemens Corporation(西门子公司)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV