On the Comprehensibility of Multi-structured Financial Documents using LLMs and Pre-processing Tools
专题命中 其他多模态 :multi-modal(abstract);MLLM(abstract)
Comments 15 pages, 5 figures, 9 tables
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 其他多模态 :multi-modal(abstract);MLLM(abstract)
Comments 15 pages, 5 figures, 9 tables
机构 * University of Washington(华盛顿大学) ; Google Research(谷歌研究) ; UCLA(加州大学洛杉矶分校) ; Google DeepMind(谷歌DeepMind)
专题命中 其他多模态 :multimodal(abstract);分类 cs.CV、cs.AI
Comments Accepted to the ICCV'25 Workshop "Vision Foundation Models and Generative AI for Accessibility: Challenges and Opportunities"
机构 * The University of Adelaide(阿德莱德大学) ; Macquarie University(麦考瑞大学) ; CSIRO’s Data 61 and The University of New South Wales(CSIRO的数据61与新南威尔士大学)
专题命中 其他多模态 :multimodal(abstract);分类 cs.AI
Comments 9 pages
专题命中 其他多模态 :multi-modal(abstract);分类 cs.AI
Comments Accepted to AAAI/ACM AIES 2025
专题命中 其他多模态 :multi-modal(abstract)