LitForager: Exploring Multimodal Literature Foraging Strategies in Immersive Sensemaking
专题命中 多模态Agent :multimodal(title,abstract)
Comments 11 pages, 10 figures, Accepted to IEEE ISMAR 2025 (TVCG)
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态Agent :multimodal(title,abstract)
Comments 11 pages, 10 figures, Accepted to IEEE ISMAR 2025 (TVCG)
机构 * Laboratory of Brain Atlas and Brain-inspired Intelligence, Institute of Automation Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所脑图谱与类脑智能实验室) ; School of Artificial Intelligence, University of Chinese Academy of Sciences (UCAS)(中国科学院大学人工智能学院) ; School of Systems Science, Beijing Normal University(北京师范大学系统科学学院) ; School of Psychological and Cognitive Sciences & Beijing Key Laboratory of Behavior and Mental Health, Peking University(北京大学心理与认知科学学院) ; IDG/McGovern Institute for Brain Research, Peking University(北京大学IDG/ McGovern脑科学研究院) ; Institute for Artificial Intelligence & Key Laboratory of Machine Perception (Ministry of Education), Peking University(北京大学人工智能研究所) ; School of Future Technology, University of Chinese Academy of Sciences (UCAS)(中国科学院大学未来技术学院)
专题命中 多模态Agent :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
机构 * Dongguk University(东国大学)
专题命中 多模态Agent :multi-modal(abstract);分类 cs.CL
专题命中 多模态Agent :multimodal(abstract);分类 cs.AI
Comments update ICRA 6 page
专题命中 多模态Agent :multimodal(abstract)
Comments 34 pages, 12 figures