Frequency-Aware Vision-Language Multimodality Generalization Network for Remote Sensing Image Classification
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
机构 * The Alan Turing Institute(艾伦·图灵研究所)
专题命中 图文多模态 :multimodal(abstract);分类 cs.AI
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments AAAI2026, with supplementary material
专题命中 图文多模态 :multi-modal(abstract);分类 cs.CV
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted in AAAI 2026
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted to AAAI-2026 Oral
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments 7 pages, 5 figures
机构 * Heinrich Heine University of Düsseldorf(海因里希-海涅大学杜塞尔多夫分校)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments Accepted in WACV 2026. Code in https://github.com/HHU-MMBS/clustermine_wacv_official 9 Tables, 11 Figures
机构 * Department of Cyber-Physical Systems, Clark Atlanta University(克雷克阿特拉大学计算机物理系统系) ; Siemens Corporation(西门子公司)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
机构 * College of Computer Science and Technology, Zhejiang University-University of Illinois Urbana-Champaign Institute(浙江大学计算机科学与技术学院) ; Stomatology Hospital, School of Stomatology, Zhejiang University School of Medicine(浙江大学口腔医院) ; Alibaba Inc(阿里巴巴集团) ; College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院) ; Angelalign Technology Inc.(Angelalign技术有限公司) ; CFAR & IHPC, Agency for Science, Technology and Research(CFAR与IHPC,新加坡科技研究局) ; Department of Orthodontics, Shanghai Ninth People’s Hospital, College of Stomatology, Shanghai Jiao Tong University(上海第九人民医院正畸科,上海交通大学口腔医学院)
专题命中 图文多模态 :cross-modal(abstract);分类 cs.CV
机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国先进科学技术研究院)
专题命中 图文多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted at TMLR 2025
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
Comments Need more refinement
机构 * The University of Manchester(曼彻斯特大学)
专题命中 图文多模态 :multimodal(abstract);分类 cs.AI
机构 * University of Amsterdam(阿姆斯特丹大学)
专题命中 图文多模态 :multimodal(abstract);分类 cs.AI
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments Published in Transactions on Machine Learning Research (09/2025)
Journal ref Transactions on Machine Learning Research (09/2025)
机构 * School of Computer Science and Engineering, Central South University(中南大学计算机科学与工程学院) ; Miner School of Computer & Information Sciences, University of Massachusetts Lowell(马萨诸塞大学洛厄尔分校Miner计算机与信息科学学院)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2025
机构 * Department of Architecture, National University of Singapore(建筑系,新加坡国立大学) ; Department of Electrical and Computer Engineering, National University of Singapore(电气与计算机工程系,新加坡国立大学) ; School of Artificial Intelligence, Shenzhen Technology University(人工智能学院,深圳科技大学) ; Department of Real Estate, National University of Singapore(房地产系,新加坡国立大学)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
Journal ref ISPRS Journal of Photogrammetry and Remote Sensing 230: 918-942, 2025
专题命中 图文多模态 :multimodal(abstract);分类 cs.AI
Comments 15 pages, 10 figures, 7 tables, conference
机构 * School of Computing, Macquarie University, Australia(计算机学院,麦考瑞大学,澳大利亚)
专题命中 图文多模态 :multimodal(abstract);分类 cs.AI
机构 * State Key Laboratory of Complex and Critical Software Environment, Beihang University(复杂与关键软件环境国家重点实验室,北京航空航天大学) ; School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院) ; School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院)
专题命中 图文多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2025
机构 * Department of Medicine I, LMU University Hospital, LMU Munich, Germany(慕尼黑大学医学院第一医学部,LMU慕尼黑大学医院) ; Lunit Inc.(Lunit公司)
专题命中 图文多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted to Computer Vision for Automated Medical Diagnosis (CVAMD) Workshop at ICCV 2025
机构 * Department of Horticultural Sciences(园艺科学系) ; Gulf Coast Research and Education Center(墨西哥湾沿岸研究与教育中心) ; University of Florida(佛罗里达大学) ; Department of Agricultural and Biological Engineering(农业与生物工程系) ; Department of Soil, Water, and Ecosystem Sciences(土壤、水与生态系统科学系) ; Citrus Research and Education Center(柑橘研究与教育中心)
专题命中 图文多模态 :multi-modal(abstract);分类 cs.AI
Comments 26 pages, 8 figures, and 2 tables
机构 * University of Chinese Academy of Science, Beijing,100190, China(中国科学院大学) ; Macquarie University(麦考瑞大学)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments Published in Pattern Recognition
机构 * Department of Computer Science(计算机科学系)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments 23 pages, 10 figures, 14 tables
机构 * Korea University(韩国大学) ; The Catholic University of Korea(韩国天主大学)
专题命中 图文多模态 :image-text(abstract);分类 cs.CV
Comments Accepted to ICCV 2025. Code is available at: https://github.com/QuIIL/ICCV2025_Ano-NAViLa
专题命中 图文多模态 :cross-modal(abstract);分类 cs.CV
机构 * DRAGON Lab at Department of Mechanical Engineering, The University of Tokyo(东京大学机械工程系DRAGON实验室)
专题命中 图文多模态 :multimodal(abstract);分类 cs.AI
Comments 18 pages, 10 figures
Journal ref Advanced Intelligent Systems, Oct. 2025
机构 * School of Computer Science, University of South China(南方大学计算机科学学院) ; New Laboratory of Pattern Recognition, MAIS, CASIA(模式识别新实验室,MAIS,CASIA) ; School of Intelligence Science and Technology, Nanjing University(智能科学与技术学院,南京大学) ; Department of Computer Science and Engineering, University of California, Merced(加州大学默塞德分校计算机科学与工程系) ; Department of Computer Science and Engineering, Yonsei University(延世大学计算机科学与工程系)
专题命中 图文多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted by TPAMI
机构 * Key Lab of Intell. Info. Process., Inst. of Comput. Tech., CAS(智能信息处理重点实验室,计算技术研究所,中国科学院) ; University of Chinese Academy of Sciences(中国科学院大学)
专题命中 图文多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2025
专题命中 图文多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted into the ICAAI 2025 - The 9th International Conference on Advances in Artificial Intelligence