RL-U$^2$Net: A Dual-Branch UNet with Reinforcement Learning-Assisted Multimodal Feature Fusion for Accurate 3D Whole-Heart Segmentation
专题命中 多模态训练与对齐 :multimodal(title);multi-modal(abstract);cross-modal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态训练与对齐 :multimodal(title);multi-modal(abstract);cross-modal(abstract);分类 cs.CV
机构 * Zhejiang Key Laboratory of Accessible Perception and Intelligent Systems, Zhejiang University(浙江可感知智能系统重点实验室,浙江大学) ; Alibaba Group(阿里巴巴集团) ; Hangzhou High-Tech Zone (Binjiang) Institute of Blockchain and DataSecurity(杭州高新技术区(滨江)区块链与数据安全研究院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.AI
Comments Accepted at EMNLP 2025
机构 * Image Processing Lab, Sharif University of Technology(沙斐大学技术实验室)
专题命中 多模态训练与对齐 :cross-modal(title,abstract);分类 cs.CV
Comments 11 pages, 5 figures, 8 tables
机构 * Ant Group(蚂蚁集团)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract)
Comments 12 pages, 5 figures
机构 * Columbia University(哥伦比亚大学)
专题命中 多模态训练与对齐 :multimodal(abstract);cross-modal(abstract);分类 cs.AI
Comments More videos can be found on our website:https://binghao-huang.github.io/touch_in_the_wild/