OmniVaT: Single Domain Generalization for Multimodal Visual-Tactile Learning
OmniVaT:单域泛化用于多模态视觉-触觉学习
机构 * Fujian Key Laboratory for Intelligent Processing and Wireless Transmission of Media Information(福建智能媒体信息处理与无线传输重点实验室) ; College of Physics and Information Engineering(物理与信息工程学院) ; Fuzhou University(福州市大学) ; College of Computer and Data Science(计算机与数据科学学院) ; MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition(MoE脑启发智能感知与认知重点实验室) ; University of Science and Technology of China(中国科学技术大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV
AI总结 OmniVaT通过多模态分数傅里叶适配器和离散树生成模块,首次实现单域泛化多模态视觉-触觉学习任务,提升跨领域适应性与泛化性能。