Viper-F1: Fast and Fine-Grained Multimodal Understanding with Cross-Modal State-Space Modulation
Viper-F1:基于交叉模态状态空间调制的高效细粒度多模态理解
专题命中 图文多模态 :multimodal(title,abstract);cross-modal(title);分类 cs.CV
AI总结 Viper-F1通过引入高效液态状态空间动力学和令牌-网格相关模块,实现了高效细粒度多模态理解。
Comments arXiv admin comment: This version has been removed by arXiv administrators as the submitter did not have the rights to agree to the license at the time of submission