3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments 34 pages, 21 figures
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments Accepted as IJCAI2023 paper
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments Accepted for publication in Elsevier Applied Soft Computing Journal, 36 pages, 18 figures
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.AI
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments Winner of CVPR2023 Long-form Video Understanding and Generation Challenge (Track 3)
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments Project Website: https://colin97.github.io/OpenShape/
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments 13 pages, 16 figures. This paper has been accepted by IEEE TMI
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments 16 pages, 9 figures, video is https://youtu.be/acy0SNLfahg, accepted in HCI International 2023
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments preprint
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments Accepted at MICCAI 2018
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments T4V Workshop
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments Accepted by CVPR 2023
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments 10 pages, 5 figures
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments CVPR 2023
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments Accepted at ICML 2023
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments This work has been accepted to the Conference on Robot Learning (CoRL) 2022
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV
Comments Accepted at CVPR 2023, project page at https://lattas.github.io/fitme , 17 pages including supplementary material
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI
专题命中 多模态训练与对齐 :image-text(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CL
Comments Accepted at The 17th International Conference on Document Analysis and Recognition (ICDAR 2023)
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI
Journal ref 32nd International Joint Conference on Artificial Intelligence, IJCAI 2023
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.AI
Journal ref Nat Commun 13, 3293 (2022)
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CL