AnySynth: Harnessing the Power of Image Synthetic Data Generation for Generalized Vision-Language Tasks
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Project Page: https://hara012.github.io/MaGRITTe-project
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments project page: https://sketch-agent.csail.mit.edu/
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.CL
Comments Accepted to WACV 2025
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments Github repo: https://github.com/jayLEE0301/vq_bet_official
Journal ref PMLR 235:26991-27008, 2024
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments Accepted by EMNLP 2024 (Oral Presentation); Project Page: https://haoningwu3639.github.io/MatchTime/
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments arXiv admin note: substantial text overlap with arXiv:2408.13335
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Accepted by 2024 5th International Conference on Information Science, Parallel and Distributed Systems
Journal ref Proceedings of the 2024 5th International Conference on Information Science, Parallel and Distributed Systems (ISPDS), 2024, pp. 77-81
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
Comments 34 pages
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Accepted to ECCV 2024
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments Presented at the CHI 2024 Workshop "Building a Metaverse for All: Opportunities and Challenges for Future Inclusive and Accessible Virtual Environments", May 11, 2024, Honolulu, Hawaii
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments NeurIPS 2024
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
Comments Accepted at the GenLaw (Generative AI + Law) workshop at ICML'24
专题命中 多模态生成 :MLLM(abstract);分类 cs.CV
Comments NeurIPS 2024, Code Available: https://github.com/lingxiao-li/Bifrost
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments Under review
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments 2024 16th IIAI International Congress on Advanced Applied Informatics (IIAI-AAI)
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
Comments This article has been accepted for publication in a future issue of IEEE Transactions on Medical Imaging (TMI), but has not been fully edited. Content may change prior to final publication. Citation information: DOI: https://doi.org/10.1109/TMI.2024.3473745 . Code: https://github.com/wuyongjianCODE/AttriPrompter
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments Project Page: https://malteprinzler.github.io/projects/joker/
专题命中 多模态生成 :MLLM(abstract);分类 cs.CV
Comments 14 pages