Typographic Attacks in a Multi-Image Setting
专题命中 其他VLM :vision-language model(abstract)
Comments Accepted by NAACL2025. Our code is available at https://github.com/XiaomengWang-AI/Typographic-Attacks-in-a-Multi-Image-Setting
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 其他VLM :vision-language model(abstract)
Comments Accepted by NAACL2025. Our code is available at https://github.com/XiaomengWang-AI/Typographic-Attacks-in-a-Multi-Image-Setting
专题命中 其他VLM :vision-language model(abstract)
Comments Accepted to NAACL 2025
专题命中 其他VLM :multimodal large language model(abstract)
Comments Accepted to HRI 2025
专题命中 其他VLM :vision-language model(abstract)
Comments This work was presented at the WARN, Weighing the Benefits of Autonomous Robot Personalization, workshop at the 33rd IEEE RO-MAN 2024 conference
专题命中 其他VLM :multimodal large language model(abstract)
Comments preprint
专题命中 其他VLM :multimodal large language model(abstract)
Comments 11 pages, 2 figures, 2 tables
专题命中 其他VLM :vision-language model(abstract)
专题命中 其他VLM :MLLM(abstract)
Comments 22 pages, 10 figures
专题命中 其他VLM :vision language model(abstract)
Comments Accepted at IROS 2024
专题命中 其他VLM :multimodal large language model(abstract)
Comments Accepted to EMNLP 2024 Findings
专题命中 其他VLM :visual language model(abstract)
Comments CIKM 2024 Workshop on Industrial Recommendation Systems
专题命中 其他VLM :vision-language model(abstract)
Comments 5 pages
专题命中 其他VLM :vision-language model(abstract)
Comments 6 pages
专题命中 其他VLM :vision-language model(abstract)
专题命中 其他VLM :MLLM(abstract)
专题命中 其他VLM :vision-language model(abstract)
专题命中 其他VLM :MLLM(abstract)
专题命中 其他VLM :multimodal large language model(abstract)
Comments Short paper accepted by AGILE 2024 conference (https://agile-gi.eu/conference-2024)
专题命中 其他VLM :vision-language model(abstract)
专题命中 其他VLM :vision language model(abstract)
Comments 13 pages, 5 figures
专题命中 其他VLM :multimodal large language model(abstract)
Comments Work in Progress
专题命中 其他VLM :vision-language model(abstract)
Comments For accessibility tagged pdf, please refer to the ancillary file
专题命中 其他VLM :vision language model(abstract)