Guarding the Meaning: Self-Supervised Training for Semantic Robustness in Guard Models
机构 * Qualcomm AI Research(高通人工智能研究)
专题命中 幻觉与鲁棒性 :grounding(abstract);分类 cs.AI、cs.LG
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * Qualcomm AI Research(高通人工智能研究)
专题命中 幻觉与鲁棒性 :grounding(abstract);分类 cs.AI、cs.LG
专题命中 幻觉与鲁棒性 :visual question answering(abstract);分类 cs.CV
机构 * Institute of Trustworthy Embodied AI(可信具身人工智能研究院) ; Fudan University(复旦大学) ; Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心) ; Columbia University(哥伦比亚大学)
专题命中 幻觉与鲁棒性 :vision-language model(abstract);分类 cs.CV
机构 * University of Padua(帕多瓦大学)
专题命中 VLM训练与架构 :MLLM(title,abstract);LLaVA(abstract);multimodal large language model(abstract);分类 cs.CV
专题命中 VLM训练与架构 :grounding(abstract);multimodal large language model(abstract);分类 cs.CV、cs.AI
专题命中 VLM训练与架构 :vision-language model(abstract);VLM(abstract);分类 cs.CV、cs.AI
Comments Accepted by AAAI 2026
专题命中 VLM训练与架构 :vision-language model(abstract);VLM(abstract);分类 cs.LG
Comments 11 pages, 3 figures
机构 * College of Intelligent Engineering and Automation, Beijing University of Posts and Telecommunications(智能工程与自动化学院,北京邮电大学)
专题命中 VLM训练与架构 :multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
机构 * Idiap Research Institute(日内瓦研究所)
专题命中 VLM训练与架构 :vision-language model(abstract);分类 cs.CV
专题命中 VLM训练与架构 :vision-language model(abstract);分类 cs.CV
专题命中 VLM训练与架构 :grounding(abstract)
Comments 9 pages, 3 figures,submitted to AAMAS 2026
机构 * Beijing Institute of Technology (BIT)(北京理工大学) ; School of Computer Science, University of Nottingham(诺丁汉大学计算机学院)
专题命中 VLM训练与架构 :VLM(abstract)
机构 * Department of Computer Science and Electrical Engineering University of Maryland Baltimore County(计算机科学与电气工程系马里兰大学巴尔的摩县)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted in IJCNLP-AACL 2025 (also presented in MAGMAR 2025 at ACL 2025)