Show or Tell? Effectively prompting Vision-Language Models for semantic segmentation
展示还是讲述?有效提示视觉-语言模型进行语义分割
机构 * IBM Research(IBM研究院) ; ETH Zurich(苏黎世联邦理工学院)
AI总结 本文提出PromptMatcher,通过结合文本和视觉提示提升VLMs在语义分割中的性能,实现优于现有方法的改进。
Journal ref Transactions on Machine Learning Research, 2025