TruthLens: Visual Grounding for Universal DeepFake Reasoning
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);MLLM(abstract);分类 cs.CV、cs.AI
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);MLLM(abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV
Comments Work under review in NeurIPS 2025 with the title "Are we using Motion in Referring Segmentation? A Motion-Centric Evaluation"
专题命中 视觉定位与Grounding :MLLM(title);multimodal large language model(abstract);分类 cs.CV
机构 * College of Software Technology, Zhejiang University(浙江大学软件技术学院) ; College of Computer Science, Zhejiang University(浙江大学计算机科学学院)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV
机构 * Duke University(杜克大学)
专题命中 视觉定位与Grounding :vision language model(title,abstract);分类 cs.CV
Comments The paper has been accepted to the 2025 IEEE Conference on Virtual Reality and 3D User Interfaces (IEEE VR), and selected for publication in the 2025 IEEE Transactions on Visualization and Computer Graphics (TVCG) special issue
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
Comments Accepted by ACMMM2025
机构 * Enuma, Inc.(Enuma公司) ; Korea University(韩国大学)
专题命中 视觉定位与Grounding :multimodal large language model(abstract);MLLM(abstract)
机构 * Media Arts & Technology UC Santa Barbara(媒体艺术与技术大学圣芭芭拉分校)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments to be published in IEEE VISAP 2025
机构 * University of Glasgow(格拉斯哥大学) ; Jadavpur University(贾瓦德普尔大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments I want to revisit some of the experiments in this paper, specifically figure 5