Steering and Rectifying Latent Representation Manifolds in Frozen Multi-modal LLMs for Video Anomaly Detection
在冻结的多模态大语言模型中引导和校正潜在表示流形以进行视频异常检测
专题命中 视频多模态 :multi-modal(title,abstract);MLLM(abstract);分类 cs.CV
AI总结 SteerVAD通过引导和校正冻结多模态大语言模型的潜在表示流形,提升视频异常检测的性能,仅需1%训练数据即达最优效果。
Comments Accepted by ICLR 2026