AdCare-VLM: Towards a Unified and Pre-aligned Latent Representation for Healthcare Video Understanding
机构 * University of Georgia(佐治亚大学)
专题命中 视觉问答 :VLM(title,abstract);vision language model(abstract);LLaVA(abstract);visual question answering(abstract)
Comments 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: 7th International Workshop on Large Scale Holistic Video Understanding: Toward Video Foundation Models
Journal ref Neural Information Processing Systems (NeurIPS 2025)