Training-Free Label Space Alignment for Universal Domain Adaptation
机构 * Department of Artificial Intelligence, Korea University(人工智能系,韩国大学)
专题命中 VLM训练与架构 :vision-language model(abstract);分类 cs.CV、cs.AI
Comments 22 pages, 12 figures
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * Department of Artificial Intelligence, Korea University(人工智能系,韩国大学)
专题命中 VLM训练与架构 :vision-language model(abstract);分类 cs.CV、cs.AI
Comments 22 pages, 12 figures
机构 * Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) ; Guangdong Provincial Key Laboratory of Novel Security Intelligence Technologies(广东省新型安全智能技术重点实验室)
专题命中 其他VLM :multimodal large language model(abstract);MLLM(abstract);分类 cs.AI、cs.LG
Comments EMNLP 2025 Main Conference
机构 * Beijing Key Laboratory of Intelligent Information Technology, School of Computer Science & Technology, Beijing Institute of Technology, China(北京智能信息科技重点实验室,计算机科学与技术学院,北京理工大学,中国) ; Guangdong Provincial Key Laboratory of Machine Perception and Intelligent Computing, Shenzhen MSU-BIT University, Shenzhen, China(广东省机器感知与智能计算重点实验室,深圳MSU-BIT大学,深圳,中国) ; NVIDIA
专题命中 其他VLM :vision language model(abstract);分类 cs.CV
Comments NeurIPS 2025, Project: https://github.com/YaoChengTang/3D-Visual-Illusion-Depth-Estimation