Level Up Your Tutorials: VLMs for Game Tutorials Quality Assessment
机构 * Politecnico di Torino(托里尼理工大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted at ECCV 2024 CV2 Workshop
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * Politecnico di Torino(托里尼理工大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted at ECCV 2024 CV2 Workshop
机构 * Meituan(美团)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
机构 * The University of Tokyo, Graduate School of Arts and Sciences(东京大学艺术与科学研究生院) ; Center for Information and Neural Networks (CiNet), National Institute of Information and Communications Technology(信息与神经网络中心(CiNet),信息与通信技术国家研究所) ; The University of Osaka, Graduate School of Frontier Biosciences(大阪大学前沿生命科学研究生院) ; The University of Tokyo, Faculty of Engineering(东京大学工学部) ; The University of Osaka, Graduate School of Engineering Science(大阪大学工学研究院) ; The University of Osaka, Graduate School of Medicine(大阪大学医学研究院)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.AI
Comments 25 pages, 7 figures
机构 * Institute for AI Industry Research (AIR), Tsinghua University(人工智能产业研究所(AIR),清华大学) ; PharMolix Inc.(PharMolix公司)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
机构 * Beijing University of Posts and Telecommunications(北京邮电大学) ; China Mobile Research Institute(中国移动研究院)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
Comments Under review
机构 * Connected and Autonomous Vehicles Lab (CAV-Lab)(连接与自动驾驶车辆实验室) ; University of Surrey(萨里大学) ; Centre for Vision, Speech and Signal Processing (CVSSP)(视觉、语音和信号处理中心)
专题命中 其他VLM :vision language model(abstract);分类 cs.CV
机构 * Seoul National University(首尔国立大学)
专题命中 其他VLM :MLLM(abstract);分类 cs.AI
机构 * Department of Computer Science University of California Santa Barbara(计算机科学系加州大学圣芭芭拉分校) ; Graduate Program in Dynamical Neuroscience University of California Santa Barbara(动态神经科学研究生项目加州大学圣芭芭拉分校) ; Department of Psychological and Brain Sciences University of California Santa Barbara(心理学与脑科学系加州大学圣芭芭拉分校)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted to ACM Transactions on Graphics (SIGGRAPH 2025)
机构 * Intelligent Computing and Machine Learning Lab, School of ASEE, Beihang University(北京航空航天大学自动化学院智能计算与机器学习实验室) ; Xiaohongshu(小红书) ; School of Sino-French Engineer, Beihang University(北京航空航天大学中法工程师学院) ; College of Engineering and Computer Science, VinUniversity(Vin大学工程与计算机科学学院)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
机构 * Collage of Computer Science, Sichuan University(计算机科学学院,四川大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.LG
机构 * Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(香港科技大学电子与计算机工程系) ; Guangdong Cardiovascular Institute, Guangdong Provincial People’s Hospital (Guangdong Academy of Medical Sciences), Southern Medical University(广东省心血管病研究所,广东省人民医院(广东省医学科学院)) ; Department of Ultrasound, The First Affiliated Hospital of Guangzhou Medical University(广州市第一人民医院超声科) ; Department of Pulmonary Circulation, Shanghai Pulmonary Hospital, Tongji University School of Medicine(上海 pulmonary 医院,同济大学医学院) ; Department of Echocardiography, Fuwai Hospital Chinese Academy of Medical Sciences(阜外医院中国医学科学院) ; Guangdong Provincial Key Laboratory of South China Structural Heart Disease(广东省南方结构性心脏病重点实验室) ; State Key Laboratory of Cardiovascular Disease, Department of Echocardiography, National Center for Cardiovascular Diseases, Fuwai Hospital, Chinese Academy of Medical Sciences and Peking Union Medical College(心血管疾病国家重点实验室,国家心血管病中心,阜外医院,中国医学科学院和北京协和医学院) ; Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
机构 * School of Computer Science, University of South China(南方科技大学计算机科学学院) ; New Laboratory of Pattern Recognition, MAIS, CASIA(模式识别新实验室,MAIS,CASIA) ; School of Artificial Intelligence, UCAS(人工智能学院,UCAS) ; The Laboratory of Cognition and Decision Intelligence for Complex Systems, CASIA(复杂系统认知与决策智能实验室,CASIA) ; OPPO AI Center(OPPO AI中心)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
机构 * Grupo de Tratamiento de Imágenes (GTI), Information Processing and Telecommunications Center, ETSI Telecomunicación, Universidad Politécnica de Madrid(图像处理小组(GTI)、信息处理与电信中心、电信工程学院、马德里理工大学)
专题命中 其他VLM :visual language model(abstract);分类 cs.CV
Comments 6 pages, 3 figures,Accepted for the poster session at the CV4Animals workshop: Computer Vision for Animal Behavior Tracking and Modeling In conjunction with Computer Vision and Pattern Recognition 2024
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments v2: corrected the author list
机构 * EPFL(瑞士联邦理工学院) ; CNAM(法国国家科学与技术研究中心) ; INRIA(法国国家信息与自动化研究所) ; CIHEAM-IAMM(CIHEAM- IAMM) ; Univ. of Montpellier(蒙彼利埃大学)
专题命中 其他VLM :vision language model(abstract);分类 cs.CV
Comments Accepted at EarthVision 2025 (CVPRW 2025)
机构 * National Key Laboratory for Novel Software Technology, Nanjing University, China(新型软件技术国家实验室,南京大学,中国) ; School of Artificial Intelligence, Nanjing University, China(人工智能学院,南京大学,中国)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.AI
Comments A technical report for the MMCTR Challenge held by EReL@MIR Workshop at WWW 2025
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
Comments CVPR 2025. Project page and codes: https://eval3d.github.io/
机构 * Indian Institute of Technology Bombay(印度理工学院班加罗尔分校) ; Mohamed Bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
机构 * University of Guelph(圭尔夫大学)
专题命中 其他VLM :MLLM(abstract);分类 cs.AI
Comments 11 pages
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.AI
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments 11 pages, 4 figures, 2 tables
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Our dataset and code will be publicly available at https://github.com/SiyuanYan1/Derm1M
专题命中 其他VLM :vision language model(abstract);分类 cs.CV
Comments To appear in ETRA '25: Proceedings of the 2025 Symposium on Eye Tracking Research and Applications
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments CVPR 2025
专题命中 其他VLM :MLLM(abstract);分类 cs.CV
Comments Github page: https://github.com/linkangheng/PR1
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV