A Training-Free, Task-Agnostic Framework for Enhancing MLLM Performance on High-Resolution Images
机构 * Yonsei University(延世大学)
Comments Accepted at CVPR 2025 Workshop on Emergent Visual Abilities and Limits of Foundation Models
期刊&会议
Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision
机构 * Yonsei University(延世大学)
Comments Accepted at CVPR 2025 Workshop on Emergent Visual Abilities and Limits of Foundation Models
机构 * SoftBank Corp. AI & Data Technology Planning Division(软银公司人工智能与数据技术规划部门)
Comments The 2nd place solution for the EgoExo4D Proficiency Estimation Challenge at the CVPR EgoVis Workshop 2025
机构 * Beihang University(北航) ; Zhongguancun Laboratory(中关村实验室) ; Hefei Comprehensive National Science Center(合肥综合性国家科学中心) ; ETH Zürich(苏黎世联邦理工学院) ; Pengcheng Laboratory(鹏城实验室) ; Tsinghua University(清华大学) ; University of California, Berkeley(加州大学伯克利分校) ; Johns Hopkins University(约翰霍普金斯大学) ; University of Oxford(牛津大学) ; Meta ; Google Brain(谷歌脑) ; Nanyang Technological University(南洋理工大学) ; Aarhus University(哥本哈根大学) ; China University of Mining(中国矿业大学) ; University of Electronic Science(电子科学大学) ; School of Electronic, Electrical(电子、电气与通信工程学院) ; Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, CAS(智能信息处理国家重点实验室) ; School of Computer Science(计算机科学学院) ; Key Laboratory of Big Data Mining(大数据挖掘国家重点实验室) ; Zhejiang Gongshang University(浙江工商大学) ; Binjiang Institute of Zhejiang University(浙江大学滨江学院) ; Zhejiang University(浙江大学) ; The Chinese University of Hong Kong(香港中文大学)
Comments AdvML@CVPR Challenge Report
Comments 10 pages, 10 figures, Mechanistic Interpretability for Vision at CVPR 2025
机构 * School of Information Science and Technology, ShanghaiTech University(信息科学与技术学院,上海科技大学) ; Sun Yat-sen University(中山大学) ; Wuhan University(武汉大学)
Comments This work is accepted by CVPR 2025
机构 * Amazon(亚马逊)
Comments Presented at the workshop Three questions about virtual try-on at CVPR 2025
机构 * Department of ECE(电子工程系) ; IPAI ; INMC ; Seoul National University(首尔国立大学)
Comments Accepted to CVPR 2025
机构 * Stanford University(斯坦福大学) ; Microsoft(微软公司) ; Georgia Institute of Technology(佐治亚理工学院)
Comments CVPR 2025 (oral). The first two authors contributed equally. Project website: https://yuegao.me/FluidNexus
机构 * School of Information Science and Technology, ShanghaiTech University(信息科学与技术学院,上海科技大学) ; Sun Yat-sen University(孙中山大学)
Comments Accepted to CVPR 2025
机构 * School of Information Science and Technology, ShanghaiTech University(信息科学与技术学院,上海科技大学) ; Wuhan University(武汉大学) ; Sun Yat-sen University(中山大学)
Comments Accepted to CVPR 2025
机构 * Cornell University(康奈尔大学) ; University of Texas at Austin(德克萨斯大学奥斯汀分校)
Comments VisCon @ CVPR 2025
机构 * Anhui Provincial Key Laboratory of Multimodal Cognitive Computation, School of Computer Science and Technology, Anhui University(安徽省多模态认知计算重点实验室,计算机科学与技术学院,安徽大学) ; Nanjing University of Aeronautics and Astronautics(南京航空航天大学) ; Nanjing University of Science and Technology(南京理工大学)
Comments IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025
机构 * Northeastern University(东北大学)
Comments CVPR 2025
机构 * KAIST(韩国科学技术院) ; Google DeepMind(谷歌DeepMind) ; Georgia Tech(佐治亚理工学院) ; Luma AI ; Google Research(谷歌研究)
Comments CVPR 2025 Workshop on AI for Content Creation (Oral)
机构 * Technical University of Munich(慕尼黑技术大学) ; Munich Center for Machine Learning(慕尼黑机器学习中心)
Comments CVPR 2025. Code availabe at https://ceveloper.github.io/publications/skeletondiffusion
机构 * National Key Laboratory for Novel Software Technology, Nanjing University, China(国家新型软件技术重点实验室,南京大学,中国) ; School of Artificial Intelligence, Nanjing University, China(人工智能学院,南京大学,中国)
Comments CVPR 2025. The code is publicly available at https://github.com/wujx2001/QwT
机构 * Technical University of Munich(慕尼黑技术大学) ; Siemens AG(西门子股份公司) ; Munich Center for Machine Learning(慕尼黑机器学习中心)
Comments Accepted by CVPR 2025
机构 * IDLab(ID实验室) ; Department of Information Technology at Ghent University(根特大学信息科技系) ; imec
Comments Accepted to the 2024 IEEE CVPR Workshop on Fair, Data-efficient, and Trusted Computer Vision. Code available at https://github.com/sdeconinck/ModelAgnosticDataAttribution
机构 * LTCI, Télécom Paris, Institut Polytechnique de Paris(LTCI,巴黎电信学院,巴黎理工学院) ; Technical University of Munich(慕尼黑技术大学) ; Helmholtz Munich(海德堡-慕尼黑研究所) ; Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML)) ; Noah’s Ark Lab, Paris(诺亚实验室,巴黎) ; Aalto University, Finland, Dept. of Electrical Engineering and Automation(阿莱克大学,芬兰,电气工程与自动化系) ; Nomagic(Nomagic公司) ; Aalto University, Finland, Department of Computer Science(阿莱克大学,芬兰,计算机科学系) ; University of Manchester, UK, Department of Computer Science(曼彻斯特大学,英国,计算机科学系)
Comments Code available at https://github.com/qbouniot/AffScoreDeep
Journal ref Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR), 2025, pp. 25250-25260
机构 * KTH Royal Institute of Technology(皇家理工学院) ; Sorbonne University(索邦大学)
Comments Accepted as a non-archival paper at the CVPR 2025 Humanoid Agents Workshop. Project page: https://groundedgestures.github.io
机构 * Dept. of Artificial Intelligence, Korea University, Seoul, Korea(人工智能系,韩国大学)
Comments CVPR 2025 (highlight)
机构 * Massachusetts Institute of Technology(麻省理工学院) ; Amazon Web Services(亚马逊网络服务)
Comments 8 pages, 5 figures, accepted to the 11th IEEE International Workshop on Computer Vision in Sports (CVSports) at CVPR 2025; supplementary appendix included
Comments An earlier version appeared in the CVPR 2025 Workshop on Generative Models for Computer Vision
机构 * University of Waterloo(滑铁卢大学)
Comments accepted in CVPR 2025, project page https://fanguw.github.io/FruitNinja3D
机构 * LG CNS AI Research(LG CNS人工智能研究)
Comments CVPR 2025 camera ready. Project page: https://lgcnsai.github.io/apt
Journal ref Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR), 2025, pp. 28619-28628
机构 * The University of Hong Kong(香港大学) ; Tsinghua University(清华大学) ; Vast ; Shanghai AI Laboratory(上海人工智能实验室)
Comments Accepted by TPAMI, extension of CVPR 2024 paper DreamComposer
机构 * HKU MMLab(香港大学多模态实验室) ; SJTU(上海交通大学) ; D-Robotics(D机器人) ; AgileX Robotics(AgileX机器人) ; Huawei Germany(华为德国分公司) ; SZU(深圳大学) ; THU(清华大学) ; NJU(南京大学) ; VIVO ; Jingdong Technology Information Technology Co., Ltd.(京东科技信息技术有限公司) ; Horizon Robotics ; SUST(上海大学) ; USST(上海理工大学) ; SYSU(南方科技大学) ; UESTC(电子科技大学) ; SCU(四川大学) ; Dexmal ; SWJTU(西南交通大学) ; HKUST(香港理工大学) ; BJTU(北京交通大学) ; BUAA(北京航空航天大学) ; NEU(东北大学) ; SHU(上海大学) ; Sangfor Technologies Inc.(深信服技术有限公司) ; CAUC(中国科学院大学) ; HUST(华中科技大学) ; Reconova Technologies Co.(Reconova技术有限公司)
Comments Challenge Webpage: https://robotwin-benchmark.github.io/cvpr-2025-challenge/
机构 * University of Waterloo(滑铁卢大学)
Comments published at CVPR 2025
Comments This paper has been accepted for publication at the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, Vancouver, 2023. Copyright IEEE
机构 * Tsinghua University(清华大学) ; Brown University(布朗大学) ; University of Liverpool(利物浦大学) ; Microsoft Research Asia(微软亚洲研究院) ; Microsoft(微软公司)
Comments Accepted by CVPR 2025. Project Page: https://bizgen-msra.github.io