arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2507.17240 2025-07-24 cs.CV

Perceptual Classifiers: Detecting Generative Images using Perceptual Features

Krishna Srikar Durbha, Asvin Kumar Venkataramanan, Rajesh Sureddi, Alan C. Bovik

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校) University of Colorado Boulder(科罗拉多大学波德分校)

Comments 8 pages, 6 figures, 3 tables, ICCV VQualA Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17221 2025-07-24 cs.LG cs.CV

Dataset Distillation as Data Compression: A Rate-Utility Perspective

Youneng Bao, Yiping Liu, Zhuo Chen, Yongsheng Liang, Mu Li, Kede Ma

机构 * Harbin Institute of Technology(哈尔滨工业大学) Peng Cheng Laboratory(鹏城实验室) City University of Hong Kong(香港城市大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17050 2025-07-24 cs.CV

Toward Scalable Video Narration: A Training-free Approach Using Multimodal Large Language Models

Tz-Ying Wu, Tahani Trigui, Sharath Nittur Sridhar, Anand Bodas, Subarna Tripathi

机构 * Intel(英特尔)

Comments Accepted to CVAM Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16946 2025-07-24 cs.CV

Toward Long-Tailed Online Anomaly Detection through Class-Agnostic Concepts

Chiao-An Yang, Kuan-Chuan Peng, Raymond A. Yeh

机构 * Department of Computer Science, Purdue University(普渡大学计算机科学系) Mitsubishi Electric Research Laboratories(三菱电机研究实验室)

Comments This paper is accepted to ICCV 2025. The supplementary material is included. The long-tailed online anomaly detection dataset is available at https://doi.org/10.5281/zenodo.16283852

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02591 2025-07-24 cs.CV

AuroraLong: Bringing RNNs Back to Efficient Open-Ended Video Understanding

Weili Xu, Enxin Song, Wenhao Chai, Xuexiang Wen, Tian Ye, Gaoang Wang

机构 * Zhejiang University(浙江大学) University of Washington(华盛顿大学) HKUST (GZ)(香港科技大学(广州)) Zhejiang University / Shanghai AI Lab(浙江大学/上海人工智能实验室)

Comments ICCV 2025 Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07986 2025-07-24 cs.CV

Rethinking Cross-Modal Interaction in Multimodal Diffusion Transformers

Zhengyao Lv, Tianlin Pan, Chenyang Si, Zhaoxi Chen, Wangmeng Zuo, Ziwei Liu, Kwan-Yee K. Wong

机构 * The University of Hong Kong(香港大学) Nanjing University(南京大学) University of Chinese Academy of Sciences(中国科学院大学) Nanyang Technological University(南洋理工大学) Harbin Institute of Technology(哈尔滨工业大学)

Comments Accepted by ICCV 2025; Project Page: https://vchitect.github.io/TACA/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16790 2025-07-23 cs.CV

Enhancing Domain Diversity in Synthetic Data Face Recognition with Dataset Fusion

Anjith George, Sebastien Marcel

机构 * Idiap Research Institute(Idiap研究机构)

Comments Accepted in ICCV Workshops 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16782 2025-07-23 cs.CV

Task-Specific Zero-shot Quantization-Aware Training for Object Detection

Changhao Li, Xinrui Chen, Ji Wang, Kang Zhao, Jianfei Chen

机构 * School of Computational Science and Engineering, Georgia Institute of Technology(计算科学与工程学院,佐治亚理工学院) Shenzhen International Graduate School, Tsinghua University(深圳国际研究生院,清华大学) School of Software, Tsinghua University(软件学院,清华大学) Dept. of Comp. Sci. and Tech., Institute for AI, Tsinghua-Bosch Joint ML Center, Tsinghua University(计算机科学与技术系,人工智能研究院,清华大学-博世联合机器学习中心,清华大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16389 2025-07-23 cs.CV cs.AI

From Flat to Round: Redefining Brain Decoding with Surface-Based fMRI and Cortex Structure

Sijin Yu, Zijiao Chen, Wenxuan Wu, Shengxian Chen, Zhongliang Liu, Jingxin Nie, Xiaofen Xing, Xiangmin Xu, Xin Zhang

机构 * South China University of Technology(华南理工大学) Stanford University(斯坦福大学) South China Normal University(华南师范大学) Pazhou Lab(琶洲实验室)

Comments 18 pages, 14 figures, ICCV Findings 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16213 2025-07-23 cs.CV cs.AI

Advancing Visual Large Language Model for Multi-granular Versatile Perception

Wentao Xiang, Haoxian Tan, Cong Wei, Yujie Zhong, Dengjie Li, Yujiu Yang

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学) Meituan Inc.(美团公司)

Comments To appear in ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16206 2025-07-23 cs.LG cs.AI

METER: Multi-modal Evidence-based Thinking and Explainable Reasoning -- Algorithm and Benchmark

Xu Yang, Qi Zhang, Shuming Jiang, Yaowen Xu, Zhaofan Zou, Hao Sun, Xuelong Li

机构 * Institute of Artificial Intelligence (TeleAI), China Telecom(电信人工智能研究院) Institute of Artificial Intelligence and Robotics(IAIR), Xi’an Jiaotong University(人工智能与机器人研究院) Advanced Technique of Artificial Intelligence(ATAI), Chongqing University of Technology(人工智能先进技术研究院)

Comments 9 pages,3 figures ICCV format

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16154 2025-07-23 cs.CV cs.AI

LSSGen: Leveraging Latent Space Scaling in Flow and Diffusion for Efficient Text to Image Generation

Jyun-Ze Tang, Chih-Fan Hsu, Jeng-Lin Li, Ming-Ching Chang, Wei-Chao Chen

机构 * Inventec Corporation(英威达公司) University at Albany, State University of New York(纽约州立大学阿尔巴尼分校)

Comments ICCV AIGENS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05056 2025-07-23 cs.CV cs.AI

INTER: Mitigating Hallucination in Large Vision-Language Models by Interaction Guidance Sampling

Xin Dong, Shichao Dong, Jin Wang, Jing Huang, Li Zhou, Zenghui Sun, Lihua Jing, Jingsong Lan, Xiaoyong Zhu, Bo Zheng

机构 * University of Chinese Academy of Sciences(中国科学院大学) The University of Hong Kong(香港大学) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01756 2025-07-23 cs.CV

Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis

Peng Zheng, Junke Wang, Yi Chang, Yizhou Yu, Rui Ma, Zuxuan Wu

机构 * School of Artificial Intelligence, Jilin University(吉林大学人工智能学院) Shanghai Innovation Institute(上海创新研究院) Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究院) Department of Computer Science, The University of Hong Kong(香港大学计算机科学系)

Comments iccv 2025, camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23468 2025-07-23 cs.CV

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments

Xuan Yao, Junyu Gao, Changsheng Xu

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences (CASIA)(多模态人工智能系统国家重点实验室(MAIS)、自动化研究所、中国科学院(CASIA)) School of Artificial Intelligence, University of Chinese Academy of Sciences (UCAS)(人工智能学院、中国科学院大学(UCAS)) Peng Cheng Laboratory, ShenZhen, China(鹏城实验室、深圳中国)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22139 2025-07-23 cs.CV

Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs

Shaojie Zhang, Jiahui Yang, Jianqin Yin, Zhenbo Luo, Jian Luan

机构 * MiLM Plus, Xiaomi Inc.(小米公司)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21401 2025-07-23 cs.CV

Curve-Aware Gaussian Splatting for 3D Parametric Curve Reconstruction

Zhirui Gao, Renjiao Yi, Yaqiao Dai, Xuening Zhu, Wei Chen, Chenyang Zhu, Kai Xu

机构 * National University of Defense Technology(国防科技大学)

Comments Accepted by ICCV 2025, Code: https://github.com/zhirui-gao/Curve-Gaussian

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17262 2025-07-23 cs.CV cs.LG eess.IV

Unsupervised Joint Learning of Optical Flow and Intensity with Event Cameras

Shuang Guo, Friedhelm Hamann, Guillermo Gallego

机构 * TU Berlin and Robotics Institute Germany(柏林技术大学和机器人研究所) Science of Intelligence Excellence Cluster(智能科学卓越中心) Einstein Center for Digital Future(爱因斯坦数字未来研究中心)

Comments 13 pages, 8 figures, 9 tables. Project page: https://github.com/tub-rip/E2FAI . IEEE/CVF International Conference on Computer Vision (ICCV), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07601 2025-07-23 cs.CV cs.LG

Balanced Image Stylization with Style Matching Score

Yuxin Jiang, Liming Jiang, Shuai Yang, Jia-Wei Liu, Ivor Tsang, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学Show实验室) Agency for Science, Technology and Research (A*STAR)(科技研究局) Nanyang Technological University(南洋理工大学) Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)

Comments ICCV 2025. Code: https://github.com/showlab/SMS Project page: https://yuxinn-j.github.io/projects/SMS.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05182 2025-07-23 cs.CV

MGSR: 2D/3D Mutual-boosted Gaussian Splatting for High-fidelity Surface Reconstruction under Various Light Conditions

Qingyuan Zhou, Yuehu Gong, Weidong Yang, Jiaze Li, Yeqi Luo, Baixin Xu, Shuhao Li, Ben Fei, Ying He

机构 * Fudan University(复旦大学) Nanyang Technological University(南洋理工大学) The Chinese University of Hong Kong(香港中文大学)

Comments Accepted at ICCV'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.02358 2025-07-23 cs.CV

MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm

Ziyan Guo, Zeyu Hu, De Wen Soh, Na Zhao

机构 * Singapore University of Technology and Design(新加坡科技设计大学) LIGHTSPEED

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08629 2025-07-23 cs.CV cs.LG

FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models

Vladimir Kulikov, Matan Kleiner, Inbar Huberman-Spiegelglas, Tomer Michaeli

机构 * Technion – Israel Institute of Technology(技术学院 – 以色列理工学院)

Comments ICCV 2025. Project's webpage at https://matankleiner.github.io/flowedit/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13613 2025-07-23 cs.CV cs.GR

MEGA: Memory-Efficient 4D Gaussian Splatting for Dynamic Scenes

Xinjie Zhang, Zhening Liu, Yifan Zhang, Xingtong Ge, Dailan He, Tongda Xu, Yan Wang, Zehong Lin, Shuicheng Yan, Jun Zhang

机构 * iComAI Lab, The Hong Kong University of Science and Technology(香港科技大学iComAI实验室) Skywork AI The Chinese University of Hong Kong(香港中文大学) National University of Singapore(新加坡国立大学) Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12332 2025-07-23 cs.CV

MC-Bench: A Benchmark for Multi-Context Visual Grounding in the Era of MLLMs

Yunqiu Xu, Linchao Zhu, Yi Yang

机构 * ReLER Lab, CCAI Zhejiang University(ReLER实验室,中国浙江大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16015 2025-07-23 cs.CV

Is Tracking really more challenging in First Person Egocentric Vision?

Matteo Dunnhofer, Zaira Manigrasso, Christian Micheloni

机构 * University of Udine(乌迪内大学) York University(约克大学)

Comments 2025 IEEE/CVF International Conference on Computer Vision (ICCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04123 2025-07-23 cs.CV cs.AI

Towards Accurate and Efficient 3D Object Detection for Autonomous Driving: A Mixture of Experts Computing System on Edge

Linshen Liu, Boyan Su, Junyue Jiang, Guanlin Wu, Cong Guo, Ceyu Xu, Hao Frank Yang

机构 * Johns Hopkins University(约翰霍普金斯大学) Duke University(杜克大学) HKUST(香港科技大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13684 2025-07-23 cs.CV

FiVE: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models

Minghan Li, Chenxi Xie, Yichen Wu, Lei Zhang, Mengyu Wang

机构 * Harvard AI and Robotics Lab, Harvard University(哈佛人工智能与机器人实验室,哈佛大学) Broad Institute(博德研究所) Hong Kong Polytechnic University(香港理工大学) School of Engineering and Applied Sciences, Harvard University(哈佛大学工程与应用科学学院) City University of Hong Kong(香港城市大学) Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(自然与人工智能研究学院,哈佛大学)

Comments 24 pages, 14 figures, 16 tables

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15803 2025-07-22 cs.CV cs.AI cs.LG

ConformalSAM: Unlocking the Potential of Foundational Segmentation Models in Semi-Supervised Semantic Segmentation with Conformal Prediction

Danhui Chen, Ziquan Liu, Chuxi Yang, Dan Wang, Yan Yan, Yi Xu, Xiangyang Ji

机构 * Dalian University of Technology(大连理工大学) Queen Mary University of London(伦敦玛丽女王大学) Washington State University(华盛顿州立大学) Tsinghua University(清华大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15775 2025-07-22 gr-qc astro-ph.IM cs.AI

Learning Null Geodesics for Gravitational Lensing Rendering in General Relativity

Mingyuan Sun, Zheng Fang, Jiaxu Wang, Kunyi Zhang, Qiang Zhang, Renjing Xu

机构 * Northeastern University(东北大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Sichuan University(四川大学) Beijing Innovation Center of Humanoid Robotics Co., Ltd.(北京人形机器人创新中心有限公司)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15686 2025-07-22 cs.CV cs.AI

LINR-PCGC: Lossless Implicit Neural Representations for Point Cloud Geometry Compression

Wenjie Huang, Qi Yang, Shuting Xia, He Huang, Zhu Li, Yiling Xu

机构 * Shanghai Jiao Tong University(上海交通大学) University of Missouri-Kansas City(密苏里大学-凯斯城分校)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏