arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2025-07-23 至 2025-07-23 共收录 21
2507.16790 2025-07-23 cs.CV

Enhancing Domain Diversity in Synthetic Data Face Recognition with Dataset Fusion

Anjith George, Sebastien Marcel

机构 * Idiap Research Institute(Idiap研究机构)

Comments Accepted in ICCV Workshops 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16782 2025-07-23 cs.CV

Task-Specific Zero-shot Quantization-Aware Training for Object Detection

Changhao Li, Xinrui Chen, Ji Wang, Kang Zhao, Jianfei Chen

机构 * School of Computational Science and Engineering, Georgia Institute of Technology(计算科学与工程学院,佐治亚理工学院) Shenzhen International Graduate School, Tsinghua University(深圳国际研究生院,清华大学) School of Software, Tsinghua University(软件学院,清华大学) Dept. of Comp. Sci. and Tech., Institute for AI, Tsinghua-Bosch Joint ML Center, Tsinghua University(计算机科学与技术系,人工智能研究院,清华大学-博世联合机器学习中心,清华大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16389 2025-07-23 cs.CV cs.AI

From Flat to Round: Redefining Brain Decoding with Surface-Based fMRI and Cortex Structure

Sijin Yu, Zijiao Chen, Wenxuan Wu, Shengxian Chen, Zhongliang Liu, Jingxin Nie, Xiaofen Xing, Xiangmin Xu, Xin Zhang

机构 * South China University of Technology(华南理工大学) Stanford University(斯坦福大学) South China Normal University(华南师范大学) Pazhou Lab(琶洲实验室)

Comments 18 pages, 14 figures, ICCV Findings 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16213 2025-07-23 cs.CV cs.AI

Advancing Visual Large Language Model for Multi-granular Versatile Perception

Wentao Xiang, Haoxian Tan, Cong Wei, Yujie Zhong, Dengjie Li, Yujiu Yang

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学) Meituan Inc.(美团公司)

Comments To appear in ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16206 2025-07-23 cs.LG cs.AI

METER: Multi-modal Evidence-based Thinking and Explainable Reasoning -- Algorithm and Benchmark

Xu Yang, Qi Zhang, Shuming Jiang, Yaowen Xu, Zhaofan Zou, Hao Sun, Xuelong Li

机构 * Institute of Artificial Intelligence (TeleAI), China Telecom(电信人工智能研究院) Institute of Artificial Intelligence and Robotics(IAIR), Xi’an Jiaotong University(人工智能与机器人研究院) Advanced Technique of Artificial Intelligence(ATAI), Chongqing University of Technology(人工智能先进技术研究院)

Comments 9 pages,3 figures ICCV format

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16154 2025-07-23 cs.CV cs.AI

LSSGen: Leveraging Latent Space Scaling in Flow and Diffusion for Efficient Text to Image Generation

Jyun-Ze Tang, Chih-Fan Hsu, Jeng-Lin Li, Ming-Ching Chang, Wei-Chao Chen

机构 * Inventec Corporation(英威达公司) University at Albany, State University of New York(纽约州立大学阿尔巴尼分校)

Comments ICCV AIGENS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05056 2025-07-23 cs.CV cs.AI

INTER: Mitigating Hallucination in Large Vision-Language Models by Interaction Guidance Sampling

Xin Dong, Shichao Dong, Jin Wang, Jing Huang, Li Zhou, Zenghui Sun, Lihua Jing, Jingsong Lan, Xiaoyong Zhu, Bo Zheng

机构 * University of Chinese Academy of Sciences(中国科学院大学) The University of Hong Kong(香港大学) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01756 2025-07-23 cs.CV

Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis

Peng Zheng, Junke Wang, Yi Chang, Yizhou Yu, Rui Ma, Zuxuan Wu

机构 * School of Artificial Intelligence, Jilin University(吉林大学人工智能学院) Shanghai Innovation Institute(上海创新研究院) Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究院) Department of Computer Science, The University of Hong Kong(香港大学计算机科学系)

Comments iccv 2025, camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23468 2025-07-23 cs.CV

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments

Xuan Yao, Junyu Gao, Changsheng Xu

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences (CASIA)(多模态人工智能系统国家重点实验室(MAIS)、自动化研究所、中国科学院(CASIA)) School of Artificial Intelligence, University of Chinese Academy of Sciences (UCAS)(人工智能学院、中国科学院大学(UCAS)) Peng Cheng Laboratory, ShenZhen, China(鹏城实验室、深圳中国)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22139 2025-07-23 cs.CV

Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs

Shaojie Zhang, Jiahui Yang, Jianqin Yin, Zhenbo Luo, Jian Luan

机构 * MiLM Plus, Xiaomi Inc.(小米公司)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21401 2025-07-23 cs.CV

Curve-Aware Gaussian Splatting for 3D Parametric Curve Reconstruction

Zhirui Gao, Renjiao Yi, Yaqiao Dai, Xuening Zhu, Wei Chen, Chenyang Zhu, Kai Xu

机构 * National University of Defense Technology(国防科技大学)

Comments Accepted by ICCV 2025, Code: https://github.com/zhirui-gao/Curve-Gaussian

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17262 2025-07-23 cs.CV cs.LG eess.IV

Unsupervised Joint Learning of Optical Flow and Intensity with Event Cameras

Shuang Guo, Friedhelm Hamann, Guillermo Gallego

机构 * TU Berlin and Robotics Institute Germany(柏林技术大学和机器人研究所) Science of Intelligence Excellence Cluster(智能科学卓越中心) Einstein Center for Digital Future(爱因斯坦数字未来研究中心)

Comments 13 pages, 8 figures, 9 tables. Project page: https://github.com/tub-rip/E2FAI . IEEE/CVF International Conference on Computer Vision (ICCV), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07601 2025-07-23 cs.CV cs.LG

Balanced Image Stylization with Style Matching Score

Yuxin Jiang, Liming Jiang, Shuai Yang, Jia-Wei Liu, Ivor Tsang, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学Show实验室) Agency for Science, Technology and Research (A*STAR)(科技研究局) Nanyang Technological University(南洋理工大学) Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)

Comments ICCV 2025. Code: https://github.com/showlab/SMS Project page: https://yuxinn-j.github.io/projects/SMS.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05182 2025-07-23 cs.CV

MGSR: 2D/3D Mutual-boosted Gaussian Splatting for High-fidelity Surface Reconstruction under Various Light Conditions

Qingyuan Zhou, Yuehu Gong, Weidong Yang, Jiaze Li, Yeqi Luo, Baixin Xu, Shuhao Li, Ben Fei, Ying He

机构 * Fudan University(复旦大学) Nanyang Technological University(南洋理工大学) The Chinese University of Hong Kong(香港中文大学)

Comments Accepted at ICCV'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.02358 2025-07-23 cs.CV

MotionLab: Unified Human Motion Generation and Editing via the Motion-Condition-Motion Paradigm

Ziyan Guo, Zeyu Hu, De Wen Soh, Na Zhao

机构 * Singapore University of Technology and Design(新加坡科技设计大学) LIGHTSPEED

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08629 2025-07-23 cs.CV cs.LG

FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models

Vladimir Kulikov, Matan Kleiner, Inbar Huberman-Spiegelglas, Tomer Michaeli

机构 * Technion – Israel Institute of Technology(技术学院 – 以色列理工学院)

Comments ICCV 2025. Project's webpage at https://matankleiner.github.io/flowedit/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13613 2025-07-23 cs.CV cs.GR

MEGA: Memory-Efficient 4D Gaussian Splatting for Dynamic Scenes

Xinjie Zhang, Zhening Liu, Yifan Zhang, Xingtong Ge, Dailan He, Tongda Xu, Yan Wang, Zehong Lin, Shuicheng Yan, Jun Zhang

机构 * iComAI Lab, The Hong Kong University of Science and Technology(香港科技大学iComAI实验室) Skywork AI The Chinese University of Hong Kong(香港中文大学) National University of Singapore(新加坡国立大学) Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12332 2025-07-23 cs.CV

MC-Bench: A Benchmark for Multi-Context Visual Grounding in the Era of MLLMs

Yunqiu Xu, Linchao Zhu, Yi Yang

机构 * ReLER Lab, CCAI Zhejiang University(ReLER实验室,中国浙江大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16015 2025-07-23 cs.CV

Is Tracking really more challenging in First Person Egocentric Vision?

Matteo Dunnhofer, Zaira Manigrasso, Christian Micheloni

机构 * University of Udine(乌迪内大学) York University(约克大学)

Comments 2025 IEEE/CVF International Conference on Computer Vision (ICCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04123 2025-07-23 cs.CV cs.AI

Towards Accurate and Efficient 3D Object Detection for Autonomous Driving: A Mixture of Experts Computing System on Edge

Linshen Liu, Boyan Su, Junyue Jiang, Guanlin Wu, Cong Guo, Ceyu Xu, Hao Frank Yang

机构 * Johns Hopkins University(约翰霍普金斯大学) Duke University(杜克大学) HKUST(香港科技大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13684 2025-07-23 cs.CV

FiVE: A Fine-grained Video Editing Benchmark for Evaluating Emerging Diffusion and Rectified Flow Models

Minghan Li, Chenxi Xie, Yichen Wu, Lei Zhang, Mengyu Wang

机构 * Harvard AI and Robotics Lab, Harvard University(哈佛人工智能与机器人实验室,哈佛大学) Broad Institute(博德研究所) Hong Kong Polytechnic University(香港理工大学) School of Engineering and Applied Sciences, Harvard University(哈佛大学工程与应用科学学院) City University of Hong Kong(香港城市大学) Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(自然与人工智能研究学院,哈佛大学)

Comments 24 pages, 14 figures, 16 tables

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏