arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2412.19179 2025-07-08 cs.CV cs.AI cs.LG 79%

Mask Approximation Net: A Novel Diffusion Model Approach for Remote Sensing Change Captioning

Dongwei Sun, Jing Yao, Wu Xue, Changsheng Zhou, Pedram Ghamisi, Xiangyong Cao

机构 * School of Computer Science and Technology and the Ministry of Education Key Lab for Intelligent Networks and Network Security, Xi’an Jiaotong University(计算机科学与技术学院和教育部智能网络与网络安全重点实验室,西安交通大学) Aerospace Information Research Institute, Chinese Academy of Sciences(航天信息研究所,中国科学院) Space Engineering University(航天工程大学) School of Mathematics and Statistics, Guangdong University of Technology(数学与统计学院,广东工业大学) Helmholtz-Zentrum Dresden-Rossendorf(德累斯顿-罗斯托克亥姆霍兹中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04856 2025-07-08 cs.CV 79%

Semantically Consistent Discrete Diffusion for 3D Biological Graph Modeling

Chinmay Prabhakar, Suprosanna Shit, Tamaz Amiranashvili, Hongwei Bran Li, Bjoern Menze

机构 * Department of Quantitative Biomedicine, University of Zurich, Switzerland(苏黎世大学定量生物医学系) ETH AI Center, ETH Zurich, Switzerland(苏黎世联邦理工学院AI中心) Department of Computer Science, Technical University of Munich, Germany(慕尼黑技术大学计算机科学系) Athinoula A. Martinos Center, Harvard Medical School, USA(哈佛医学院阿提尼娅·马丁斯中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04547 2025-07-08 eess.IV cs.CV 79%

FB-Diff: Fourier Basis-guided Diffusion for Temporal Interpolation of 4D Medical Imaging

Xin You, Runze Yang, Chuyan Zhang, Zhongliang Jiang, Jie Yang, Nassir Navab

机构 * Computer Aided Medical Procedures, Technical University of Munich(技术大学慕尼黑计算机辅助医学程序) Institute of Medical Robotics, Shanghai Jiao Tong University(上海交通大学医学机器人研究所) Munich Center for Machine Learning, Munich(慕尼黑机器学习中心) School of Computing, Macquarie University(麦考瑞大学计算机学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04403 2025-07-08 cs.CV 79%

Sat2City: 3D City Generation from A Single Satellite Image with Cascaded Latent Diffusion

Tongyan Hua, Lutao Jiang, Ying-Cong Chen, Wufan Zhao

机构 * HKUST(GZ)(香港科技大学(广州))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14706 2025-07-08 cs.CV 79%

Iterative Camera-LiDAR Extrinsic Optimization via Surrogate Diffusion

Ni Ou, Zhuo Chen, Xinru Zhang, Junzheng Wang

机构 * School of Automation, Beijing Institute of Technology(自动化学院,北京理工大学) Centre for Robotics Research, Department of Engineering, King’s College London(机器人研究中心,工程系,伦敦大学国王学院) School of Integrated Circuits and Electronics, Beijing Institute of Technology(集成电路与电子学院,北京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 7 pages, 4 figures, accepted by IROS 2025. arXiv admin note: substantial text overlap with arXiv:2411.10936

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06035 2025-07-08 cs.CV cs.AI 79%

HAVIR: HierArchical Vision to Image Reconstruction using CLIP-Guided Versatile Diffusion

Shiyi Zhang, Dong Liang, Hairong Zheng, Yihang Zhou

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments We have decided to withdraw this paper because the baseline methods used for comparison are outdated and do not reflect the current state-of-the-art. This significantly affects the validity of our performance claims and conclusions. We plan to conduct a more comprehensive evaluation and submit a revised version in the future

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01380 2025-07-08 cs.CV cs.AI 79%

Playing with Transformer at 30+ FPS via Next-Frame Diffusion

Xinle Cheng, Tianyu He, Jiayi Xu, Junliang Guo, Di He, Jiang Bian

机构 * Peking University(北京大学) Microsoft Research(微软研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://nextframed.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07003 2025-07-08 cs.CV 79%

CMD: Controllable Multiview Diffusion for 3D Editing and Progressive Generation

Peng Li, Suizhi Ma, Jialiang Chen, Yuan Liu, Congyi Zhang, Wei Xue, Wenhan Luo, Alla Sheffer, Wenping Wang, Yike Guo

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Johns Hopkins University(约翰霍普金斯大学) University of British Columbia(不列颠哥伦比亚大学) Texas A\&M University(德克萨斯A&M大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments SIGGRAPH 2025, Page: https://penghtyx.github.io/CMD/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11262 2025-07-08 cs.CV eess.IV 79%

Dark Noise Diffusion: Noise Synthesis for Low-Light Image Denoising

Liying Lu, Raphaël Achddou, Sabine Süsstrunk

机构 * School of Computer and Communication Sciences, École Polytechnique Fédérale de Lausanne(计算机与通信科学学院,瑞士联邦理工学院洛桑分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10936 2025-07-08 cs.CV 79%

Iterative Camera-LiDAR Extrinsic Optimization via Surrogate Diffusion

Ni Ou, Zhuo Chen, Xinru Zhang, Junzheng Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This article is an earlier version of the article arXiv:2506.14706, and most of its content is duplicated.

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.06553 2025-07-08 cs.CV 79%

HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Xiaogang Peng, Yiming Xie, Zizhao Wu, Varun Jampani, Deqing Sun, Huaizu Jiang

机构 * Northeastern University(东北大学) Hangzhou Dianzi University(杭州电子科技大学) Stability AI Google Research(谷歌研究)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://neu-vi.github.io/HOI-Diff/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03393 2025-07-08 cs.CV 79%

Masked Temporal Interpolation Diffusion for Procedure Planning in Instructional Videos

Yufan Zhou, Zhaobo Qi, Lingshuai Lin, Junqi Jing, Tingting Chai, Beichen Zhang, Shuhui Wang, Weigang Zhang

机构 * Harbin Institute of Technology(哈尔滨工业大学) Key Lab of Intell. Info. Process., Inst. of Comput. Tech., CAS(计算技术研究所智能信息处理重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03502 2025-07-08 cs.CV cs.SY eess.SY 79%

CHIME: Conditional Hallucination and Integrated Multi-scale Enhancement for Time Series Diffusion Model

Yuxuan Chen, Haipeng Xie

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15784 2025-07-08 cs.CV 79%

RL4Med-DDPO: Reinforcement Learning for Controlled Guidance Towards Diverse Medical Image Generation using Vision-Language Foundation Models

Parham Saremi, Amar Kumar, Mohamed Mohamed, Zahra TehraniNasab, Tal Arbel

机构 * Center for Intelligent Machines, McGill University, Montreal, Canada(智能机器中心,麦吉尔大学,加拿大) Mila - Quebec AI institute, Montreal, Canada(魁北克人工智能研究所,加拿大)

专题命中 扩散模型 :image generation(title);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02860 2025-07-04 cs.CV 79%

Less is Enough: Training-Free Video Diffusion Acceleration via Runtime-Adaptive Caching

Xin Zhou, Dingkang Liang, Kaijin Chen, Tianrui Feng, Xiwu Chen, Hongkai Lin, Yikang Ding, Feiyang Tan, Hengshuang Zhao, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学) MEGVII Technology(美科七科技) University of Hong Kong(香港大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments The code is made available at https://github.com/H-EmbodVis/EasyCache. Project page: https://h-embodvis.github.io/EasyCache/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02813 2025-07-04 cs.CV 79%

LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion

Fangfu Liu, Hao Li, Jiawei Chi, Hanyang Wang, Minghui Yang, Fudong Wang, Yueqi Duan

机构 * Tsinghua University(清华大学) NTU(国立科技大学) Ant Group(蚂蚁集团)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://liuff19.github.io/LangScene-X

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02687 2025-07-04 cs.CV cs.AI 79%

APT: Adaptive Personalized Training for Diffusion Models with Limited Data

JungWoo Chae, Jiyoon Kim, JaeWoong Choi, Kyungyul Kim, Sangheum Hwang

机构 * LG CNS AI Research(LG CNS人工智能研究)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025 camera ready. Project page: https://lgcnsai.github.io/apt

Journal ref Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR), 2025, pp. 28619-28628

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02405 2025-07-04 cs.CV 79%

PosDiffAE: Position-aware Diffusion Auto-encoder For High-Resolution Brain Tissue Classification Incorporating Artifact Restoration

Ayantika Das, Moitreya Chaudhuri, Koushik Bhat, Keerthi Ram, Mihail Bota, Mohanasankar Sivaprakasam

机构 * Department of Electrical Engineering, Indian Institute of Technology Madras (IITM)(电子工程系,印度理工学院Madras(IITM)) Sudha Gopalakrishnan Brain Centre (SGBC), IITM(Sudha Gopalakrishnan脑中心(SGBC),IITM)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Published in IEEE Journal of Biomedical and Health Informatics (Early Access Available) https://ieeexplore.ieee.org/document/10989734

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09993 2025-07-04 cs.CV cs.AI cs.LG 79%

Text-Aware Image Restoration with Diffusion Models

Jaewon Min, Jin Hyeon Kim, Paul Hyunbin Cho, Jaeeun Lee, Jihye Park, Minkyu Park, Sangpil Kim, Hyunhee Park, Seungryong Kim

机构 * KAIST AI(韩国科学技术院人工智能研究所) Korea University(韩国大学) Yonsei University(延世大学) Samsung Electronics(三星电子)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://cvlab-kaist.github.io/TAIR/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23115 2025-07-04 cs.CV 79%

Diffusion-Based Generative Models for 3D Occupancy Prediction in Autonomous Driving

Yunshen Wang, Yicheng Liu, Tianyuan Yuan, Yingshi Liang, Xiuyu Yang, Honggang Zhang, Hang Zhao

机构 * Institute for Interdisciplinary Information Sciences, Tsinghua University(清华大学交叉信息研究院) Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICRA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02545 2025-07-04 cs.CV 79%

MAD: Makeup All-in-One with Cross-Domain Diffusion Model

Bo-Kai Ruan, Hong-Han Shuai

机构 * National Yang Ming Chiao Tung University(阳明交通大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by CVPRW2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15248 2025-07-04 cs.CV 79%

Enhancing Fetal Plane Classification Accuracy with Data Augmentation Using Diffusion Models

Yueying Tian, Elif Ucurum, Xudong Han, Rupert Young, Chris Chatwin, Philip Birch

机构 * School of Engineering and Informatics University of Sussex(工程与信息学院 英国 Sussex 大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02299 2025-07-04 cs.CV 79%

DreamComposer++: Empowering Diffusion Models with Multi-View Conditions for 3D Content Generation

Yunhan Yang, Shuo Chen, Yukun Huang, Xiaoyang Wu, Yuan-Chen Guo, Edmund Y. Lam, Hengshuang Zhao, Tong He, Xihui Liu

机构 * The University of Hong Kong(香港大学) Tsinghua University(清华大学) Vast Shanghai AI Laboratory(上海人工智能实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by TPAMI, extension of CVPR 2024 paper DreamComposer

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02129 2025-07-04 cs.LG cs.CV 79%

Generative Latent Diffusion for Efficient Spatiotemporal Data Reduction

Xiao Li, Liangji Zhu, Anand Rangarajan, Sanjay Ranka

机构 * University of Florida(佛罗里达大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15057 2025-07-04 eess.IV cs.CV 79%

Non-rigid Motion Correction for MRI Reconstruction via Coarse-To-Fine Diffusion Models

Frederic Wang, Jonathan I. Tamir

机构 * Department of Computing and Mathematical Sciences, Caltech(计算与数学科学部,加利福尼亚理工学院) Chandra Family Department of Electrical and Computer Engineering, UT Austin(查纳德家庭电子与计算机工程系,德克萨斯大学奥斯汀分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICIP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01953 2025-07-03 cs.CV 79%

FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model

Yukang Cao, Chenyang Si, Jinghao Wang, Ziwei Liu

机构 * S-Lab, Nanyang Technological University(南洋理工大学S实验室) Nanjing University(南京大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025. Project page: https://yukangcao.github.io/FreeMorph/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01422 2025-07-03 cs.CV cs.AI 79%

DocShaDiffusion: Diffusion Model in Latent Space for Document Image Shadow Removal

Wenjie Liu, Bingshu Wang, Ze Wang, C. L. Philip Chen

机构 * School of Software, Northwestern Polytechnical University(软件学院,西北工业大学) School of Computer Science and Engineering, South China University of Technology and Pazhou Lab(计算机科学与工程学院,华南理工大学及琶洲实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01275 2025-07-03 cs.CV 79%

Frequency Domain-Based Diffusion Model for Unpaired Image Dehazing

Chengxu Liu, Lu Qi, Jinshan Pan, Xueming Qian, Ming-Hsuan Yang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04320 2025-07-03 cs.CV cs.LG 79%

ConceptAttention: Diffusion Transformers Learn Highly Interpretable Features

Alec Helbling, Tuna Han Salih Meral, Ben Hoover, Pinar Yanardag, Duen Horng Chau

机构 * ibm(IBM研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Oral Presentation at ICML 2025, Best Paper Award at CVPR Workshop on Visual Concepts

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08700 2025-07-03 eess.IV cs.CV cs.HC cs.LG 79%

Diffusion-based Iterative Counterfactual Explanations for Fetal Ultrasound Image Quality Assessment

Paraskevas Pegios, Manxi Lin, Nina Weng, Morten Bo Søndergaard Svendsen, Zahra Bashir, Siavash Bigdeli, Anders Nymark Christensen, Martin Tolsgaard, Aasa Feragen

机构 * Technical University of Denmark(技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏