arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2505.11800 2025-05-20 cs.CV eess.IV 79%

Self-Learning Hyperspectral and Multispectral Image Fusion via Adaptive Residual Guided Subspace Diffusion Model

Jian Zhu, He Wang, Yang Xu, Zebin Wu, Zhihui Wei

机构 * School of Computer Science and Engineering, Nanjing University of Science and Technology, Nanjing, China(计算机科学与工程学院,南京理工大学,南京,中国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments cvpr

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07775 2025-05-20 cs.LG cs.CV 79%

Efficient Diversity-Preserving Diffusion Alignment via Gradient-Informed GFlowNets

Zhen Liu, Tim Z. Xiao, Weiyang Liu, Yoshua Bengio, Dinghuai Zhang

机构 * Mila, Université de Montréal(蒙特利尔大学Mila实验室) Max Planck Institute for Intelligent Systems - Tübingen(智能系统马克斯·普朗克研究所(图宾根)) The Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)) University of Tübingen(图宾根大学) University of Cambridge(剑桥大学) Microsoft Research(微软研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Technical Report (36 pages, 31 figures), Accepted at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.21759 2025-05-20 cs.CV 79%

IntLoRA: Integral Low-rank Adaptation of Quantized Diffusion Models

Hang Guo, Yawei Li, Tao Dai, Shu-Tao Xia, Luca Benini

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10733 2025-05-20 cs.CV cs.AI 79%

Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

Junyu Chen, Han Cai, Junsong Chen, Enze Xie, Shang Yang, Haotian Tang, Muyang Li, Yao Lu, Song Han

机构 * MIT(麻省理工学院) Tsinghua University(清华大学) NVIDIA(英伟达)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICLR 2025. The first two authors contributed equally to this work. Fix Typo

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08151 2025-05-20 cs.CV cs.LG 79%

Progressive Autoregressive Video Diffusion Models

Desai Xie, Zhan Xu, Yicong Hong, Hao Tan, Difan Liu, Feng Liu, Arie Kaufman, Yang Zhou

机构 * Stony Brook University(石溪大学) Adobe Research(Adobe研究)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 7 figures. Code and video results are available at https://desaixie.github.io/pa-vdm/. v2: Accepted to CVPRW 2025. Updated figures, tables, notations, and text in all sections. Added comparison with more baseline methods, FVD metric results, user study, and discussion on parallel works

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.08730 2025-05-20 cs.CV 79%

DiffusionAD: Norm-guided One-step Denoising Diffusion for Anomaly Detection

Hui Zhang, Zheng Wang, Dan Zeng, Zuxuan Wu, Yu-Gang Jiang

机构 * Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究院) School of Computer Science, Zhejiang University of Technology(浙江工业大学计算机学院) School of Communication & Information Engineering, Shanghai University(上海大学通信与信息工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11278 2025-05-19 stat.ML cs.CV cs.LG stat.ME 79%

A Fourier Space Perspective on Diffusion Models

Fabian Falck, Teodora Pandeva, Kiarash Zahirnia, Rachel Lawrence, Richard Turner, Edward Meeds, Javier Zazo, Sushrut Karmalkar

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08197 2025-05-19 cs.CV 79%

Visual Watermarking in the Era of Diffusion Models: Advances and Challenges

Junxian Duan, Jiyang Guan, Wenkui Yang, Ran He

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA, Beijing, China(多模态人工智能系统国家重点实验室,中国科学院自动化研究所,北京)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18440 2025-05-19 astro-ph.GA cs.CV 79%

Understanding Galaxy Morphology Evolution Through Cosmic Time via Redshift Conditioned Diffusion Models

Andrew Lizarraga, Eric Hanchen Jiang, Jacob Nowack, Yun Qi Li, Ying Nian Wu, Bernie Boscoe, Tuan Do

机构 * Department of Statistics and Data Science, UCLA(UCLA统计与数据科学系) Department of Computer Science, Southern Oregon University(南方俄勒冈大学计算机科学系) Department of Physics and Astronomy, University of Washington(华盛顿大学物理与天文学系) Department of Physics and Astronomy, UCLA(UCLA物理与天文学系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09985 2025-05-16 eess.IV cs.CV 79%

Ordered-subsets Multi-diffusion Model for Sparse-view CT Reconstruction

Pengfei Yu, Bin Huang, Minghui Zhang, Weiwen Wu, Shaoyu Wang, Qiegen Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09965 2025-05-16 cs.CV 79%

MambaControl: Anatomy Graph-Enhanced Mamba ControlNet with Fourier Refinement for Diffusion-Based Disease Trajectory Prediction

Hao Yang, Tao Tan, Shuai Tan, Weiqin Yang, Kunyan Cai, Calvin Chen, Yue Sun

机构 * Faculty of Applied Sciences, Macao Polytechnic University(应用科学学院,澳门理工学院) Department of Electrical Engineering, Zhejiang University(电气工程学院,浙江大学) Department of Computer Science, The University of Adelaide(计算机科学学院,阿德莱德大学) Department of Computer Science, University of Birmingham(计算机科学学院,伯明翰大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09858 2025-05-16 cs.CV 79%

Mission Balance: Generating Under-represented Class Samples using Video Diffusion Models

Danush Kumar Venkatesh, Isabel Funke, Micha Pfeiffer, Fiona Kolbinger, Hanna Maria Schmeiser, Juergen Weitz, Marius Distler, Stefanie Speidel

机构 * NCT/UCC Dresden(德累斯顿NCT/UCC) DKFZ Heidelberg(海德堡DKFZ) Faculty of Medicine & University Hospital Carl Gustav Carus TU Dresden(德累斯顿技术大学医学学院及卡尔·古斯塔夫·卡尔医院) HZDR Dresden, Germany(德累斯顿HZDR) Department of Translational Surgical Oncology, NCT/UCC Dresden, Faculty of Medicine & University Hospital Carl Gustav Carus Germany(德累斯顿NCT/UCC医学系及卡尔·古斯塔夫·卡尔医院) The Centre for Tactile Internet with Human-in-the-Loop (CeTI), TUD Dresden, Germany(德累斯顿TUD tactile internet中心) Weldon School of Biomedical Engineering, Regenstrief Center for Healthcare Engineering (RCHE), Purdue University, USA(美国普渡大学生物医学工程学院及Regenstrief医疗工程中心) Department of Biostatistics and Health Data Science, Richard M. Fairbanks School of Public Health, Indiana University, USA(美国印第安纳大学公共健康学院生物统计学与健康数据科学系) Department of Visceral, Thoracic and Vascular Surgery, University Hospital & Faculty of Medicine Carl Gustav Carus, TUD Germany(德累斯顿技术大学 visceral、胸腔及血管外科系及卡尔·古斯塔夫·卡尔医院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Early accept at MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09140 2025-05-15 cs.CV 79%

TopoDiT-3D: Topology-Aware Diffusion Transformer with Bottleneck Structure for 3D Point Cloud Generation

Zechao Guan, Feng Yan, Shuai Du, Lin Ma, Qingshan Liu

机构 * Southeast University(东南大学) Meituan(美团)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08833 2025-05-15 cs.CV cs.LG 79%

Generative AI for Urban Planning: Synthesizing Satellite Imagery via Diffusion Models

Qingyi Wang, Yuebing Liang, Yunhan Zheng, Kaiyuan Xu, Jinhua Zhao, Shenhao Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15511 2025-05-15 stat.CO cs.CV cs.LG 79%

Bayesian computation with generative diffusion models by Multilevel Monte Carlo

Abdul-Lateef Haji-Ali, Marcelo Pereyra, Luke Shaw, Konstantinos Zygalakis

机构 * School of Mathematical and Computer Sciences, Heriot-Watt University(赫瑞斯泰学院数学与计算机科学系) Departament de Matemàtiques and IMAC, Universitat Jaume I(数学系和IMAC, Jaime I大学) School of Mathematics, University of Edinburgh(爱丁堡大学数学系) Maxwell Institute for Mathematical Sciences, Edinburgh(爱丁堡数学科学研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 13 images

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01953 2025-05-14 eess.IV cs.CV cs.LG 79%

Deep Representation Learning for Unsupervised Clustering of Myocardial Fiber Trajectories in Cardiac Diffusion Tensor Imaging

Mohini Anand, Xavier Tricoche

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 5 figures. An extended journal manuscript is in preparation

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08281 2025-05-14 cs.CV eess.IV 79%

Ultra Lowrate Image Compression with Semantic Residual Coding and Compression-aware Diffusion

Anle Ke, Xu Zhang, Tong Chen, Ming Lu, Chao Zhou, Jiawen Gu, Zhan Ma

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08235 2025-05-14 cs.CV 79%

EventDiff: A Unified and Efficient Diffusion Model Framework for Event-based Video Frame Interpolation

Hanle Zheng, Xujie Han, Zegang Peng, Shangbin Zhang, Guangxun Du, Zhuo Zou, Xilin Wang, Jibin Wu, Hao Guo, Lei Deng

机构 * Center for Brain Inspired Computing Research (CBICR), Department of Precision Instrument, Tsinghua University(脑启发计算研究中心(CBICR)、精密仪器系,清华大学) College of Computer Science and Technology, Taiyuan University of Technology(计算机科学与技术学院,太原理工大学) The Information Science Academy of China Electronics Technology Group Corporation(中国电子科技集团有限公司信息科学学院) State Key Laboratory of Integrated Chips and Systems, School of Information Science and Technology, Fudan University(集成电路与系统集成国家重点实验室,复旦大学信息科学学院) Engineering Laboratory of Power Equipment Reliability in Complicated Coastal Environments, Tsinghua Shenzhen International Graduate School, Tsinghua University(复杂近海环境电力设备可靠性工程实验室,清华大学深圳国际研究生院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07866 2025-05-14 eess.IV cs.AI cs.CV 79%

Computationally Efficient Diffusion Models in Medical Imaging: A Comprehensive Review

Abdullah, Tao Huang, Ickjai Lee, Euijoon Ahn

机构 * James Cook University(詹姆斯库克大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments pages 36, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21650 2025-05-14 cs.CV 79%

HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generation

Haiyang Zhou, Wangbo Yu, Jiawen Guan, Xinhua Cheng, Yonghong Tian, Li Yuan

机构 * School of Electronic and Computer Engineering, Peking University(北京大学电子与计算机工程学院) Peng Cheng Laboratory(鹏城实验室) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Homepage: https://zhouhyocean.github.io/holotime/ Code: https://github.com/PKU-YuanGroup/HoloTime

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15026 2025-05-14 cs.CV cs.CR 79%

Gaussian Shading++: Rethinking the Realistic Deployment Challenge of Performance-Lossless Image Watermark for Diffusion Models

Zijin Yang, Xin Zhang, Kejiang Chen, Kai Zeng, Qiyi Yao, Han Fang, Weiming Zhang, Nenghai Yu

机构 * University of Science and Technology of China(中国科学技术大学) Anhui Province Key Laboratory of Digital Security(安徽省数字安全重点实验室) Department of Information Engineering and Mathematics(信息工程与数学系) University of Siena(锡耶纳大学) School of Computing(计算学院) National University of Singapore(新加坡国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 18 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04698 2025-05-14 cs.CV 79%

ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning

Yuzhou Huang, Ziyang Yuan, Quande Liu, Qiulin Wang, Xintao Wang, Ruimao Zhang, Pengfei Wan, Di Zhang, Kun Gai

机构 * Sun Yat-sen University(中山大学) Kuaishou Technology(快手科技) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Tsinghua University(清华大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://yuzhou914.github.io/ConceptMaster/. Update and release MCVC Evaluation Set

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07652 2025-05-13 cs.CV 79%

ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models

Ozgur Kara, Krishna Kumar Singh, Feng Liu, Duygu Ceylan, James M. Rehg, Tobias Hinz

机构 * UIUC(伊利诺伊大学香槟分校) Adobe(Adobe公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07548 2025-05-13 cs.LG cs.AI cs.CV 79%

Noise Optimized Conditional Diffusion for Domain Adaptation

Lingkun Luo, Shiqiang Hu, Liming Chen

机构 * School of Aeronautics and Astronautics, Shanghai Jiao Tong University, Shanghai, China(上海交通大学航空宇航学院) LIRIS, CNRS UMR 5205, Ecole Centrale de Lyon, France(法国里尔研究所) Institut Universitaire de France (IUF), France(法国国家科学研究中心(IUF))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 9 pages, 4 figures This work has been accepted by the International Joint Conference on Artificial Intelligence (IJCAI 2025)

Journal ref IJCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07481 2025-05-13 cs.CV 79%

Addressing degeneracies in latent interpolation for diffusion models

Erik Landolsi, Fredrik Kahl

机构 * Chalmers University of Technology(查尔姆斯理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07477 2025-05-13 cs.LG cs.CV 79%

You Only Look One Step: Accelerating Backpropagation in Diffusion Sampling with Gradient Shortcuts

Hongkun Dou, Zeyu Li, Xingyu Jiang, Hongjue Li, Lijun Yang, Wen Yao, Yue Deng

机构 * School of Astronautics, Beihang University(北航航天学院) Defense Innovation Institute, Chinese Academy of Military Science(国防创新研究院,中国军事科学院) Institute of Artificial Intelligence, Beihang University(北航人工智能学院) Beijing Zhongguancun Academy(中关村学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13859 2025-05-13 cs.CV 79%

Less is More: Improving Motion Diffusion Models with Sparse Keyframes

Jinseok Bae, Inwoo Hwang, Young Yoon Lee, Ziyu Guo, Joseph Liu, Yizhak Ben-Shabat, Young Min Kim, Mubbasir Kapadia

机构 * Seoul National University(首尔国立大学) Roblox(Roblox公司) The Chinese University of Hong Kong(香港中文大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16003 2025-05-13 cs.CV physics.ao-ph 79%

Improving Tropical Cyclone Forecasting With Video Diffusion Models

Zhibo Ren, Pritthijit Nath, Pancham Shukla

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Recommended for spotlight presentation at the ICLR 2025 workshop on Tackling Climate Change with Machine Learning. 7 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07071 2025-05-13 cs.CV 79%

Semantic-Guided Diffusion Model for Single-Step Image Super-Resolution

Zihang Liu, Zhenyu Zhang, Hao Tang

机构 * Beijing Institute of Technology(北京理工大学) Nanjing University(南京大学) School of Computer Science, Peking University(北京大学计算机学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07057 2025-05-13 cs.CV 79%

DAPE: Dual-Stage Parameter-Efficient Fine-Tuning for Consistent Video Editing with Diffusion Models

Junhao Xia, Chaoyang Zhang, Yecheng Zhang, Chengyang Zhou, Zhichang Wang, Bochun Liu, Dongshuo Yin

机构 * Tsinghua University(清华大学) Duke University(杜克大学) Peking University(北京大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏