arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86714 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70159 篇

2408.06157 2025-11-18 cs.CV 83%

3D-free meets 3D priors: Novel View Synthesis from a Single Image with Pretrained Diffusion Guidance

Taewon Kang, Divya Kothandaraman, Dinesh Manocha, Ming C. Lin

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to The 40th Annual AAAI Conference on Artificial Intelligence (AAAI-26), AAAI 2026 Workshop on AI for Environmental Science (AI4ES). Due to arXiv's 1,920-character limit, the abstract here is shortened. Please refer to the paper (View PDF) to read the full abstract. 14 pages, 13 figures, v5: AAAI-26 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11231 2025-11-17 cs.CV cs.AI 83%

3D Gaussian and Diffusion-Based Gaze Redirection

Abiram Panchalingam, Indu Bodala, Stuart Middleton

机构 * School of Electronics and Computer Science, University of Southampton(电子与计算机科学学院,南安普顿大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08207 2025-11-14 cs.CV cs.LG 83%

DICE: Discrete Inversion Enabling Controllable Editing for Multinomial Diffusion and Masked Generative Models

Xiaoxiao He, Quan Dao, Ligong Han, Song Wen, Minhao Bai, Di Liu, Han Zhang, Martin Renqiang Min, Felix Juefei-Xu, Chaowei Tan, Bo Liu, Kang Li, Hongdong Li, Junzhou Huang, Faez Ahmed, Akash Srivastava, Dimitris Metaxas

机构 * Rutgers University(新泽西罗格斯大学) MIT-IBM Watson AI Lab(MIT-IBM沃森人工智能实验室) Red Hat AI Innovation(红帽AI创新) Google DeepMind(谷歌DeepMind) NYU(纽约大学) Walmart Global Tech(沃尔玛全球技术) NEC Labs America(NEC美国实验室) Massachusetts Institute of Technology(麻省理工学院) ANU(澳大利亚国立大学) UT Arlington(德克萨斯大学阿灵顿分校)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Project webpage: https://hexiaoxiao-cs.github.io/DICE/. This paper was accepted to CVPR 2025 but later desk-rejected post camera-ready, due to a withdrawal from ICLR made 14 days before reviewer assignment

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08090 2025-11-12 cs.CV cs.AI 83%

StableMorph: High-Quality Face Morph Generation with Stable Diffusion

Wassim Kabbani, Kiran Raja, Raghavendra Ramachandra, Christoph Busch

机构 * Norwegian University of Science and Technology(挪威科学技术大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Journal ref International Joint Conference on Biometrics 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07806 2025-11-12 cs.CV 83%

PC-Diffusion: Aligning Diffusion Models with Human Preferences via Preference Classifier

Shaomeng Wang, He Wang, Xiaolu Wei, Longquan Dai, Jinhui Tang

机构 * School of Computer Science and Engineering, Nanjing University of Science and Technology(计算机科学与工程学院,南京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 10 pages, 3 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07499 2025-11-12 cs.CV cs.AI 83%

Toward the Frontiers of Reliable Diffusion Sampling via Adversarial Sinkhorn Attention Guidance

Kwanyoung Kim

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted to AAAI 26

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24434 2025-11-11 cs.LG cs.CV 83%

Graph Flow Matching: Enhancing Image Generation with Neighbor-Aware Flow Fields

Md Shahriar Rahim Siddiqui, Moshe Eliasof, Eldad Haber

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments The 40th Annual AAAI Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18877 2025-11-11 cs.CV cs.CR cs.LG 83%

Mitigating Sexual Content Generation via Embedding Distortion in Text-conditioned Diffusion Models

Jaesin Ahn, Heechul Jung

机构 * Department of Artificial Intelligence(人工智能系) Kyungpook National University(全北国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments NeurIPS 2025 accepted. Official code: https://github.com/amoeba04/des

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09935 2025-11-11 eess.IV cs.CV physics.med-ph 83%

Physics-informed DeepCT: Sinogram Wavelet Decomposition Meets Masked Diffusion

Zekun Zhou, Tan Liu, Bing Yu, Yanru Gong, Liu Shi, Qiegen Liu

机构 * School of Mathematics and Computer Sciences, Nanchang University, Nanchang, China(南昌大学数学与计算机科学学院) School of Information Engineering, Nanchang University, Nanchang, China(南昌大学信息工程学院) Key Laboratory of Advanced Medical Imaging and Intelligent Computing of Guizhou Province, China(贵州省先进医学影像与智能计算重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04117 2025-11-07 cs.CV 83%

Tortoise and Hare Guidance: Accelerating Diffusion Model Inference with Multirate Integration

Yunghee Lee, Byeonghyun Pak, Junwha Hong, Hoseong Kim

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments 21 pages, 8 figures. NeurIPS 2025. Project page: https://yhlee-add.github.io/THG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27171 2025-11-06 cs.CV cs.AI 83%

H2-Cache: A Novel Hierarchical Dual-Stage Cache for High-Performance Acceleration of Generative Diffusion Models

Mingyu Sung, Il-Min Kim, Sangseok Yun, Jae-Mo Kang

机构 * Department of Artificial Intelligence, Kyungpook National University(人工智能系,庆尚国立大学) Department of Electrical and Computer Engineering, Queen’s University(电气与计算机工程系,皇后大学) Department of Information and Communications Engineering, Pukyong National University(信息与通信工程系,浦项国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02462 2025-11-05 cs.CV 83%

KAO: Kernel-Adaptive Optimization in Diffusion for Satellite Image

Teerapong Panboonyuen

机构 * Chulalongkorn University(朱拉隆功大学)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments 18 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15724 2025-11-05 cs.CV 83%

A Practical Investigation of Spatially-Controlled Image Generation with Transformers

Guoxuan Xia, Harleen Hanspal, Petru-Daniel Tudosiu, Shifeng Zhang, Sarah Parisot

机构 * Huawei Noah’s Ark Lab(华为诺亚实验室)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments TMLR https://openreview.net/forum?id=loT6xhgLYK

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01175 2025-11-05 cs.CV 83%

Diffusion Transformer meets Multi-level Wavelet Spectrum for Single Image Super-Resolution

Peng Du, Hui Li, Han Xu, Paul Barom Jeon, Dongwook Lee, Daehyun Ji, Ran Yang, Feng Zhu

机构 * Samsung R&D Institute China Xi’an (SRCX)(三星中国研发中心西安(SRCX)) Samsung Electronics Co., LTD., South Korea(三星电子有限公司,韩国)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments ICCV 2025 Oral Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21015 2025-11-05 cs.CV cs.LG quant-ph 83%

MediQ-GAN: Quantum-Inspired GAN for High Resolution Medical Image Generation

Qingyue Jiao, Yongcan Tang, Jun Zhuang, Jason Cong, Yiyu Shi

机构 * Independent Researcher(独立研究者)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01645 2025-11-04 cs.CV 83%

Enhancing Diffusion-based Restoration Models via Difficulty-Adaptive Reinforcement Learning with IQA Reward

Xiaogang Xu, Ruihang Chu, Jian Wang, Kun Zhou, Wenjie Shu, Harry Yang, Ser-Nam Lim, Hao Chen, Liang Lin

机构 * The Chinese University of Hong Kong(香港中文大学) Tsinghua University(清华大学) Snap Research Shenzhen University(深圳大学) HKUST(香港科技大学) University of Central Florida(佛罗里达大学) UC Davis(加州大学戴维斯分校) Sun Yat-Sen University(孙中山大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01466 2025-11-04 cs.CV 83%

SecDiff: Diffusion-Aided Secure Deep Joint Source-Channel Coding Against Adversarial Attacks

Changyuan Zhao, Jiacheng Wang, Ruichen Zhang, Dusit Niyato, Hongyang Du, Zehui Xiong, Dong In Kim, Ping Zhang

机构 * College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) Department of Electrical and Electronic Engineering, University of Hong Kong(电子与电气工程系,香港大学) School of Electronics, Electrical Engineering and Computer Science, Queen’s University Belfast(电子、电气工程与计算机科学学院,女王学院贝尔法斯特分校) Department of Electrical and Computer Engineering, Sungkyunkwan University(电气与计算机工程系,成均馆大学) State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications(网络与交换技术国家重点实验室,北京邮电大学)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments 13 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07140 2025-11-04 cs.CV 83%

FIRE: Robust Detection of Diffusion-Generated Images via Frequency-Guided Reconstruction Error

Beilin Chu, Xuan Xu, Xin Wang, Yufei Zhang, Weike You, Linna Zhou

机构 * School of CyberSpace Security, Beijing University of Posts and Telecommunications(网络安全学院,北京邮电大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 14 pages, 14 figures. Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00429 2025-11-04 cs.CV cs.AI 83%

Enhancing Frequency Forgery Clues for Diffusion-Generated Image Detection

Daichi Zhang, Tong Zhang, Shiming Ge, Sabine Süsstrunk

机构 * School of Computer and Communication Sciences, EPFL(瑞士联邦理工学院计算机与通信科学学院) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) University of Chinese Academy of Sciences(中国科学院大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10978 2025-11-04 cs.CV cs.AI cs.LG 83%

Where and How to Perturb: On the Design of Perturbation Guidance in Diffusion and Flow Models

Donghoon Ahn, Jiwon Kang, Sanghyun Lee, Minjae Kim, Jaewon Min, Wooseok Jang, Sangwu Lee, Sayak Paul, Susung Hong, Seungryong Kim

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted at NeurIPS 2025. Project page: https://cvlab-kaist.github.io/HeadHunter/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17955 2025-11-04 cs.CV 83%

Diffusion Classifiers Understand Compositionality, but Conditions Apply

Yujin Jeong, Arnas Uselis, Seong Joon Oh, Anna Rohrbach

机构 * TU Darmstadt(图恩大学) University of Tübingen(图宾根大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments NeurIPS 2025 Datasets and Benchmarks

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.12245 2025-11-04 cs.CV 83%

Scalable Autoregressive Image Generation with Mamba

Haopeng Li, Jinyue Yang, Kexin Wang, Xuerui Qiu, Yuhong Chou, Xin Li, Guoqi Li

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments 9 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04687 2025-11-04 cs.CV cs.AI 83%

Targeted Attack Improves Protection against Unauthorized Diffusion Customization

Boyang Zheng, Chumeng Liang, Xiaoyu Wu

机构 * Shanghai Jiao Tong University(上海交通大学) University of Southern California(南加州大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments ICLR 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.12935 2025-11-04 cs.LG cs.AI cs.CV 83%

ERA-Solver: Error-Robust Adams Solver for Fast Sampling of Diffusion Probabilistic Models

Shengming Li, Luping Liu, Runnan Li, Xu Tan

机构 * Zhejiang University(浙江大学) Microsoft Azure Speech(微软Azure语音) Microsoft Research Asia(微软亚洲研究院)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00011 2025-11-04 cs.CV cs.AI cs.HC 83%

Generative human motion mimicking through feature extraction in denoising diffusion settings

Alexander Okupnik, Johannes Schneider, Kyriakos Flouris

机构 * University of Liechtenstein(列支敦士登大学) University of Cambridge(剑桥大学)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26391 2025-10-31 cs.CV 83%

EEG-Driven Image Reconstruction with Saliency-Guided Diffusion Models

Igor Abramov, Ilya Makarov

机构 * Ivannikov Institute for System Programming of the Russian Academy of Sciences(俄罗斯科学院伊万诺夫系统编程研究所) Research Center for Trusted Artificial Intelligence(可信人工智能研究中心) AI Talent Hub, ITMO University(ITMO大学人工智能人才中心) Moscow, Russia(俄罗斯莫斯科)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Demo paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24805 2025-10-30 q-bio.QM cs.AI cs.CV 83%

CT-Less Attenuation Correction Using Multiview Ensemble Conditional Diffusion Model on High-Resolution Uncorrected PET Images

Alexandre St-Georges, Gabriel Richard, Maxime Toussaint, Christian Thibaudeau, Etienne Auger, Étienne Croteau, Stephen Cunnane, Roger Lecomte, Jean-Baptiste Michaud

机构 * Department of Medical Imaging and Radiation Sciences(医学影像与放射科学系) Université de Sherbrooke(Sherbrooke大学) Sherbrooke Molecular Imaging Center(Sherbrooke分子影像中心) CRCHUS Nantes Université(Nantes大学) Laboratoire CRCI2NA, INSERM, CNRS(CRCI2NA实验室,INSERM,CNRS) Imaging Research and Technology (IR&T) Inc.(影像研究与技术公司) Research Center on Aging, Department of Medicine(衰老研究中心,医学系) Department of Electrical Engineering and Computer Engineering(电气工程与计算机工程系)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments This is a preprint and not the final version of this paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22366 2025-10-28 cs.CV cs.AI 83%

T2SMark: Balancing Robustness and Diversity in Noise-as-Watermark for Diffusion Models

Jindong Yang, Han Fang, Weiming Zhang, Nenghai Yu, Kejiang Chen

机构 * University of Science and Technology of China(中国科学技术大学) Anhui Province Key Laboratory of Digital Security(安徽省数字安全重点实验室) National University of Singapore(新加坡国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17131 2025-10-28 cs.CV cs.AI 83%

GOOD: Training-Free Guided Diffusion Sampling for Out-of-Distribution Detection

Xin Gao, Jiyao Liu, Guanghao Li, Yueming Lyu, Jianxiong Gao, Weichen Yu, Ningsheng Xu, Liang Wang, Caifeng Shan, Ziwei Liu, Chenyang Si

机构 * Nanjing University(南京大学) Fudan University(复旦大学) Carnegie Mellon University(卡内基梅隆大学) Chinese Academy of Sciences(中国科学院) Nanyang Technological University(南洋理工大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 28 pages, 16 figures, conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09998 2025-10-27 cs.LG cs.CV 83%

Adaptive Non-uniform Timestep Sampling for Accelerating Diffusion Model Training

Myunsoo Kim, Donghyeon Ki, Seong-Woong Shim, Byung-Jun Lee

机构 * Korea University(韩国大学) Gauss Labs Inc.(Gauss实验室公司)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏