arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86714 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70159 篇

2506.21544 2025-07-01 cs.CV 83%

DeOcc-1-to-3: 3D De-Occlusion from a Single Image via Self-Supervised Multi-View Diffusion

Yansong Qu, Shaohui Dai, Xinyang Li, Yuze Wang, You Shen, Liujuan Cao, Rongrong Ji

机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室、厦门大学) State Key Laboratory of Virtual Reality Technology and Systems, Beihang University(虚拟现实技术与系统国家重点实验室、北京航空航天大学)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Project page: \url{https://quyans.github.io/DeOcc123/}

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09039 2025-07-01 cs.CV cs.AI cs.LG 83%

Sculpting Memory: Multi-Concept Forgetting in Diffusion Models via Dynamic Mask and Concept-Aware Optimization

Gen Li, Yang Xiao, Jie Ji, Kaiyuan Deng, Bo Hui, Linke Guo, Xiaolong Ma

机构 * Clemson University(克莱姆森大学) University of Tulsa(塔尔萨大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments ICCV2025(Accept)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13349 2025-07-01 cs.CV 83%

MSF: Efficient Diffusion Model Via Multi-Scale Latent Factorize

Haohang Xu, Longyu Chen, Yichen Zhang, Shuangrui Ding, Zhipeng Zhang

机构 * Huawei Inc.(华为公司) Huazhong University of Science & Technology(华中科技大学) The Chinese University of Hong Kong(香港中文大学) School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21977 2025-06-30 eess.IV cs.CV 83%

StableCodec: Taming One-Step Diffusion for Extreme Image Compression

Tianyu Zhang, Xin Luo, Li Li, Dong Liu

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21681 2025-06-30 cs.CV cs.LG 83%

TanDiT: Tangent-Plane Diffusion Transformer for High-Quality 360° Panorama Generation

Hakan Çapuk, Andrew Bond, Muhammed Burak Kızıl, Emir Göçen, Erkut Erdem, Aykut Erdem

机构 * Koç University(科克大学) Hacettepe University(哈切特佩大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21270 2025-06-27 cs.CV 83%

Video Virtual Try-on with Conditional Diffusion Transformer Inpainter

Cheng Zou, Senlin Cheng, Bolei Xu, Dandan Zheng, Xiaobo Li, Jingdong Chen, Ming Yang

机构 * Ant Group(蚂蚁集团)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01923 2025-06-27 cs.CV cs.AI cs.LG 83%

TaxaDiffusion: Progressively Trained Diffusion Model for Fine-Grained Species Generation

Amin Karimi Monsefi, Mridul Khurana, Rajiv Ramnath, Anuj Karpatne, Wei-Lun Chao, Cheng Zhang

机构 * The Ohio State University(俄亥俄州立大学) Virginia Tech(弗吉尼亚理工大学) Texas A&M University(德克萨斯大学阿姆斯特朗分校)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08914 2025-06-26 cs.CV cs.AI 83%

Diffusion Models Through a Global Lens: Are They Culturally Inclusive?

Zahra Bayramli, Ayhan Suleymanzade, Na Min An, Huzama Ahmad, Eunsu Kim, Junyeong Park, James Thorne, Alice Oh

机构 * School of Computing, KAIST(韩国成均馆大学计算机学院) Graduate School of AI, KAIST(韩国成均馆大学人工智能研究生院)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 17 pages, 17 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19455 2025-06-25 eess.IV cs.CV 83%

Angio-Diff: Learning a Self-Supervised Adversarial Diffusion Model for Angiographic Geometry Generation

Zhifeng Wang, Renjiao Yi, Xin Wen, Chenyang Zhu, Kai Xu, Kunlun He

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16776 2025-06-25 eess.IV cs.CV cs.LG 83%

Diff-Def: Diffusion-Generated Deformation Fields for Conditional Atlases

Sophie Starck, Vasiliki Sideri-Lampretsa, Bernhard Kainz, Martin J. Menten, Tamara T. Mueller, Daniel Rueckert

机构 * School of Computation, Information and Technology and the School of Medicine and Health, TUM Klinikum, Technical University of Munich(计算信息学院和医学健康学院,慕尼黑技术大学克林克医院,慕尼黑技术大学) Munich Center for Machine Learning (MCML), Munich, Germany(慕尼黑机器学习中心(MCML),德国慕尼黑) Department of Computing, Imperial College London, UK(计算学院,伦敦帝国学院,英国) FAU Erlangen-Nürnberg, Germany(埃尔兰根-纽伦堡大学,德国)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17975 2025-06-24 cs.CV 83%

Enabling PSO-Secure Synthetic Data Sharing Using Diversity-Aware Diffusion Models

Mischa Dombrowski, Bernhard Kainz

机构 * Department of Computing, Imperial College London, London, UK(伦敦帝国学院计算机系)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17705 2025-06-24 cs.CV 83%

DreamJourney: Perpetual View Generation with Video Diffusion Models

Bo Pan, Yang Chen, Yingwei Pan, Ting Yao, Wei Chen, Tao Mei

机构 * State Key Lab of CAD&CG, Zhejiang University(浙江大学CAD与CG国家重点实验室) Laboratory of Art and Archaeology Image(艺术与考古图像实验室)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.04585 2025-06-24 cs.CV cs.LG 83%

EDA-DM: Enhanced Distribution Alignment for Post-Training Quantization of Diffusion Models

Xuewen Liu, Zhikai Li, Junrui Xiao, Mengjuan Chen, Jianquan Li, Qingyi Gu

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Code: http://github.com/BienLuky/EDA-DM

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16743 2025-06-23 cs.CV 83%

Noise-Informed Diffusion-Generated Image Detection with Anomaly Attention

Weinan Guan, Wei Wang, Bo Peng, Ziwen He, Jing Dong, Haonan Cheng

机构 * School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所模式识别实验室) Nanjing University of Information Science and Technology(南京信息科学技术大学) State Key Laboratory of Media Convergence and Communication, Communication University of China(中国传媒大学媒体融合与传播国家重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted by TIFS 2025. Our code is availabel at https://github.com/WeinanGuan/NASA-Swin

Journal ref IEEE Trans. Inf. Forensics Security, vol.20, pp. 5256-5268, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14919 2025-06-19 cs.CV cs.LG 83%

Frequency-Calibrated Membership Inference Attacks on Medical Image Diffusion Models

Xinkai Zhao, Yuta Tokuoka, Junichiro Iwasawa, Keita Oda

机构 * Graduate School of Informatics, Nagoya University, Japan(名古屋大学信息科学研究生院, 日本) Preferred Networks, Inc., Japan(Preferred Networks, Inc., 日本)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14560 2025-06-18 cs.CV cs.LG 83%

Risk Estimation of Knee Osteoarthritis Progression via Predictive Multi-task Modelling from Efficient Diffusion Model using X-ray Images

David Butler, Adrian Hilton, Gustavo Carneiro

机构 * Centre for Vision, Speech and Signal Processing(视觉、语音和信号处理中心)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13484 2025-06-17 cs.CV eess.IV 83%

Deep Diffusion Models and Unsupervised Hyperspectral Unmixing for Realistic Abundance Map Synthesis

Martina Pastorino, Michael Alibani, Nicola Acito, Gabriele Moser

机构 * DITEN, University of Genoa(迪滕,热那亚大学) DII, University of Pisa(迪II,比萨大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments CVPRw2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13614 2025-06-17 stat.ML cs.CV cs.LG 83%

Exploiting the Exact Denoising Posterior Score in Training-Free Guidance of Diffusion Models

Gregory Bellchambers

机构 * PhysicsX Ltd.(PhysicsX有限公司)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07225 2025-06-17 cs.CV 83%

CAT: Contrastive Adversarial Training for Evaluating the Robustness of Protective Perturbations in Latent Diffusion Models

Sen Peng, Mingyue Wang, Jianfei He, Jijia Yang, Xiaohua Jia

机构 * Department of Computer Science, City University of Hong Kong, Kowloon, Hong Kong SAR(香港城市大学计算机科学系)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02099 2025-06-17 cs.CV cs.AI 83%

AccDiffusion v2: Towards More Accurate Higher-Resolution Diffusion Extrapolation

Zhihang Lin, Mingbao Lin, Wengyi Zhan, Rongrong Ji

机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University, China(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学,中国) Shanghai Innovation Institute, Shanghai, China(上海创新研究院,上海,中国) Rakuten Asia Pte. Ltd., Singapore 048946(Rakuten Asia(新加坡)) Institute of Artificial Intelligence, Xiamen University, Xiamen 361005, China(厦门大学人工智能研究院,厦门,中国)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 13 pages. arXiv admin note: text overlap with arXiv:2407.10738

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07903 2025-06-16 cs.LG cs.AI cs.CV 83%

Diffuse Everything: Multimodal Diffusion Models on Arbitrary State Spaces

Kevin Rojas, Yuchen Zhu, Sichen Zhu, Felix X. -F. Ye, Molei Tao

机构 * Machine Learning Center, Georgia Institute of Technology, Atlanta, GA School of Mathematics, Georgia Institute of Technology, Atlanta, GA Department of Mathematics \& Statistics, SUNY Albany, NY

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to ICML 2025. Code available at https://github.com/KevinRojas1499/Diffuse-Everything

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10233 2025-06-13 eess.IV cs.CV 83%

Conditional diffusion models for guided anomaly detection in brain images using fluid-driven anomaly randomization

Ana Lawry Aguila, Peirong Liu, Oula Puonti, Juan Eugenio Iglesias

机构 * Athinoula A. Martinos Center for Biomedical Imaging, Massachusetts General Hospital and Harvard Medical School, Boston, USA(Athinoula A. Martinos 生物医学成像中心、麻省总医院和哈佛医学院,波士顿,美国) Danish Research Centre for Magnetic Resonance, Department of Radiology and Nuclear Medicine(丹麦磁共振研究中心、放射学与核医学部门) Copenhagen University Hospital– Amager and Hvidovre(哥本哈根大学医院–阿马格和赫维多尔)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10038 2025-06-13 cs.GR cs.AI cs.LG 83%

Ambient Diffusion Omni: Training Good Models with Bad Data

Giannis Daras, Adrian Rodriguez-Munoz, Adam Klivans, Antonio Torralba, Constantinos Daskalakis

机构 * Massachusetts Institute of Technology(麻省理工学院) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.GR

Comments Preprint, work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14398 2025-06-12 cs.CV 83%

Dynamic Negative Guidance of Diffusion Models

Felix Koulischer, Johannes Deleu, Gabriel Raya, Thomas Demeester, Luca Ambrogioni

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Paper accepted at ICLR 2025 (poster). Our implementation is available at https://github.com/FelixKoulischer/Dynamic-Negative-Guidance.git

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.07944 2025-06-12 cs.CV cs.AI 83%

Effective Data Augmentation With Diffusion Models

Brandon Trabucco, Kyle Doherty, Max Gurinas, Ruslan Salakhutdinov

机构 * Carnegie Mellon University(卡内基梅隆大学) MPG Ranch(MPG牧场) University Of Chicago(芝加哥大学) Laboratory Schools(实验室学校)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Update to ICLR 2024 manuscript (https://openreview.net/forum?id=ZWzUA9zeAg), add leafy spurge citations

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07590 2025-06-10 cs.CV cs.LG 83%

Explore the vulnerability of black-box models via diffusion models

Jiacheng Shi, Yanfu Zhang, Huajie Shao, Ashley Gao

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.17929 2025-06-10 cs.CV cs.LG 83%

GLASS: Guided Latent Slot Diffusion for Object-Centric Learning

Krishnakant Singh, Simone Schaub-Meyer, Stefan Roth

机构 * Department of Computer Science, TU Darmstadt(图宾根大学计算机科学系)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments CVPR 2025. Project Page: http://visinf.github.io/glass/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05934 2025-06-09 cs.CV cs.AI 83%

FADE: Frequency-Aware Diffusion Model Factorization for Video Editing

Yixuan Zhu, Haolin Wang, Shilin Ma, Wenliang Zhao, Yansong Tang, Lei Chen, Jie Zhou

机构 * Department of Automation, Tsinghua University(自动化系,清华大学) Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学)

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

Comments Accepted by IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09625 2025-06-09 cs.CV 83%

Illusion3D: 3D Multiview Illusion with 2D Diffusion Priors

Yue Feng, Vaibhav Sanjay, Spencer Lutz, Badour AlBahar, Songwei Ge, Jia-Bin Huang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Project page: https://3d-multiview-illusion.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02528 2025-06-04 cs.CV 83%

RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers

Yan Gong, Yiren Song, Yicheng Li, Chenglin Li, Yin Zhang

机构 * Zhe Jiang University(浙江大学) National University of Singapore(新加坡国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏