arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-07-29 至 2025-07-29 共收录 92 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 5 篇

2502.15278 2025-07-29 cs.CV cs.AI 88%

CopyJudge: Automated Copyright Infringement Identification and Mitigation in Text-to-Image Diffusion Models

Shunchang Liu, Zhuan Shi, Lingjuan Lyu, Yaochu Jin, Boi Faltings

机构 * EPFL(苏黎世联邦理工学院) Mila - Quebec AI Institute Mcgill University(蒙特利尔麦吉尔大学人工智能研究所) Westlake University(西湖大学)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Accepted by ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.17046 2025-07-29 cs.CV 88%

Text-to-Image Generation Via Energy-Based CLIP

Roy Ganz, Michael Elad

机构 * Electrical Engineering Department Technion(技术学院电子工程系) Computer Science Department Technion(技术学院计算机科学系)

专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV

Comments Accepted to TMLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20934 2025-07-29 cs.CV 86%

Exploring text-to-image generation for historical document image retrieval

Melissa Cote, Alexandra Branzan Albu

机构 * University of Victoria(维多利亚大学)

专题命中 文生图 :text-to-image(title,abstract);image generation(title);分类 cs.CV

Comments Accepted and presented as an extended abstract (double-blind review process) at the 2025 Scandinavian Conference on Image Analysis (SCIA). 4 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20782 2025-07-29 cs.CV cs.AI 70%

Investigation of Accuracy and Bias in Face Recognition Trained with Synthetic Data

Pavel Korshunov, Ketan Kotwal, Christophe Ecabert, Vidit Vidit, Amir Mohammadi, Sebastien Marcel

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments Accepted for publication in IEEE International Joint Conference on Biometrics (IJCB), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.08789 2025-07-29 cs.CV 70%

Generative AI in Agriculture: Creating Image Datasets Using DALL.E's Advanced Large Language Model Capabilities

Ranjan Sapkota, Manoj Karkee

专题命中 文生图 :image generation(abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 图像编辑 3 篇

2505.23145 2025-07-29 cs.CV cs.AI cs.LG 83%

FlowAlign: Trajectory-Regularized, Inversion-Free Flow-based Image Editing

Jeongsol Kim, Yeobin Hong, Jonghyun Park, Jong Chul Ye

机构 * KAIST(韩国科学技术院)

专题命中 图像编辑 :image editing(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07317 2025-07-29 cs.CV 79%

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation

Sherry X. Chen, Yi Wei, Luowei Zhou, Suren Kumar

机构 * Samsung AI Center(三星AI中心) Mountain View University of California, Santa Barbara(山景城加州大学圣巴巴拉分校)

专题命中 图像编辑 :image editing(title,abstract);分类 cs.CV

Comments International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21033 2025-07-29 cs.CV 57%

GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Yuhan Wang, Siwei Yang, Bingchen Zhao, Letian Zhang, Qing Liu, Yuyin Zhou, Cihang Xie

机构 * University of California, Santa Cruz(加州大学圣克鲁兹分校) The University of Edinburgh(爱丁堡大学) Adobe Project Page(Adobe项目页面)

专题命中 图像编辑 :image editing(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 扩散模型 55 篇

2502.01189 2025-07-29 eess.IV cs.AI cs.CV cs.IT eess.SP math.IT 88%

Compressed Image Generation with Denoising Diffusion Codebook Models

Guy Ohayon, Hila Manor, Tomer Michaeli, Michael Elad

机构 * Faculty of Computer Science, Technion -- Israel Institute of Technology, Haifa, Israel(计算机科学系,技术离子理工学院——以色列理工学院,海法,以色列) Faculty of Electrical and Computer Engineering, Technion -- Israel Institute of Technology, Haifa, Israel(电气与计算机工程系,技术离子理工学院——以色列理工学院,海法,以色列)

专题命中 扩散模型 :image generation(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Published in the International Conference on Machine Learning (ICML) 2025. Code and demo are available at https://ddcm-2025.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20478 2025-07-29 cs.LG 88%

Conditional Diffusion Models for Global Precipitation Map Inpainting

Daiko Kishikawa, Yuka Muto, Shunji Kotsuki

专题命中 扩散模型 :diffusion(title,abstract);inpainting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17226 2025-07-29 cs.CV 83%

DDB: Diffusion Driven Balancing to Address Spurious Correlations

Aryan Yazdan Parast, Basim Azam, Naveed Akhtar

机构 * The University of Melbourne(墨尔本大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20363 2025-07-29 cs.CV 79%

Generative Pre-training for Subjective Tasks: A Diffusion Transformer-Based Framework for Facial Beauty Prediction

Djamel Eddine Boukhari, Ali chemsa

机构 * LGEERE Laboratory(LGEERE实验室) Department of Electrical Engineering,University of El Oued(电气工程系,埃尔欧德大学) Scientific and Technical Research Centre for Arid Areas(干旱地区科学技术研究中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20158 2025-07-29 cs.CV 79%

AnimeColor: Reference-based Animation Colorization with Diffusion Transformers

Yuhong Zhang, Liyao Wang, Han Wang, Danni Wu, Zuzeng Lin, Feng Wang, Li Song

机构 * Shanghai Jiao Tong University(上海交通大学) Tianjin University(天津大学) Communication University of China(中国通信大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19970 2025-07-29 eess.IV cs.CV cs.LG 79%

SkinDualGen: Prompt-Driven Diffusion for Simultaneous Image-Mask Generation in Skin Lesions

Zhaobin Xu

机构 * Shandong University, Qingdao, China(山东大学,青岛)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19874 2025-07-29 cs.CV 79%

All-in-One Medical Image Restoration with Latent Diffusion-Enhanced Vector-Quantized Codebook Prior

Haowei Chen, Zhiwen Yang, Haotian Hou, Hui Zhang, Bingzheng Wei, Gang Zhou, Yan Xu

机构 * School of Biological Science and Medical Engineering(生物科学与医学工程学院) State Key Laboratory of Software Development Environment(软件开发环境国家重点实验室) Key Laboratory of Biomechanics and Mechanobiology of Ministry of Education(教育部生物力学与机械生物学重点实验室) Beijing Advanced Innovation Center for Biomedical Engineering(北京生物医学工程先进创新中心) Beihang University(北京航空航天大学) Department of Biomedical Engineering(生物医学工程系) Tsinghua University(清华大学) ByteDance Inc.(字节跳动公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 11pages, 3figures, MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03140 2025-07-29 cs.CV 79%

Model Reveals What to Cache: Profiling-Based Feature Reuse for Video Diffusion Models

Xuran Ma, Yexin Liu, Yaofu Liu, Xianfeng Wu, Mingzhe Zheng, Zihao Wang, Ser-Nam Lim, Harry Yang

机构 * Hong Kong University of Science and Technology(香港科技大学) Everlyn AI University of Central Florida(佛罗里达中央大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08737 2025-07-29 cs.CV cs.AI 79%

Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models

In Cho, Youngbeom Yoo, Subin Jeon, Seon Joo Kim

机构 * Yonsei University(延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06201 2025-07-29 cs.CV cs.AI cs.CL 79%

Explainable Synthetic Image Detection through Diffusion Timestep Ensembling

Yixin Wu, Feiran Zhang, Tianyuan Shi, Ruicheng Yin, Zhenghua Wang, Zhenliang Gan, Xiaohua Wang, Changze Lv, Xiaoqing Zheng, Xuanjing Huang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 16 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06334 2025-07-29 cs.CV 79%

TriDi: Trilateral Diffusion of 3D Humans, Objects, and Interactions

Ilya A. Petrov, Riccardo Marin, Julian Chibane, Gerard Pons-Moll

机构 * University of Tübingen, Germany(图宾根大学) Tübingen AI Center, Germany(图宾根人工智能中心) Technical University of Munich, Germany(慕尼黑技术大学) Munich Center for Machine Learning, Germany(慕尼黑机器学习中心) Max Planck Institute for Informatics, Saarland Informatics Campus, Germany(马克斯·普朗克信息研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 2025 IEEE/CVF International Conference on Computer Vision (ICCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00623 2025-07-29 cs.CV 79%

A Lesson in Splats: Teacher-Guided Diffusion for 3D Gaussian Splats Generation with 2D Supervision

Chensheng Peng, Ido Sobol, Masayoshi Tomizuka, Kurt Keutzer, Chenfeng Xu, Or Litany

机构 * UC Berkeley(伯克利大学) Technion(技术学院) NVIDIA(英伟达)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04004 2025-07-29 eess.IV cs.CV 79%

Synomaly Noise and Multi-Stage Diffusion: A Novel Approach for Unsupervised Anomaly Detection in Medical Images

Yuan Bi, Lucie Huang, Ricarda Clarenbach, Reza Ghotbi, Angelos Karlas, Nassir Navab, Zhongliang Jiang

机构 * Computer Aided Medical Procedures, Technical University of Munich, Munich, Germany(技术大学慕尼黑计算机辅助医疗程序) Munich Center of Machine Learning, Munich, Germany(慕尼黑机器学习中心) Department for Vascular and Endovascular Surgery, rechts der Isar University Hospital, Technical University of Munich, Munich, Germany(技术大学慕尼黑血管和内血管外科部门)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19789 2025-07-29 cs.CV 79%

TransFlow: Motion Knowledge Transfer from Video Diffusion Models to Video Salient Object Detection

Suhwan Cho, Minhyeok Lee, Jungho Lee, Sunghun Yang, Sangyoun Lee

机构 * GenGenAI Yonsei University(延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCVW 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19770 2025-07-29 cs.CV 79%

MoFRR: Mixture of Diffusion Models for Face Retouching Restoration

Jiaxin Liu, Qichao Ying, Zhenxing Qian, Sheng Li, Runqi Zhang, Jian Liu, Xinpeng Zhang

机构 * Fudan University(复旦大学) Ant Group(蚂蚁集团)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20635 2025-07-29 cond-mat.mtrl-sci 78%

High-fidelity modeling of interface crossing in the diffusion welding process at the polycrystalline scale

Camille Godinot, Emmanuel Rigal, Frédéric Bernard, Philippe Emonot, Pierre-Eric Frayssines, Luc Védie, Marc Bernacki

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20602 2025-07-29 math.AP 78%

Deriving sub-diffusion equations

Benoît Perthame, Min Tang

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20524 2025-07-29 cs.NI 78%

A Lyapunov-Guided Diffusion-Based Reinforcement Learning Approach for UAV-Assisted Vehicular Networks with Delayed CSI Feedback

Zhang Liu, Lianfen Huang, Zhibin Gao, Xianbin Wang, Dusit Niyato, Xuemin, Shen

专题命中 扩散模型 :diffusion(title,abstract)

Comments 13 pages, 11 figures, transactions paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20477 2025-07-29 cs.IT eess.SP math.IT 78%

Rethinking Multi-User Communication in Semantic Domain: Enhanced OMDMA by Shuffle-Based Orthogonalization and Diffusion Denoising

Maojun Zhang, Guangxu Zhu, Xiaoming Chen, Kaibin Huang, Zhaoyang Zhang

专题命中 扩散模型 :diffusion(title,abstract)

Comments 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19882 2025-07-29 cs.AI 78%

Causality-aligned Prompt Learning via Diffusion-based Counterfactual Generation

Xinshu Li, Ruoyu Wang, Erdun Gao, Mingming Gong, Lina Yao

机构 * The University of New South Wales(新南威尔士大学) The University of Adelaide(阿德莱德大学) The University of Melbourne(墨尔本大学) CSIRO’s Data 61(CSIRO数据61)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10774 2025-07-29 math.NA cs.NA 78%

Isoparametric finite element methods for mean curvature flow and surface diffusion

Harald Garcke, Robert Nürnberg, Simon Praetorius, Ganghui Zhang

专题命中 扩散模型 :diffusion(title,abstract)

Comments 29 pages, 19 figures

Journal ref J. Comput. Phys. 539 (2025) 114248

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07815 2025-07-29 q-bio.BM cs.LG 78%

Mask prior-guided denoising diffusion improves inverse protein folding

Peizhen Bai, Filip Miljković, Xianyuan Liu, Leonardo De Maria, Rebecca Croasdale-Wood, Owen Rackham, Haiping Lu

机构 * School of Computer Science, University of Sheffield(谢菲尔德大学计算机科学学院) Biologics Engineering, Oncology R&D, AstraZeneca(阿斯利康生物制药工程与肿瘤研发部) Medicinal Chemistry, Research and Early Development, Cardiovascular, Renal and Metabolism, BioPharmaceuticals R&D, AstraZeneca(阿斯利康医药化学与研发部) School of Biological Sciences, University of Southampton(南安普顿大学生物科学学院) Centre for Machine Intelligence, University of Sheffield(谢菲尔德大学智能中心)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏