arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-07-29 至 2025-07-29 共收录 55 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 55 篇

2502.01189 2025-07-29 eess.IV cs.AI cs.CV cs.IT eess.SP math.IT 88%

Compressed Image Generation with Denoising Diffusion Codebook Models

Guy Ohayon, Hila Manor, Tomer Michaeli, Michael Elad

机构 * Faculty of Computer Science, Technion -- Israel Institute of Technology, Haifa, Israel(计算机科学系,技术离子理工学院——以色列理工学院,海法,以色列) Faculty of Electrical and Computer Engineering, Technion -- Israel Institute of Technology, Haifa, Israel(电气与计算机工程系,技术离子理工学院——以色列理工学院,海法,以色列)

专题命中 扩散模型 :image generation(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Published in the International Conference on Machine Learning (ICML) 2025. Code and demo are available at https://ddcm-2025.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20478 2025-07-29 cs.LG 88%

Conditional Diffusion Models for Global Precipitation Map Inpainting

Daiko Kishikawa, Yuka Muto, Shunji Kotsuki

专题命中 扩散模型 :diffusion(title,abstract);inpainting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17226 2025-07-29 cs.CV 83%

DDB: Diffusion Driven Balancing to Address Spurious Correlations

Aryan Yazdan Parast, Basim Azam, Naveed Akhtar

机构 * The University of Melbourne(墨尔本大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20363 2025-07-29 cs.CV 79%

Generative Pre-training for Subjective Tasks: A Diffusion Transformer-Based Framework for Facial Beauty Prediction

Djamel Eddine Boukhari, Ali chemsa

机构 * LGEERE Laboratory(LGEERE实验室) Department of Electrical Engineering,University of El Oued(电气工程系,埃尔欧德大学) Scientific and Technical Research Centre for Arid Areas(干旱地区科学技术研究中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20158 2025-07-29 cs.CV 79%

AnimeColor: Reference-based Animation Colorization with Diffusion Transformers

Yuhong Zhang, Liyao Wang, Han Wang, Danni Wu, Zuzeng Lin, Feng Wang, Li Song

机构 * Shanghai Jiao Tong University(上海交通大学) Tianjin University(天津大学) Communication University of China(中国通信大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19970 2025-07-29 eess.IV cs.CV cs.LG 79%

SkinDualGen: Prompt-Driven Diffusion for Simultaneous Image-Mask Generation in Skin Lesions

Zhaobin Xu

机构 * Shandong University, Qingdao, China(山东大学,青岛)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19874 2025-07-29 cs.CV 79%

All-in-One Medical Image Restoration with Latent Diffusion-Enhanced Vector-Quantized Codebook Prior

Haowei Chen, Zhiwen Yang, Haotian Hou, Hui Zhang, Bingzheng Wei, Gang Zhou, Yan Xu

机构 * School of Biological Science and Medical Engineering(生物科学与医学工程学院) State Key Laboratory of Software Development Environment(软件开发环境国家重点实验室) Key Laboratory of Biomechanics and Mechanobiology of Ministry of Education(教育部生物力学与机械生物学重点实验室) Beijing Advanced Innovation Center for Biomedical Engineering(北京生物医学工程先进创新中心) Beihang University(北京航空航天大学) Department of Biomedical Engineering(生物医学工程系) Tsinghua University(清华大学) ByteDance Inc.(字节跳动公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 11pages, 3figures, MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03140 2025-07-29 cs.CV 79%

Model Reveals What to Cache: Profiling-Based Feature Reuse for Video Diffusion Models

Xuran Ma, Yexin Liu, Yaofu Liu, Xianfeng Wu, Mingzhe Zheng, Zihao Wang, Ser-Nam Lim, Harry Yang

机构 * Hong Kong University of Science and Technology(香港科技大学) Everlyn AI University of Central Florida(佛罗里达中央大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08737 2025-07-29 cs.CV cs.AI 79%

Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models

In Cho, Youngbeom Yoo, Subin Jeon, Seon Joo Kim

机构 * Yonsei University(延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06201 2025-07-29 cs.CV cs.AI cs.CL 79%

Explainable Synthetic Image Detection through Diffusion Timestep Ensembling

Yixin Wu, Feiran Zhang, Tianyuan Shi, Ruicheng Yin, Zhenghua Wang, Zhenliang Gan, Xiaohua Wang, Changze Lv, Xiaoqing Zheng, Xuanjing Huang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 16 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06334 2025-07-29 cs.CV 79%

TriDi: Trilateral Diffusion of 3D Humans, Objects, and Interactions

Ilya A. Petrov, Riccardo Marin, Julian Chibane, Gerard Pons-Moll

机构 * University of Tübingen, Germany(图宾根大学) Tübingen AI Center, Germany(图宾根人工智能中心) Technical University of Munich, Germany(慕尼黑技术大学) Munich Center for Machine Learning, Germany(慕尼黑机器学习中心) Max Planck Institute for Informatics, Saarland Informatics Campus, Germany(马克斯·普朗克信息研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 2025 IEEE/CVF International Conference on Computer Vision (ICCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00623 2025-07-29 cs.CV 79%

A Lesson in Splats: Teacher-Guided Diffusion for 3D Gaussian Splats Generation with 2D Supervision

Chensheng Peng, Ido Sobol, Masayoshi Tomizuka, Kurt Keutzer, Chenfeng Xu, Or Litany

机构 * UC Berkeley(伯克利大学) Technion(技术学院) NVIDIA(英伟达)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04004 2025-07-29 eess.IV cs.CV 79%

Synomaly Noise and Multi-Stage Diffusion: A Novel Approach for Unsupervised Anomaly Detection in Medical Images

Yuan Bi, Lucie Huang, Ricarda Clarenbach, Reza Ghotbi, Angelos Karlas, Nassir Navab, Zhongliang Jiang

机构 * Computer Aided Medical Procedures, Technical University of Munich, Munich, Germany(技术大学慕尼黑计算机辅助医疗程序) Munich Center of Machine Learning, Munich, Germany(慕尼黑机器学习中心) Department for Vascular and Endovascular Surgery, rechts der Isar University Hospital, Technical University of Munich, Munich, Germany(技术大学慕尼黑血管和内血管外科部门)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19789 2025-07-29 cs.CV 79%

TransFlow: Motion Knowledge Transfer from Video Diffusion Models to Video Salient Object Detection

Suhwan Cho, Minhyeok Lee, Jungho Lee, Sunghun Yang, Sangyoun Lee

机构 * GenGenAI Yonsei University(延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCVW 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19770 2025-07-29 cs.CV 79%

MoFRR: Mixture of Diffusion Models for Face Retouching Restoration

Jiaxin Liu, Qichao Ying, Zhenxing Qian, Sheng Li, Runqi Zhang, Jian Liu, Xinpeng Zhang

机构 * Fudan University(复旦大学) Ant Group(蚂蚁集团)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20635 2025-07-29 cond-mat.mtrl-sci 78%

High-fidelity modeling of interface crossing in the diffusion welding process at the polycrystalline scale

Camille Godinot, Emmanuel Rigal, Frédéric Bernard, Philippe Emonot, Pierre-Eric Frayssines, Luc Védie, Marc Bernacki

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20602 2025-07-29 math.AP 78%

Deriving sub-diffusion equations

Benoît Perthame, Min Tang

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20524 2025-07-29 cs.NI 78%

A Lyapunov-Guided Diffusion-Based Reinforcement Learning Approach for UAV-Assisted Vehicular Networks with Delayed CSI Feedback

Zhang Liu, Lianfen Huang, Zhibin Gao, Xianbin Wang, Dusit Niyato, Xuemin, Shen

专题命中 扩散模型 :diffusion(title,abstract)

Comments 13 pages, 11 figures, transactions paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20477 2025-07-29 cs.IT eess.SP math.IT 78%

Rethinking Multi-User Communication in Semantic Domain: Enhanced OMDMA by Shuffle-Based Orthogonalization and Diffusion Denoising

Maojun Zhang, Guangxu Zhu, Xiaoming Chen, Kaibin Huang, Zhaoyang Zhang

专题命中 扩散模型 :diffusion(title,abstract)

Comments 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19882 2025-07-29 cs.AI 78%

Causality-aligned Prompt Learning via Diffusion-based Counterfactual Generation

Xinshu Li, Ruoyu Wang, Erdun Gao, Mingming Gong, Lina Yao

机构 * The University of New South Wales(新南威尔士大学) The University of Adelaide(阿德莱德大学) The University of Melbourne(墨尔本大学) CSIRO’s Data 61(CSIRO数据61)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10774 2025-07-29 math.NA cs.NA 78%

Isoparametric finite element methods for mean curvature flow and surface diffusion

Harald Garcke, Robert Nürnberg, Simon Praetorius, Ganghui Zhang

专题命中 扩散模型 :diffusion(title,abstract)

Comments 29 pages, 19 figures

Journal ref J. Comput. Phys. 539 (2025) 114248

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07815 2025-07-29 q-bio.BM cs.LG 78%

Mask prior-guided denoising diffusion improves inverse protein folding

Peizhen Bai, Filip Miljković, Xianyuan Liu, Leonardo De Maria, Rebecca Croasdale-Wood, Owen Rackham, Haiping Lu

机构 * School of Computer Science, University of Sheffield(谢菲尔德大学计算机科学学院) Biologics Engineering, Oncology R&D, AstraZeneca(阿斯利康生物制药工程与肿瘤研发部) Medicinal Chemistry, Research and Early Development, Cardiovascular, Renal and Metabolism, BioPharmaceuticals R&D, AstraZeneca(阿斯利康医药化学与研发部) School of Biological Sciences, University of Southampton(南安普顿大学生物科学学院) Centre for Machine Intelligence, University of Sheffield(谢菲尔德大学智能中心)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.09336 2025-07-29 cs.LG 78%

Compositional Abilities Emerge Multiplicatively: Exploring Diffusion Models on a Synthetic Task

Maya Okawa, Ekdeep Singh Lubana, Robert P. Dick, Hidenori Tanaka

机构 * Center for Brain Science, Harvard University(哈佛大学脑科学中心) Physics & Informatics Laboratories, NTT Research, Inc.(NTT研究所物理与信息实验室) EECS Department, University of Michigan(密歇根大学电子工程与计算机科学系)

专题命中 扩散模型 :diffusion(title,abstract)

Comments 37th Conference on Neural Information Processing Systems (NeurIPS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11441 2025-07-29 cs.CV cs.LG 70%

Implementing Adaptations for Vision AutoRegressive Model

Kaif Shaikh, Franziska Boenisch, Adam Dziedzic

机构 * cispa(CISPA信息安全研究中心)

专题命中 扩散模型 :image generation(abstract);diffusion(abstract);分类 cs.CV

Comments Accepted at DIG-BUGS: Data in Generative Models Workshop @ ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17860 2025-07-29 cs.GR cs.CV cs.LG 62%

Multi-Person Interaction Generation from Two-Person Motion Priors

Wenning Xu, Shiyu Fan, Paul Henderson, Edmond S. L. Ho

机构 * University of Glasgow(格拉斯哥大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV、cs.GR

Comments SIGGRAPH 2025 Conference Papers, project page at http://wenningxu.github.io/multicharacter/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20976 2025-07-29 cs.CV 57%

Adapting Vehicle Detectors for Aerial Imagery to Unseen Domains with Weak Supervision

Xiao Fang, Minhyek Jeon, Zheyang Qin, Stanislav Panev, Celso de Melo, Shuowen Hu, Shayok Chakraborty, Fernando De la Torre

机构 * Carnegie Mellon University(卡内基梅隆大学) DEVCOM Army Research Laboratory(陆军研究实验室) Florida State University(佛罗里达州立大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20239 2025-07-29 cs.CV 57%

Decomposing Densification in Gaussian Splatting for Faster 3D Scene Reconstruction

Binxiao Huang, Zhengwu Liu, Ngai Wong

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19835 2025-07-29 cs.SD cs.MM 57%

SonicGauss: Position-Aware Physical Sound Synthesis for 3D Gaussian Representations

Chunshi Wang, Hongxing Li, Yawei Luo

机构 * Zhejiang University(浙江大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.MM

Comments Accepted by ACMMM'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08685 2025-07-29 cs.CV 57%

"Principal Components" Enable A New Language of Images

Xin Wen, Bingchen Zhao, Ismail Elezi, Jiankang Deng, Xiaojuan Qi

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09822 2025-07-29 cs.CV 57%

Dynamic Try-On: Taming Video Virtual Try-on with Dynamic Attention Mechanism

Jun Zheng, Jing Wang, Fuwei Zhao, Xujie Zhang, Xiaodan Liang

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Project Page: https://zhengjun-ai.github.io/dynamic-tryon-page/. Accepted by The 36th British Machine Vision Conference

详情

展开后加载摘要…

URL PDF HTML 收藏