arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86872 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2503.06923 2025-08-12 cs.CV cs.AI 80%

From Reusing to Forecasting: Accelerating Diffusion Models with TaylorSeers

Jiacheng Liu, Chang Zou, Yuanhuiyi Lyu, Junjie Chen, Linfeng Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Shandong University(山东大学) University of Electronic Science and Technology of China(电子科技大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 14 figures; Accepted by ICCV2025; Mainly focus on feature caching for diffusion transformers acceleration

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18260 2025-07-25 cs.CV cs.AI 80%

Exploiting Gaussian Agnostic Representation Learning with Diffusion Priors for Enhanced Infrared Small Target Detection

Junyao Li, Yahao Lu, Xingyuan Guo, Xiaoyu Xian, Tiantian Wang, Yukai Shi

机构 * School of Information Engineering, Guangdong University of Technology, Guangzhou, 510006, China(广东技术大学信息工程学院) Guangzhou National Laboratory, Guangzhou 510006, China(广州国家实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Submitted to Neural Networks. We propose the Gaussian Group Squeezer, leveraging Gaussian sampling and compression with diffusion models for channel-based data augmentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09733 2025-07-15 cs.LG cs.AI cs.CV 80%

Universal Physics Simulation: A Foundational Diffusion Approach

Bradley Camburn

机构 * Singapore University of Technology and Design(新加坡科技设计大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 3 figures. Foundational AI model for universal physics simulation using sketch-guided diffusion transformers. Achieves SSIM > 0.8 on electromagnetic field generation without requiring a priori physics encoding

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03293 2025-06-05 cs.RO cs.CV 80%

Diffusion-VLA: Generalizable and Interpretable Robot Foundation Model via Self-Generated Reasoning

Junjie Wen, Minjie Zhu, Yichen Zhu, Zhibin Tang, Jinming Li, Zhongyi Zhou, Chengmeng Li, Xiaoyu Liu, Yaxin Peng, Chaomin Shen, Feifei Feng

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICML 2025. The project page is available at: http://diffusion-vla.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09185 2025-05-27 eess.IV cs.CV 80%

Cancer-Net PCa-Seg: Benchmarking Deep Learning Models for Prostate Cancer Segmentation Using Synthetic Correlated Diffusion Imaging

Jarett Dewbury, Chi-en Amy Tai, Alexander Wong

机构 * Vision and Image Processing Group, Systems Design Engineering, University of Waterloo(滑动与图像处理组,系统设计工程,滑铁卢大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 8 pages, 2 figures, to be published in Studies in Computational Intelligence. This paper introduces Cancer-Net PCa-Seg, a comprehensive evaluation of deep learning models for prostate cancer segmentation using synthetic correlated diffusion imaging (CDI$^s$). We benchmark five state-of-the-art architectures: U-Net, SegResNet, Swin UNETR, Attention U-Net, and LightM-UNet

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02156 2025-05-13 cs.CV cs.AI 80%

Latent Feature-Guided Diffusion Models for Shadow Removal

Kangfu Mei, Luis Figueroa, Zhe Lin, Zhihong Ding, Scott Cohen, Vishal M. Patel

机构 * Johns Hopkins University(约翰霍普金斯大学) Adobe Research(Adobe研究)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments project page see https://kfmei.com/shadow-diffusion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11838 2025-04-30 cs.CV 80%

High-Resolution Frame Interpolation with Patch-based Cascaded Diffusion

Junhwa Hur, Charles Herrmann, Saurabh Saxena, Janne Kontkanen, Wei-Sheng Lai, Yichang Shih, Michael Rubinstein, David J. Fleet, Deqing Sun

机构 * University of Toronto(多伦多大学) Vector Institute(向量研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://hifi-diffusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20595 2025-03-27 cs.LG cs.CV stat.ML 80%

Diffusion Counterfactuals for Image Regressors

Trung Duc Ha, Sidney Bender

机构 * Technische Universität Berlin(柏林工业大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 24 Pages, 5 Figures, Accepted at 3rd World Conference on eXplainable Artificial Intelligence (xAI-2025), Code and reproduction instructions available on GitHub, see https://github.com/DevinTDHa/Diffusion-Counterfactuals-for-Image-Regressors

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09124 2025-03-13 cs.LG cs.CV 80%

AdvAD: Exploring Non-Parametric Diffusion for Imperceptible Adversarial Attacks

Jin Li, Ziqiang He, Anwei Luo, Jian-Fang Hu, Z. Jane Wang, Xiangui Kang

机构 * Sun Yat-Sen University(中山大学) University of British Columbia(不列颠哥伦比亚大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accept by NeurIPS 2024. Please cite this paper using the following format: J. Li, Z. He, A. Luo, J. Hu, Z. Wang, X. Kang*, "AdvAD: Exploring Non-Parametric Diffusion for Imperceptible Adversarial Attacks", the 38th Annual Conference on Neural Information Processing Systems (NeurIPS), Vancouver, Canada, Dec 9-15, 2024. Code: https://github.com/XianguiKang/AdvAD

Journal ref Advances in Neural Information Processing Systems, vol. 37, pp. 52323--52353, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.02316 2025-02-25 cs.LG cs.CV 80%

Your Diffusion Model is Secretly a Certifiably Robust Classifier

Huanran Chen, Yinpeng Dong, Shitong Shao, Zhongkai Hao, Xiao Yang, Hang Su, Jun Zhu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by NeurIPS 2024. Also named as "Diffusion Models are Certifiably Robust Classifiers"

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12146 2025-02-18 cs.CV 80%

Diffusion-Sharpening: Fine-tuning Diffusion Models with Denoising Trajectory Sharpening

Ye Tian, Ling Yang, Xinchen Zhang, Yunhai Tong, Mengdi Wang, Bin Cui

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Code: https://github.com/Gen-Verse/Diffusion-Sharpening

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00783 2025-01-19 cs.CV cs.AI 80%

Diffusion Models and Representation Learning: A Survey

Michael Fuest, Pingchuan Ma, Ming Gui, Johannes Schusterbauer, Vincent Tao Hu, Bjorn Ommer

机构 * Technical University of Munich(慕尼黑工业大学) LMU Munich(慕尼黑大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Github Repo: https://github.com/dongzhuoyao/Diffusion-Representation-Learning-Survey-Taxonomy

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.18137 2025-01-07 cs.CV cs.LG 80%

Advancing Super-Resolution in Neural Radiance Fields via Variational Diffusion Strategies

Shrey Vishen, Jatin Sarabu, Saurav Kumar, Chinmay Bharathulwar, Rithwick Lakshmanan, Vishnu Srinivas

机构 * Monta Vista High School(蒙塔维斯塔高中) Bellarmine College Preparatory(贝拉明学院预科学校) UI Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) John P Stevens High School(约翰·P·史蒂文斯高中) Pleasant Valley High School(普莱森特谷高中) Foothill High School(福希尔高中)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments All our code is available at https://github.com/shreyvish5678/Advancing-Super-Resolution-in-Neural-Radiance-Fields-via-Variational-Diffusion-Strategies

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13144 2025-01-03 cs.LG cs.CV 80%

Neural Network Diffusion

Kai Wang, Dongwen Tang, Boya Zeng, Yida Yin, Zhaopan Xu, Yukun Zhou, Zelin Zang, Trevor Darrell, Zhuang Liu, Yang You

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments We introduce a novel approach for parameter generation, named neural network parameter diffusion (\textbf{p-diff}), which employs a standard latent diffusion model to synthesize a new set of parameters

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04295 2024-12-17 cs.CV 80%

Everything to the Synthetic: Diffusion-driven Test-time Adaptation via Synthetic-Domain Alignment

Jiayi Guo, Junhao Zhao, Chaoqun Du, Yulin Wang, Chunjiang Ge, Zanlin Ni, Shiji Song, Humphrey Shi, Gao Huang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments GitHub: https://github.com/SHI-Labs/Diffusion-Driven-Test-Time-Adaptation-via-Synthetic-Domain-Alignment

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10294 2024-12-16 cs.CV 80%

Coherent 3D Scene Diffusion From a Single RGB Image

Manuel Dahnert, Angela Dai, Norman Müller, Matthias Nießner

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://www.manuel-dahnert.com/research/scene-diffusion - Accepted at NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01392 2024-12-11 cs.LG cs.CV cs.RO 80%

Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion

Boyuan Chen, Diego Marti Monso, Yilun Du, Max Simchowitz, Russ Tedrake, Vincent Sitzmann

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project website: https://boyuan.space/diffusion-forcing

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05355 2024-12-10 cs.CV cs.AI 80%

MotionShop: Zero-Shot Motion Transfer in Video Diffusion Models with Mixture of Score Guidance

Hidir Yesiltepe, Tuna Han Salih Meral, Connor Dunlop, Pinar Yanardag

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://motionshop-diffusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05275 2024-12-09 cs.CV cs.AI 80%

MotionFlow: Attention-Driven Motion Transfer in Video Diffusion Models

Tuna Han Salih Meral, Hidir Yesiltepe, Connor Dunlop, Pinar Yanardag

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://motionflow-diffusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.11243 2024-12-05 cs.CV cs.AI 80%

Multi-Sensor Diffusion-Driven Optical Image Translation for Large-Scale Applications

João Gabriel Vinholi, Marco Chini, Anis Amziane, Renato Machado, Danilo Silva, Patrick Matgen

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This is the accepted version of the manuscript published in IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing (JSTARS). Please access the final version at IEEEXplore (Open Access). DOI 10.1109/JSTARS.2024.3506032. This technology is protected by a patent filed on 23 december 2023 at Office Luxembourgeois de la propriété intellectuelle (LU505861)

Journal ref "Multi-Sensor Diffusion-Driven Optical Image Translation for Large-Scale Applications," in IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.05779 2024-11-21 cs.CV 80%

Erasing Undesirable Influence in Diffusion Models

Jing Wu, Trung Le, Munawar Hayat, Mehrtash Harandi

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Diffusion Model, Machine Unlearning

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.03150 2024-11-19 cs.CV cs.LG 80%

Video Diffusion Models: A Survey

Andrew Melnik, Michal Ljubljanac, Cong Lu, Qi Yan, Weiming Ren, Helge Ritter

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments https://github.com/ndrwmlnk/Awesome-Video-Diffusion-Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11473 2024-11-05 cs.CV cs.AI 80%

FIFO-Diffusion: Generating Infinite Videos from Text without Training

Jihwan Kim, Junoh Kang, Jinyoung Choi, Bohyung Han

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://jjihwan.github.io/projects/FIFO-Diffusion

Journal ref NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03954 2024-09-30 cs.RO cs.CV cs.LG 80%

3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations

Yanjie Ze, Gu Zhang, Kangning Zhang, Chenyuan Hu, Muhan Wang, Huazhe Xu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Published at Robotics: Science and Systems (RSS) 2024. Videos, code, and data: https://3d-diffusion-policy.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10542 2024-09-27 cs.AR cs.CV 80%

SF-MMCN: Low-Power Sever Flow Multi-Mode Diffusion Model Accelerator

Huan-Ke Hsu, I-Chyn Wey, T. Hui Teo

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 16 pages, 16 figures; extend the CNN to process Diffusion Model (possible this is the first reported hardware Diffusion Model implementation)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07186 2024-09-17 cs.CV cs.AI 80%

Enhancing Angular Resolution via Directionality Encoding and Geometric Constraints in Brain Diffusion Tensor Imaging

Sheng Chen, Zihao Tang, Mariano Cabezas, Xinyi Wang, Arkiev D'Souza, Michael Barnett, Fernando Calamante, Weidong Cai, Chenyu Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICONIP2024, Diffusion Weighted Imaging, Diffusion Tensor Imaging, Angular Resolution Enhancement, Fractional Anisotropy

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.10236 2024-08-21 eess.IV cs.CV 80%

AID-DTI: Accelerating High-fidelity Diffusion Tensor Imaging with Detail-preserving Model-based Deep Learning

Wenxin Fan, Jian Cheng, Cheng Li, Jing Yang, Ruoyou Wu, Juan Zou, Shanshan Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 12 pages, 3 figures, MICCAI 2024 Workshop on Computational Diffusion MRI. arXiv admin note: text overlap with arXiv:2401.01693, arXiv:2405.03159

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.14134 2024-08-12 cs.LG cs.CV cs.RO 80%

Diffusion Reward: Learning Rewards via Conditional Video Diffusion

Tao Huang, Guangqi Jiang, Yanjie Ze, Huazhe Xu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ECCV 2024. Project page and code: https://diffusion-reward.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.02226 2024-08-08 cs.CV 80%

ProCreate, Don't Reproduce! Propulsive Energy Diffusion for Creative Generation

Jack Lu, Ryan Teehan, Mengye Ren

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ECCV 2024. Project page: https://procreate-diffusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14837 2024-08-05 cs.CV 80%

Osmosis: RGBD Diffusion Prior for Underwater Image Restoration

Opher Bar Nathan, Deborah Levy, Tali Treibitz, Dan Rosenbaum

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ECCV 2024. Project page with results and code: https://osmosis-diffusion.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏