arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86714 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70159 篇

2503.17074 2025-03-25 cs.CV 83%

Zero-Shot Styled Text Image Generation, but Make It Autoregressive

Vittorio Pippi, Fabio Quattrini, Silvia Cascianelli, Alessio Tonioni, Rita Cucchiara

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted at CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18134 2025-03-25 cs.CV 83%

An Image-like Diffusion Method for Human-Object Interaction Detection

Xiaofei Hui, Haoxuan Qu, Hossein Rahmani, Jun Liu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13401 2025-03-25 cs.CV 83%

Zero-Shot Low Light Image Enhancement with Diffusion Prior

Joshua Cho, Sara Aghajanzadeh, Zhen Zhu, D. A. Forsyth

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.19324 2025-03-25 cs.CV cs.LG stat.ML 83%

Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion

Emiel Hoogeboom, Thomas Mensink, Jonathan Heek, Kay Lamerigts, Ruiqi Gao, Tim Salimans

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.05399 2025-03-25 cs.CV cs.LG 83%

Sequential Posterior Sampling with Diffusion Models

Tristan S. W. Stevens, Oisín Nolan, Jean-Luc Robert, Ruud J. G. van Sloun

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments 5 pages, 4 figures, preprint

Journal ref 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00684 2025-03-25 cs.CV cs.CL 83%

Deciphering Oracle Bone Language with Diffusion Models

Haisu Guan, Huanxin Yang, Xinyu Wang, Shengwei Han, Yongge Liu, Lianwen Jin, Xiang Bai, Yuliang Liu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments ACL 2024 Best Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.01274 2025-03-25 cs.CV 83%

Diffusion Models with Deterministic Normalizing Flow Priors

Mohsen Zand, Ali Etemad, Michael Greenspan

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 17 pages, 7 figures, Published in Transactions on Machine Learning Research (TMLR)

Journal ref https://openreview.net/pdf?id=ACMNVwcR6v, Transactions on Machine Learning Research (TMLR), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17155 2025-03-24 cs.CV 83%

D2C: Unlocking the Potential of Continuous Autoregressive Image Generation with Discrete Tokens

Panpan Wang, Liqiang Niu, Fandong Meng, Jinan Xu, Yufeng Chen, Jie Zhou

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16921 2025-03-24 cs.CV cs.AI 83%

When Preferences Diverge: Aligning Diffusion Models with Minority-Aware Adaptive DPO

Lingfan Zhang, Chen Liu, Chengming Xu, Kai Hu, Donghao Luo, Chengjie Wang, Yanwei Fu, Yuan Yao

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07001 2025-03-21 cs.CV cs.AI cs.LG 83%

From Image to Video: An Empirical Study of Diffusion Representations

Pedro Vélez, Luisa F. Polanía, Yi Yang, Chuhan Zhang, Rishabh Kabra, Anurag Arnab, Mehdi S. M. Sajjadi

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14358 2025-03-19 cs.CV cs.LG 83%

RFMI: Estimating Mutual Information on Rectified Flow for Text-to-Image Alignment

Chao Wang, Giulio Franzese, Alessandro Finamore, Pietro Michiardi

专题命中 扩散模型 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV

Comments to appear at ICLR 2025 Workshop on Deep Generative Model in Machine Learning: Theory, Principle and Efficacy

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13945 2025-03-19 cs.CV 83%

Make the Most of Everything: Further Considerations on Disrupting Diffusion-based Customization

Long Tang, Dengpan Ye, Sirun Chen, Xiuwen Shi, Yunna Lv, Ziyi Liu

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.13661 2025-03-19 cs.SE cs.AI cs.CV cs.LG 83%

Efficient Domain Augmentation for Autonomous Driving Testing Using Diffusion Models

Luciano Baresi, Davide Yi Xian Hu, Andrea Stocco, Paolo Tonella

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted for publication at the 47th International Conference on Software Engineering (ICSE 2025). This research was partially supported by project EMELIOT, funded by MUR under the PRIN 2020 program (n. 2020W3A5FY), by the Bavarian Ministry of Economic Affairs, Regional Development and Energy, by the TUM Global Incentive Fund, and by the EU Project Sec4AI4Sec (n. 101120393)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.15304 2025-03-19 cs.LG cs.CV 83%

Unlearning Concepts in Diffusion Model via Concept Domain Correction and Concept Preserving Gradient

Yongliang Wu, Shiji Zhou, Mingzhuo Yang, Lianzhe Wang, Heng Chang, Wenbo Zhu, Xinting Hu, Xiao Zhou, Xu Yang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments AAAI 2025 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12472 2025-03-18 cs.CV 83%

Diffusion-based Synthetic Data Generation for Visible-Infrared Person Re-Identification

Wenbo Dai, Lijing Lu, Zhihang Li

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01794 2025-03-18 cs.CV cs.AI 83%

IQA-Adapter: Exploring Knowledge Transfer from Image Quality Assessment to Diffusion-based Generative Models

Khaled Abud, Sergey Lavrushkin, Alexey Kirillov, Dmitriy Vatolin

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments GitHub repo: https://github.com/X1716/IQA-Adapter

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10687 2025-03-17 cs.CV 83%

Context-guided Responsible Data Augmentation with Diffusion Models

Khawar Islam, Naveed Akhtar

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments ICLRw

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07026 2025-03-17 cs.CV cs.AI 83%

Erase Diffusion: Empowering Object Removal Through Calibrating Diffusion Pathways

Yi Liu, Hao Zhou, Wenxiang Shang, Ran Lin, Benlei Cui

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.06664 2025-03-17 cs.CV cs.AI 83%

Decouple-Then-Merge: Finetune Diffusion Models as Multi-Task Learning

Qianli Ma, Xuefei Ning, Dongrui Liu, Li Niu, Linfeng Zhang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted by CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.05116 2025-03-14 cs.LG cs.AI cs.CV cs.HC 83%

HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning

Ayano Hiranaka, Shang-Fu Chen, Chieh-Hsin Lai, Dongjun Kim, Naoki Murata, Takashi Shibuya, Wei-Hsiang Liao, Shao-Hua Sun, Yuki Mitsufuji

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Published in International Conference on Learning Representations (ICLR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09491 2025-03-13 cs.CV eess.IV 83%

DAMM-Diffusion: Learning Divergence-Aware Multi-Modal Diffusion Model for Nanoparticles Distribution Prediction

Junjie Zhou, Shouju Wang, Yuxia Tang, Qi Zhu, Daoqiang Zhang, Wei Shao

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16965 2025-03-13 cs.CV 83%

Autoregressive Image Generation with Vision Full-view Prompt

Miaomiao Cai, Guanjie Wang, Wei Li, Zhijun Tu, Hanting Chen, Shaohui Lin, Jie Hu

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08729 2025-03-13 cs.CV cs.AI cs.LG 83%

Preserving Product Fidelity in Large Scale Image Recontextualization with Diffusion Models

Ishaan Malhi, Praneet Dutta, Ellie Talius, Sally Ma, Brendan Driscoll, Krista Holden, Garima Pruthi, Arunachalam Narayanaswamy

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08339 2025-03-12 cs.CV 83%

Diffusion Transformer Meets Random Masks: An Advanced PET Reconstruction Framework

Bin Huang, Binzhong He, Yanhan Chen, Zhili Liu, Xinyue Wang, Binxuan Li, Qiegen Liu

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08253 2025-03-12 cs.CV 83%

SARA: Structural and Adversarial Representation Alignment for Training-efficient Diffusion Models

Hesen Chen, Junyan Wang, Zhiyu Tan, Hao Li

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Technical report

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12382 2025-03-12 cs.CV 83%

DiffDoctor: Diagnosing Image Diffusion Models Before Treating

Yiyang Wang, Xi Chen, Xiaogang Xu, Sihui Ji, Yu Liu, Yujun Shen, Hengshuang Zhao

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 8 pages of main body

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17017 2025-03-12 cs.CV 83%

TED-VITON: Transformer-Empowered Diffusion Models for Virtual Try-On

Zhenchen Wan, Yanwu Xu, Zhaoqing Wang, Feng Liu, Tongliang Liu, Mingming Gong

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Project page: https://github.com/ZhenchenWan/TED-VITON

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.21826 2025-03-12 cs.CV 83%

Volumetric Conditioning Module to Control Pretrained Diffusion Models for 3D Medical Images

Suhyun Ahn, Wonjung Park, Jihoon Cho, Seunghyuck Park, Jinah Park

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments 17 pages, 18 figures, accepted @ WACV 2025

Journal ref Proceedings of the Winter Conference on Applications of Computer Vision (WACV), pp. 85-95, Feb. 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07659 2025-03-12 cs.CV 83%

MotionAura: Generating High-Quality and Motion Consistent Videos using Discrete Diffusion

Onkar Susladkar, Jishu Sen Gupta, Chirag Sehgal, Sparsh Mittal, Rekha Singhal

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted in ICLR 2025 (spotlight paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06746 2025-03-11 cs.CV 83%

Color Alignment in Diffusion

Ka Chun Shum, Binh-Son Hua, Duc Thanh Nguyen, Sai-Kit Yeung

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏