arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86714 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70159 篇

2410.03160 2024-10-07 cs.CV cs.LG 83%

Redefining Temporal Modeling in Video Diffusion: The Vectorized Timestep Approach

Yaofang Liu, Yumeng Ren, Xiaodong Cun, Aitor Artola, Yang Liu, Tieyong Zeng, Raymond H. Chan, Jean-michel Morel

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Code at https://github.com/Yaofang-Liu/FVDM

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.03958 2024-10-07 cs.CV cs.AI cs.LG 83%

Simple Drop-in LoRA Conditioning on Attention Layers Will Improve Your Diffusion Model

Joo Young Choi, Jaesung R. Park, Inkyu Park, Jaewoong Cho, Albert No, Ernest K. Ryu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.14830 2024-10-07 physics.med-ph cs.CV eess.IV physics.bio-ph q-bio.TO 83%

Dreaming of Electrical Waves: Generative Modeling of Cardiac Excitation Waves using Diffusion Models

Tanish Baranwal, Jan Lebert, Jan Christoph

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.13042 2024-10-07 cs.CV cs.AI cs.LG 83%

MosaicFusion: Diffusion Models as Data Augmenters for Large Vocabulary Instance Segmentation

Jiahao Xie, Wei Li, Xiangtai Li, Ziwei Liu, Yew Soon Ong, Chen Change Loy

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments International Journal of Computer Vision (IJCV), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14857 2024-10-03 cs.CV cs.AI cs.LG 83%

Conditional Diffusion on Web-Scale Image Pairs leads to Diverse Image Variations

Manoj Kumar, Neil Houlsby, Emiel Hoogeboom

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02148 2024-10-03 cs.CV 83%

Diffusion$^2$: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models

Zeyu Yang, Zijie Pan, Chun Gu, Li Zhang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00760 2024-10-02 cs.CV cs.AI cs.LG 83%

Smoothed Energy Guidance: Guiding Diffusion Models with Reduced Energy Curvature of Attention

Susung Hong

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.00265 2024-10-02 eess.IV cs.CV 83%

Adaptive Latent Diffusion Model for 3D Medical Image to Image Translation: Multi-modal Magnetic Resonance Imaging Study

Jonghun Kim, Hyunjin Park

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments 8 pages, 7 figures, WACV 2024 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19589 2024-10-01 cs.CV 83%

Effective Diffusion Transformer Architecture for Image Super-Resolution

Kun Cheng, Lei Yu, Zhijun Tu, Xiao He, Liyu Chen, Yong Guo, Mingrui Zhu, Nannan Wang, Xinbo Gao, Jie Hu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Code is available at https://github.com/kunncheng/DiT-SR

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.18025 2024-10-01 cs.CV cs.AI 83%

Where's Waldo: Diffusion Features for Personalized Segmentation and Retrieval

Dvir Samuel, Rami Ben-Ari, Matan Levy, Nir Darshan, Gal Chechik

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted to NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06098 2024-10-01 cs.CV cs.CL 83%

VidProM: A Million-scale Real Prompt-Gallery Dataset for Text-to-Video Diffusion Models

Wenhao Wang, Yi Yang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted by NeurIPS 2024 (Datasets and Benchmarks Track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.09316 2024-10-01 cs.CV 83%

Diffusion Models for Open-Vocabulary Segmentation

Laurynas Karazija, Iro Laina, Andrea Vedaldi, Christian Rupprecht

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments ECCV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15997 2024-09-30 cs.CV cs.AI cs.LG 83%

Improvements to SDXL in NovelAI Diffusion V3

Juan Ossa, Eren Doğan, Alex Birch, F. Johnson

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 14 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00094 2024-09-30 cs.CV cs.AI 83%

Fast ODE-based Sampling for Diffusion Models in Around 5 Steps

Zhenyu Zhou, Defang Chen, Can Wang, Chun Chen

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted by CVPR 2024 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17740 2024-09-27 cs.CV 83%

AnyLogo: Symbiotic Subject-Driven Diffusion System with Gemini Status

Jinghao Zhang, Wen Qian, Hao Luo, Fan Wang, Feng Zhao

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 13 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17565 2024-09-27 cs.CV cs.AI cs.LG 83%

Pixel-Space Post-Training of Latent Diffusion Models

Christina Zhang, Simran Motwani, Matthew Yu, Ji Hou, Felix Juefei-Xu, Sam Tsai, Peter Vajda, Zijian He, Jialiang Wang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17487 2024-09-27 cs.CV 83%

Learning Quantized Adaptive Conditions for Diffusion Models

Yuchen Liang, Yuchuan Tian, Lei Yu, Huao Tang, Jie Hu, Xiangzhong Fang, Hanting Chen

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16535 2024-09-26 cs.CV 83%

Prompt Sliders for Fine-Grained Control, Editing and Erasing of Concepts in Diffusion Models

Deepak Sridhar, Nuno Vasconcelos

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments ECCV'24 - Unlearning and Model Editing Workshop. Code: https://github.com/DeepakSridhar/promptsliders

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16280 2024-09-25 cs.CV 83%

MonoFormer: One Transformer for Both Diffusion and Autoregression

Chuyang Zhao, Yuxing Song, Wenhao Wang, Haocheng Feng, Errui Ding, Yifan Sun, Xinyan Xiao, Jingdong Wang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.11415 2024-09-24 cs.CV cs.AI cs.LG 83%

DreamSampler: Unifying Diffusion Sampling and Score Distillation for Image Manipulation

Jeongsol Kim, Geon Yeong Park, Jong Chul Ye

专题命中 扩散模型 :diffusion(title,abstract);image editing(abstract);分类 cs.CV

Comments ECCV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12539 2024-09-20 cs.CV 83%

Improving Cone-Beam CT Image Quality with Knowledge Distillation-Enhanced Diffusion Model in Imbalanced Data Settings

Joonil Hwang, Sangjoon Park, NaHyeon Park, Seungryong Cho, Jin Sung Kim

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments MICCAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.19140 2024-09-19 cs.CV cs.AI 83%

QNCD: Quantization Noise Correction for Diffusion Models

Huanpeng Chu, Wei Wu, Chengjie Zang, Kun Yuan

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted by ACMMM2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11831 2024-09-19 cs.RO cs.CV cs.LG 83%

RaggeDi: Diffusion-based State Estimation of Disordered Rags, Sheets, Towels and Blankets

Jikai Ye, Wanze Li, Shiraz Khan, Gregory S. Chirikjian

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11689 2024-09-19 cs.CV cs.AI 83%

GUNet: A Graph Convolutional Network United Diffusion Model for Stable and Diversity Pose Generation

Shuowen Liang, Sisi Li, Qingyun Wang, Cen Zhang, Kaiquan Zhu, Tian Yang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09569 2024-09-17 cs.LG cs.CV cs.CY 83%

Bias Begets Bias: The Impact of Biased Embeddings on Diffusion Models

Sahil Kuchlous, Marvin Li, Jeffrey G. Wang

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 19 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09144 2024-09-17 cs.CV 83%

PrimeDepth: Efficient Monocular Depth Estimation with a Stable Diffusion Preimage

Denis Zavadski, Damjan Kalšan, Carsten Rother

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.10647 2024-09-17 cs.CV cs.AI cs.LG 83%

A Survey on Video Diffusion Models

Zhen Xing, Qijun Feng, Haoran Chen, Qi Dai, Han Hu, Hang Xu, Zuxuan Wu, Yu-Gang Jiang

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08278 2024-09-13 cs.CV 83%

DreamHOI: Subject-Driven Generation of 3D Human-Object Interactions with Diffusion Priors

Thomas Hanwen Zhu, Ruining Li, Tomas Jakab

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Project page: https://DreamHOI.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.15288 2024-09-13 cs.CV cs.LG 83%

Memory-Efficient 3D Denoising Diffusion Models for Medical Image Processing

Florentin Bieder, Julia Wolleb, Alicia Durrer, Robin Sandkühler, Philippe C. Cattin

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted at MIDL 2023

Journal ref Medical Imaging with Deep Learning, PMLR 227:552-567, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07269 2024-09-12 cs.CV 83%

Realistic and Efficient Face Swapping: A Unified Approach with Diffusion Models

Sanoojan Baliah, Qinliang Lin, Shengcai Liao, Xiaodan Liang, Muhammad Haris Khan

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted as a conference paper at WACV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏