arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70082 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70082 篇

2310.00434 2024-05-15 cs.CV cs.GR 81%

DiffPoseTalk: Speech-Driven Stylistic 3D Facial Animation and Head Pose Generation via Diffusion Models

Zhiyao Sun, Tian Lv, Sheng Ye, Matthieu Lin, Jenny Sheng, Yu-Hui Wen, Minjing Yu, Yong-Jin Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments SIGGRAPH 2024 (Journal Track). Project page: https://diffposetalk.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08529 2024-05-14 cs.CV cs.GR 81%

GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models

Taoran Yi, Jiemin Fang, Junjie Wang, Guanjun Wu, Lingxi Xie, Xiaopeng Zhang, Wenyu Liu, Qi Tian, Xinggang Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments CVPR 2024, Project page: https://taoranyi.com/gaussiandreamer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.06778 2024-05-14 cs.CV cs.GR 81%

Shape Conditioned Human Motion Generation with Diffusion Model

Kebing Xue, Hyewon Seo

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17834 2024-05-09 cs.CV cs.GR 81%

Spice-E : Structural Priors in 3D Diffusion using Cross-Entity Attention

Etai Sella, Gal Fiebelman, Noam Atia, Hadar Averbuch-Elor

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to SIGGRAPH 2024. Project webpage: https://tau-vailab.github.io/Spice-E

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.03485 2024-05-07 cs.CV cs.GR 81%

LGTM: Local-to-Global Text-Driven Human Motion Diffusion Model

Haowen Sun, Ruikun Zheng, Haibin Huang, Chongyang Ma, Hui Huang, Ruizhen Hu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments 9 pages,7 figures, SIGGRAPH 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.04930 2024-05-03 cs.CV cs.GR cs.LG 81%

Blue noise for diffusion models

Xingchang Huang, Corentin Salaün, Cristina Vasconcelos, Christian Theobalt, Cengiz Öztireli, Gurprit Singh

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments SIGGRAPH 2024 Conference Proceedings; Project page: https://xchhuang.github.io/bndm

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10518 2024-04-23 cs.CV cs.GR cs.SD eess.AS 81%

Lodge: A Coarse to Fine Diffusion Network for Long Dance Generation Guided by the Characteristic Dance Primitives

Ronghui Li, YuXiang Zhang, Yachao Zhang, Hongwen Zhang, Jie Guo, Yan Zhang, Yebin Liu, Xiu Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by CVPR2024, Project page: https://li-ronghui.github.io/lodge

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.04747 2024-04-09 cs.SD cs.AI cs.CV cs.GR eess.AS 81%

DiffSHEG: A Diffusion-Based Approach for Real-Time Speech-driven Holistic 3D Expression and Gesture Generation

Junming Chen, Yunfei Liu, Jianan Wang, Ailing Zeng, Yu Li, Qifeng Chen

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by CVPR 2024. Project page: https://jeremycjm.github.io/proj/DiffSHEG

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02396 2024-04-04 cs.CV cs.GR cs.LG 81%

Enhancing Diffusion-based Point Cloud Generation with Smoothness Constraint

Yukun Li, Liping Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01862 2024-04-03 cs.CV cs.HC cs.MM 81%

Co-Speech Gesture Video Generation via Motion-Decoupled Diffusion Model

Xu He, Qiaochu Huang, Zhensong Zhang, Zhiwei Lin, Zhiyong Wu, Sicheng Yang, Minglei Li, Zhiyi Chen, Songcen Xu, Xiaofei Wu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments 22 pages, 8 figures, CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01050 2024-04-02 cs.CV cs.GR cs.HC cs.LG 81%

Drag Your Noise: Interactive Point-based Editing via Diffusion Semantic Propagation

Haofeng Liu, Chenshu Xu, Yifei Yang, Lihua Zeng, Shengfeng He

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16739 2024-04-02 cs.CV cs.GR 81%

As-Plausible-As-Possible: Plausibility-Aware Mesh Deformation Using 2D Diffusion Priors

Seungwoo Yoo, Kunho Kim, Vladimir G. Kim, Minhyuk Sung

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://as-plausible-as-possible.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05737 2024-04-01 cs.CV cs.AI cs.MM 81%

Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Lijun Yu, José Lezama, Nitesh B. Gundavarapu, Luca Versari, Kihyuk Sohn, David Minnen, Yong Cheng, Vighnesh Birodkar, Agrim Gupta, Xiuye Gu, Alexander G. Hauptmann, Boqing Gong, Ming-Hsuan Yang, Irfan Essa, David A. Ross, Lu Jiang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17005 2024-03-26 cs.CV cs.MM 81%

TRIP: Temporal Residual Learning with Image Noise Prior for Image-to-Video Diffusion Models

Zhongwei Zhang, Fuchen Long, Yingwei Pan, Zhaofan Qiu, Ting Yao, Yang Cao, Tao Mei

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments CVPR 2024; Project page: https://trip-i2v.github.io/TRIP/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17000 2024-03-26 cs.CV cs.MM 81%

Learning Spatial Adaptation and Temporal Coherence in Diffusion Models for Video Super-Resolution

Zhikai Chen, Fuchen Long, Zhaofan Qiu, Ting Yao, Wengang Zhou, Jiebo Luo, Tao Mei

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16024 2024-03-26 cs.LG cs.CV cs.GR 81%

A Unified Module for Accelerating STABLE-DIFFUSION: LCM-LORA

Ayush Thakur, Rashmi Vashisth

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.14729 2024-03-26 cs.CV cs.GR 81%

MAS: Multi-view Ancestral Sampling for 3D motion generation using 2D diffusion

Roy Kapon, Guy Tevet, Daniel Cohen-Or, Amit H. Bermano

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.12274 2024-03-22 cs.CV cs.AI cs.GR 81%

Intrinsic Image Diffusion for Indoor Single-view Material Estimation

Peter Kocsis, Vincent Sitzmann, Matthias Nießner

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://peter-kocsis.github.io/IntrinsicImageDiffusion/ Video: https://youtu.be/lz0meJlj5cA

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12036 2024-03-19 cs.CV cs.GR cs.LG 81%

One-Step Image Translation with Text-to-Image Models

Gaurav Parmar, Taesung Park, Srinivasa Narasimhan, Jun-Yan Zhu

专题命中 扩散模型 :text-to-image(title);diffusion(abstract);分类 cs.CV、cs.GR

Comments Github: https://github.com/GaParmar/img2img-turbo

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08459 2024-03-19 cs.CV cs.AI cs.GR cs.SD eess.AS 81%

FaceTalk: Audio-Driven Motion Diffusion for Neural Parametric Head Models

Shivangi Aneja, Justus Thies, Angela Dai, Matthias Nießner

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Paper Video: https://youtu.be/7Jf0kawrA3Q Project Page: https://shivangi-aneja.github.io/projects/facetalk/

Journal ref CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15573 2024-03-15 cs.CV cs.GR 81%

EucliDreamer: Fast and High-Quality Texturing for 3D Models with Stable Diffusion Depth

Cindy Le, Congrui Hetang, Chendi Lin, Ang Cao, Yihui He

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.16725 2024-03-15 cs.CV cs.AI cs.MM 81%

Terrain Diffusion Network: Climatic-Aware Terrain Generation with Geological Sketch Guidance

Zexin Hu, Kun Hu, Clinton Mo, Lei Pan, Zhiyong Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17723 2024-02-29 cs.CV cs.MM cs.SD eess.AS 81%

Seeing and Hearing: Open-domain Visual-Audio Generation with Diffusion Latent Aligners

Yazhou Xing, Yingqing He, Zeyue Tian, Xintao Wang, Qifeng Chen

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Accepted to CVPR 2024. Project website: https://yzxing87.github.io/Seeing-and-Hearing/

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.14253 2024-02-23 cs.CV cs.GR 81%

MVD$^2$: Efficient Multiview 3D Reconstruction for Multiview Diffusion

Xin-Yang Zheng, Hao Pan, Yu-Xiao Guo, Xin Tong, Yang Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.03445 2024-02-22 cs.CV cs.GR cs.LG 81%

Denoising Diffusion via Image-Based Rendering

Titas Anciukevičius, Fabian Manhardt, Federico Tombari, Paul Henderson

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted at ICLR 2024. Project page: https://anciukevicius.github.io/generative-image-based-rendering

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15399 2024-02-22 cs.CV cs.AI cs.GR 81%

Sin3DM: Learning a Diffusion Model from a Single 3D Textured Shape

Rundi Wu, Ruoshi Liu, Carl Vondrick, Changxi Zheng

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to ICLR 2024. Project page: https://Sin3DM.github.io, Code: https://github.com/Sin3DM/Sin3DM

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.06146 2024-01-15 cs.CV cs.GR 81%

AAMDM: Accelerated Auto-regressive Motion Diffusion Model

Tianyu Li, Calvin Qiao, Guanqiao Ren, KangKang Yin, Sehoon Ha

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.04362 2024-01-10 cs.CV cs.AI cs.GR 81%

Representative Feature Extraction During Diffusion Process for Sketch Extraction with One Example

Kwan Yun, Youngseo Kim, Kwanggyoon Seo, Chang Wook Seo, Junyong Noh

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments 8 pages(main paper), 8 pages(supplementary material)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.02804 2024-01-09 cs.CV cs.GR 81%

DiffBody: Diffusion-based Pose and Shape Editing of Human Images

Yuta Okuyama, Yuki Endo, Yoshihiro Kanamori

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to WACV 2024, project page: https://www.cgg.cs.tsukuba.ac.jp/~okuyama/pub/diffbody/

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03937 2024-01-08 cs.SD cs.CV cs.MM eess.AS 81%

Diffusion Models as Masked Audio-Video Learners

Elvis Nunez, Yanzi Jin, Mohammad Rastegari, Sachin Mehta, Maxwell Horton

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Camera-ready version for the Machine Learning for Audio Workshop at NeurIPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏