arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70082 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70082 篇

2504.12540 2025-04-18 cs.GR cs.CV cs.RO 81%

UniPhys: Unified Planner and Controller with Diffusion for Flexible Physics-Based Character Control

Yan Wu, Korrawe Karunratanakul, Zhengyi Luo, Siyu Tang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://wuyan01.github.io/uniphys-project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07210 2025-04-15 cs.GR cs.CV cs.LG 81%

MESA: Text-Driven Terrain Generation Using Latent Diffusion and Global Copernicus Data

Paul Borne--Pons, Mikolaj Czerkawski, Rosalie Martin, Romain Rouffet

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted at CVPR 2025 Workshop MORSE

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10209 2025-04-15 cs.CV cs.AI cs.GR 81%

GAF: Gaussian Avatar Reconstruction from Monocular Videos via Multi-view Diffusion

Jiapeng Tang, Davide Davoli, Tobias Kirschstein, Liam Schoneveld, Matthias Niessner

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Paper Video: https://youtu.be/QuIYTljvhyg Project Page: https://tangjiapeng.github.io/projects/GAF

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.05260 2025-04-15 cs.CV cs.GR 81%

DartControl: A Diffusion-Based Autoregressive Motion Model for Real-Time Text-Driven Motion Control

Kaifeng Zhao, Gen Li, Siyu Tang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Updated ICLR camera ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08912 2025-04-11 cs.CV cs.MM 81%

Reversing the Damage: A QP-Aware Transformer-Diffusion Approach for 8K Video Restoration under Codec Compression

Ali Mollaahmadi Dehaghi, Reza Razavi, Mohammad Moshirpour

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments 12 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16294 2025-04-10 cs.CV cs.GR cs.LG 81%

GenCAD: Image-Conditioned Computer-Aided Design Generation with Transformer-Based Contrastive Representation and Diffusion Priors

Md Ferdous Alam, Faez Ahmed

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments 24 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04956 2025-04-09 cs.GR cs.CV 81%

REWIND: Real-Time Egocentric Whole-Body Motion Diffusion with Exemplar-Based Identity Conditioning

Jihyun Lee, Weipeng Xu, Alexander Richard, Shih-En Wei, Shunsuke Saito, Shaojie Bai, Te-Li Wang, Minhyuk Sung, Tae-Kyun Kim, Jason Saragih

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to CVPR 2025, project page: https://jyunlee.github.io/projects/rewind/

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.17550 2025-04-08 cs.CV cs.MM 81%

DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder

Chenpeng Du, Qi Chen, Tianyu He, Xu Tan, Xie Chen, Kai Yu, Sheng Zhao, Jiang Bian

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Accepted to ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16915 2025-04-07 cs.CV cs.AI cs.GR cs.SD eess.AS 81%

FADA: Fast Diffusion Avatar Synthesis with Mixed-Supervised Multi-CFG Distillation

Tianyun Zhong, Chao Liang, Jianwen Jiang, Gaojie Lin, Jiaqi Yang, Zhou Zhao

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments CVPR 2025, Homepage https://fadavatar.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14485 2025-04-03 cs.GR cs.CV 81%

Lux Post Facto: Learning Portrait Performance Relighting with Conditional Video Diffusion and a Hybrid Dataset

Yiqun Mei, Mingming He, Li Ma, Julien Philip, Wenqi Xian, David M George, Xueming Yu, Gabriel Dedic, Ahmet Levent Taşel, Ning Yu, Vishal M. Patel, Paul Debevec

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12836 2025-04-03 cs.GR cs.AI cs.CV cs.HC 81%

EditRoom: LLM-parameterized Graph Diffusion for Composable 3D Room Layout Editing

Kaizhi Zheng, Xiaotong Chen, Xuehai He, Jing Gu, Linjie Li, Zhengyuan Yang, Kevin Lin, Jianfeng Wang, Lijuan Wang, Xin Eric Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01016 2025-04-02 cs.GR cs.AI cs.CV 81%

GeometryCrafter: Consistent Geometry Estimation for Open-world Videos with Diffusion Priors

Tian-Xing Xu, Xiangjun Gao, Wenbo Hu, Xiaoyu Li, Song-Hai Zhang, Ying Shan

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project webpage: https://geometrycrafter.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.24210 2025-04-01 cs.CV cs.AI cs.MM 81%

DiET-GS: Diffusion Prior and Event Stream-Assisted Motion Deblurring 3D Gaussian Splatting

Seungjun Lee, Gim Hee Lee

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments CVPR 2025. Project Page: https://diet-gs.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01316 2025-04-01 cs.CV cs.AI cs.MM 81%

Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation

Xin Yan, Yuxuan Cai, Qiuyue Wang, Yuan Zhou, Wenhao Huang, Huan Yang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments This paper is accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18597 2025-03-27 cs.CV cs.AI cs.MM 81%

DiTCtrl: Exploring Attention Control in Multi-Modal Diffusion Transformer for Tuning-Free Multi-Prompt Longer Video Generation

Minghong Cai, Xiaodong Cun, Xiaoyu Li, Wenze Liu, Zhaoyang Zhang, Yong Zhang, Ying Shan, Xiangyu Yue

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments CVPR 2025; 21 pages, 23 figures, Project page: https://onevfall.github.io/project_page/ditctrl ; GitHub repository: https://github.com/TencentARC/DiTCtrl

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19036 2025-03-26 cs.CV cs.GR 81%

PCDreamer: Point Cloud Completion Through Multi-view Diffusion Priors

Guangshun Wei, Yuan Feng, Long Ma, Chen Wang, Yuanfeng Zhou, Changjian Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://gsw-d.github.io/PCDreamer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17059 2025-03-24 cs.GR cs.CV cs.SD eess.AS 81%

DIDiffGes: Decoupled Semi-Implicit Diffusion Models for Real-time Gesture Generation from Speech

Yongkang Cheng, Shaoli Huang, Xuelin Chen, Jifeng Ning, Mingming Gong

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16801 2025-03-24 cs.GR cs.AI cs.CV 81%

Auto-Regressive Diffusion for Generating 3D Human-Object Interactions

Zichen Geng, Zeeshan Hayder, Wei Liu, Ajmal Saeed Mian

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15809 2025-03-21 cs.GR cs.CV 81%

Controlling Avatar Diffusion with Learnable Gaussian Embedding

Xuan Gao, Jingtao Zhou, Dongyu Liu, Yuqi Zhou, Juyong Zhang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project Page: https://ustc3dv.github.io/Learn2Control/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15586 2025-03-21 cs.GR cs.CV 81%

How to Train Your Dragon: Automatic Diffusion-Based Rigging for Characters with Diverse Topologies

Zeqi Gu, Difan Liu, Timothy Langlois, Matthew Fisher, Abe Davis

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to Eurographics 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13836 2025-03-19 cs.CV cs.AI cs.GR cs.LG 81%

SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing

Seokhyeon Hong, Chaelin Kim, Serin Yoon, Junghyun Nam, Sihun Cha, Junyong Noh

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments CVPR 2025; Project page https://seokhyeonhong.github.io/projects/salad/

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12957 2025-03-18 cs.CV cs.GR 81%

3DTopia-XL: Scaling High-quality 3D Asset Generation via Primitive Diffusion

Zhaoxi Chen, Jiaxiang Tang, Yuhao Dong, Ziang Cao, Fangzhou Hong, Yushi Lan, Tengfei Wang, Haozhe Xie, Tong Wu, Shunsuke Saito, Liang Pan, Dahua Lin, Ziwei Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments CVPR 2025, Code https://github.com/3DTopia/3DTopia-XL Project Page https://3dtopia.github.io/3DTopia-XL/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06784 2025-03-11 cs.GR cs.AI cs.CV cs.LG cs.RO 81%

Infinite Leagues Under the Sea: Photorealistic 3D Underwater Terrain Generation by Latent Fractal Diffusion Models

Tianyi Zhang, Weiming Zhi, Joshua Mangelson, Matthew Johnson-Roberson

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05638 2025-03-10 cs.CV cs.AI cs.GR 81%

TrajectoryCrafter: Redirecting Camera Trajectory for Monocular Videos via Diffusion Models

Mark YU, Wenbo Hu, Jinbo Xing, Ying Shan

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project webpage: https://trajectorycrafter.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00726 2025-03-04 cs.GR cs.AI cs.CV 81%

Enhancing Monocular 3D Scene Completion with Diffusion Model

Changlin Song, Jiaqi Wang, Liyun Zhu, He Weng

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments All authors had equal contribution

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04810 2025-03-04 cs.LG cs.CV cs.DC cs.MM 81%

FedBiP: Heterogeneous One-Shot Federated Learning with Personalized Latent Diffusion Models

Haokun Chen, Hang Li, Yao Zhang, Jinhe Bi, Gengyuan Zhang, Yueqi Zhang, Philip Torr, Jindong Gu, Denis Krompass, Volker Tresp

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17880 2025-02-26 cs.CR cs.CV cs.MM 81%

VVRec: Reconstruction Attacks on DL-based Volumetric Video Upstreaming via Latent Diffusion Model with Gamma Distribution

Rui Lu, Bihai Zhang, Dan Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17842 2025-02-26 cs.CV cs.LG cs.MM cs.SD eess.AS 81%

MMDisCo: Multi-Modal Discriminator-Guided Cooperative Diffusion for Joint Audio and Video Generation

Akio Hayakawa, Masato Ishii, Takashi Shibuya, Yuki Mitsufuji

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09599 2025-02-26 cs.CV cs.MM 81%

Language-Guided Diffusion Model for Visual Grounding

Sijia Chen, Baochun Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments 20 pages, 16 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01639 2025-02-04 cs.CV cs.GR cs.LG 81%

SliderSpace: Decomposing the Visual Capabilities of Diffusion Models

Rohit Gandikota, Zongze Wu, Richard Zhang, David Bau, Eli Shechtman, Nick Kolkin

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project Website: https://sliderspace.baulab.info

详情

展开后加载摘要…

URL PDF HTML 收藏