arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70082 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70082 篇

2410.06149 2024-10-10 cs.CV cs.MM eess.IV 81%

Toward Scalable Image Feature Compression: A Content-Adaptive and Diffusion-Based Approach

Sha Guo, Zhuo Chen, Yang Zhao, Ning Zhang, Xiaotong Li, Lingyu Duan

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Journal ref in Proceedings of the 31st ACM International Conference on Multimedia, pp. 1431-1442, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.05412 2024-10-10 cs.LG cs.CV cs.MM cs.SD eess.AS 81%

CMMD: Contrastive Multi-Modal Diffusion for Video-Audio Conditional Modeling

Ruihan Yang, Hannes Gamper, Sebastian Braun

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.19244 2024-10-04 cs.CV cs.MM 81%

Radio Frequency Signal based Human Silhouette Segmentation: A Sequential Diffusion Approach

Penghui Wen, Kun Hu, Dong Yuan, Zhiyuan Ning, Changyang Li, Zhiyong Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01366 2024-10-03 cs.CV cs.MM 81%

Harnessing the Latent Diffusion Model for Training-Free Image Style Transfer

Kento Masui, Mayu Otani, Masahiro Nomura, Hideki Nakayama

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.13049 2024-10-03 eess.AS cs.CV cs.MM cs.SD 81%

DiffSSD: A Diffusion-Based Dataset For Speech Forensics

Kratika Bhagtani, Amit Kumar Singh Yadav, Paolo Bestagini, Edward J. Delp

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Submitted to IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01295 2024-10-03 cs.CV cs.GR 81%

LaGeM: A Large Geometry Model for 3D Representation Learning and Diffusion

Biao Zhang, Peter Wonka

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments For more information: https://1zb.github.io/LaGeM

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.20502 2024-10-01 cs.LG cs.AI cs.CV cs.GR 81%

COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models

Divyanshu Daiya, Damon Conover, Aniket Bera

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments 9 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17145 2024-09-26 cs.CV cs.GR cs.LG 81%

DreamWaltz-G: Expressive 3D Gaussian Avatars from Skeleton-Guided 2D Diffusion

Yukun Huang, Jianan Wang, Ailing Zeng, Zheng-Jun Zha, Lei Zhang, Xihui Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://yukun-huang.github.io/DreamWaltz-G/

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.15889 2024-09-23 eess.IV cs.CV cs.IT cs.LG cs.MM math.IT 81%

High Perceptual Quality Wireless Image Delivery with Denoising Diffusion Models

Selim F. Yilmaz, Xueyan Niu, Bo Bai, Wei Han, Lei Deng, Deniz Gunduz

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments 6 pages, 5 figures. Published at INFOCOM 2024 Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08947 2024-09-18 cs.CV cs.GR 81%

A Diffusion Approach to Radiance Field Relighting using Multi-Illumination Synthesis

Yohan Poirier-Ginter, Alban Gauthier, Julien Philip, Jean-Francois Lalonde, George Drettakis

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project site https://repo-sam.inria.fr/fungraph/generative-radiance-field-relighting/

Journal ref Computer Graphics Forum, Volume 43 (2024), Number 4

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18159 2024-08-21 cs.CV cs.GR 81%

Human-Aware 3D Scene Generation with Spatially-constrained Diffusion Models

Xiaolin Hong, Hongwei Yi, Fazhi He, Qiong Cao

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.09384 2024-08-20 cs.CV cs.MM 81%

FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model

Ziyu Yao, Xuxin Cheng, Zhiqi Huang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Accepted by ACM Multimedia 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.05416 2024-08-13 cs.CV cs.AI cs.MM 81%

High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model

Weizhi Zhong, Junfan Lin, Peixin Chen, Liang Lin, Guanbin Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments submitted to IEEE Transactions on Image Processing(TIP)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20337 2024-07-31 cs.CV cs.AI cs.MM 81%

Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities

Lorenzo Baraldi, Federico Cocchi, Marcella Cornia, Lorenzo Baraldi, Alessandro Nicolosi, Rita Cucchiara

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments ECCV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.13759 2024-07-26 cs.CV cs.GR 81%

Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion

Boyang Deng, Richard Tucker, Zhengqi Li, Leonidas Guibas, Noah Snavely, Gordon Wetzstein

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments *Equal Contributions; Fixed few duplicated references from 1st upload; Project Page: https://boyangdeng.com/streetscapes

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17050 2024-07-25 cs.CV cs.GR 81%

Surf-D: Generating High-Quality Surfaces of Arbitrary Topologies Using Diffusion Models

Zhengming Yu, Zhiyang Dou, Xiaoxiao Long, Cheng Lin, Zekun Li, Yuan Liu, Norman Müller, Taku Komura, Marc Habermann, Christian Theobalt, Xin Li, Wenping Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted to ECCV 2024. Project Page: https://yzmblog.github.io/projects/SurfD/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12034 2024-07-22 cs.CV cs.GR cs.LG 81%

VFusion3D: Learning Scalable 3D Generative Models from Video Diffusion Models

Junlin Han, Filippos Kokkinos, Philip Torr

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments ECCV 2024. Project page: https://junlinhan.github.io/projects/vfusion3d.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.05967 2024-07-19 cs.CV cs.GR cs.LG 81%

Distilling Diffusion Models into Conditional GANs

Minguk Kang, Richard Zhang, Connelly Barnes, Sylvain Paris, Suha Kwak, Jaesik Park, Eli Shechtman, Jun-Yan Zhu, Taesung Park

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://mingukkang.github.io/Diffusion2GAN/ (ECCV2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12783 2024-07-18 cs.CV cs.GR 81%

SMooDi: Stylized Motion Diffusion Model

Lei Zhong, Yiming Xie, Varun Jampani, Deqing Sun, Huaizu Jiang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments ECCV 2024. Project page: https://neu-vi.github.io/SMooDi/

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.10736 2024-07-16 cs.CV cs.AI cs.MM 81%

When Synthetic Traces Hide Real Content: Analysis of Stable Diffusion Image Laundering

Sara Mandelli, Paolo Bestagini, Stefano Tubaro

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.16864 2024-06-25 cs.CV cs.AI cs.GR 81%

StableNormal: Reducing Diffusion Variance for Stable and Sharp Normal

Chongjie Ye, Lingteng Qiu, Xiaodong Gu, Qi Zuo, Yushuang Wu, Zilong Dong, Liefeng Bo, Yuliang Xiu, Xiaoguang Han

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments HF Demo: hf.co/Stable-X, Video: https://www.youtube.com/watch?v=sylXTxG_U2U

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17113 2024-06-25 cs.CV cs.GR 81%

Transparent Image Layer Diffusion using Latent Transparency

Lvmin Zhang, Maneesh Agrawala

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments 44 pages, 37 figures, github.com/layerdiffusion/LayerDiffuse

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06539 2024-06-12 cs.CV cs.GR 81%

MatFusion: A Generative Diffusion Model for SVBRDF Capture

Sam Sartor, Pieter Peers

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Journal ref ACM SIGGRAPH Asia 2023 Conference Proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04338 2024-06-12 cs.CV cs.AI cs.GR 81%

Physics3D: Learning Physical Properties of 3D Gaussians via Video Diffusion

Fangfu Liu, Hanyang Wang, Shunyu Yao, Shengjun Zhang, Jie Zhou, Yueqi Duan

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://liuff19.github.io/Physics3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06508 2024-06-11 cs.CV cs.AI cs.GR 81%

Monkey See, Monkey Do: Harnessing Self-attention in Motion Diffusion for Zero-shot Motion Transfer

Sigal Raab, Inbar Gat, Nathan Sala, Guy Tevet, Rotem Shalev-Arkushin, Ohad Fried, Amit H. Bermano, Daniel Cohen-Or

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Video: https://www.youtube.com/watch?v=s5oo3sKV0YU, Project page: https://monkeyseedocg.github.io, Code: https://github.com/MonkeySeeDoCG/MoMo-code

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06465 2024-06-11 cs.CV cs.AI cs.CL cs.LG cs.MM 81%

AID: Adapting Image2Video Diffusion Models for Instruction-guided Video Prediction

Zhen Xing, Qi Dai, Zejia Weng, Zuxuan Wu, Yu-Gang Jiang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04673 2024-06-10 cs.CV cs.AI cs.MM eess.AS 81%

MeLFusion: Synthesizing Music from Image and Language Cues using Diffusion Models

Sanjoy Chowdhury, Sayan Nag, K J Joseph, Balaji Vasan Srinivasan, Dinesh Manocha

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Accepted at CVPR 2024 as Highlight paper. Webpage: https://schowdhury671.github.io/melfusion_cvpr2024/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11126 2024-05-27 cs.CV cs.GR cs.LG 81%

Flexible Motion In-betweening with Diffusion Models

Setareh Cohan, Guy Tevet, Daniele Reda, Xue Bin Peng, Michiel van de Panne

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments SIGGRAPH 2024. For project page and code, see https://setarehc.github.io/CondMDI/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10540 2024-05-24 cs.CV cs.GR 81%

VecFusion: Vector Font Generation with Diffusion

Vikas Thamizharasan, Difan Liu, Shantanu Agarwal, Matthew Fisher, Michael Gharbi, Oliver Wang, Alec Jacobson, Evangelos Kalogerakis

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09539 2024-05-17 eess.IV cs.CV cs.MM 81%

MMFusion: Multi-modality Diffusion Model for Lymph Node Metastasis Diagnosis in Esophageal Cancer

Chengyu Wu, Chengkai Wang, Yaqi Wang, Huiyu Zhou, Yatao Zhang, Qifeng Wang, Shuai Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Early accepted to MICCAI 2024 (6/6/5)

详情

展开后加载摘要…

URL PDF HTML 收藏