arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2505.06890 2025-05-13 cs.LG cs.CV eess.IV 79%

Image Classification Using a Diffusion Model as a Pre-Training Model

Kosuke Ukita, Ye Xiaolong, Tsuyoshi Okita

机构 * Kosuke Ukita 1, a , Ye Xiaolong 1, b , Tsuyoshi Okita 1, C(未知机构)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06603 2025-05-13 cs.CV 79%

ReplayCAD: Generative Diffusion Replay for Continual Anomaly Detection

Lei Hu, Zhiyong Gan, Ling Deng, Jinglin Liang, Lingyu Liang, Shuangping Huang, Tianshui Chen

机构 * South China University of Technology(华南理工大学) China United Network Communications Corporation Limited Guangdong Branch(中国联合网络通信集团有限公司广东分公司) Pazhou Laboratory(琶洲实验室) Guangdong University of Technology(广东工业大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by IJCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15831 2025-05-13 cs.CV 79%

EDEN: Enhanced Diffusion for High-quality Large-motion Video Frame Interpolation

Zihao Zhang, Haoran Chen, Haoyu Zhao, Guansong Lu, Yanwei Fu, Hang Xu, Zuxuan Wu

机构 * Shanghai Key Lab of Intell. Info. Processing, School of CS, Fudan University(上海智能信息处理关键实验室,复旦大学计算机学院) Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心) Noah’s Ark Lab, Huawei(华为诺亚实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06166 2025-05-12 cs.CV 79%

DiffLocks: Generating 3D Hair from a Single Image using Diffusion Models

Radu Alexandru Rosu, Keyu Wu, Yao Feng, Youyi Zheng, Michael J. Black

机构 * Meshcapade Zhejiang University(浙江大学) Stanford University(斯坦福大学) Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06055 2025-05-12 cs.CV 79%

Towards Better Cephalometric Landmark Detection with Diffusion Data Generation

Dongqian Guo, Wencheng Han, Pang Lyu, Yuxi Zhou, Jianbing Shen

机构 * State Key Laboratory of Internet of Things for Smart City, Department of Computer and Information Science, University of Macau(物联网智能城市国家重点实验室,澳门大学计算机与信息科学系) Department of Orthopaedic Surgery, Zhongshan Hospital, Fudan University(复旦大学中山医院骨科部) Department of Periodontology, Justus-Liebig-University of Giessen, Germany(吉森大学牙科系) Department of Periodontics, Stomatology Hospital of Guangzhou Medical University(广州医科大学口腔医院牙科部)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05853 2025-05-12 cs.CV 79%

PICD: Versatile Perceptual Image Compression with Diffusion Rendering

Tongda Xu, Jiahao Li, Bin Li, Yan Wang, Ya-Qin Zhang, Yan Lu

机构 * AIR, Tsinghua University(清华大学人工智能研究院) Microsoft Research Asia(微软亚洲研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05732 2025-05-12 cs.LG cs.CV 79%

Automated Learning of Semantic Embedding Representations for Diffusion Models

Limai Jiang, Yunpeng Cai

机构 * Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Extended version of the paper published in SDM25

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05475 2025-05-09 cs.CV 79%

SVAD: From Single Image to 3D Avatar via Synthetic Data Generation with Video Diffusion and Data Augmentation

Yonwoo Choi

机构 * SECERN AI

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by CVPR 2025 SyntaGen Workshop, Project Page: https://yc4ny.github.io/SVAD/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05473 2025-05-09 cs.CV 79%

DiffusionSfM: Predicting Structure and Motion via Ray Origin and Endpoint Diffusion

Qitao Zhao, Amy Lin, Jeff Tan, Jason Y. Zhang, Deva Ramanan, Shubham Tulsiani

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025. Project website: https://qitaozhao.github.io/DiffusionSfM

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05137 2025-05-09 cs.LG cs.CV 79%

Research on Anomaly Detection Methods Based on Diffusion Models

Yi Chen

机构 * College of Physics and Electronic Information Engineering(物理与电子信息工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 6 pages, 3 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04660 2025-05-09 cs.CL cs.CV 79%

AI-Generated Fall Data: Assessing LLMs and Diffusion Model for Wearable Fall Detection

Sana Alamgeer, Yasine Souissi, Anne H. H. Ngu

机构 * Texas State University(德克萨斯州立大学) University of North Carolina(北卡罗来纳大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04522 2025-05-09 eess.IV cs.CV 79%

Text2CT: Towards 3D CT Volume Generation from Free-text Descriptions Using Diffusion Model

Pengfei Guo, Can Zhao, Dong Yang, Yufan He, Vishwesh Nath, Ziyue Xu, Pedro R. A. S. Bassi, Zongwei Zhou, Benjamin D. Simon, Stephanie Anne Harmon, Baris Turkbey, Daguang Xu

机构 * NVIDIA Johns Hopkins University(约翰霍普金斯大学) National Institutes of Health(国家卫生研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04281 2025-05-09 cs.CV eess.IV 79%

TS-Diff: Two-Stage Diffusion Model for Low-Light RAW Image Enhancement

Yi Li, Zhiyuan Zhang, Jiangnan Xia, Jianghan Cheng, Qilong Wu, Junwei Li, Yibin Tian, Hui Kong

机构 * College of Information Science and Electronic Engineering, Zhejiang University, China(浙江大学信息科学与电子工程学院) School of Computing and Information Systems, Singapore Management University, Singapore(新加坡管理大学计算机与信息科学学院) College of Mechatronics and Control Engineering, Shenzhen University, China(深圳大学机电与控制工程学院) Faculty of Science and Technology, University of Macau, China(澳门大学科技学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments International Joint Conference on Neural Networks (IJCNN)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21487 2025-05-09 cs.CV 79%

DGSolver: Diffusion Generalist Solver with Universal Posterior Sampling for Image Restoration

Hebaixu Wang, Jing Zhang, Haonan Guo, Di Wang, Jiayi Ma, Bo Du

机构 * Wuhan University(武汉大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14423 2025-05-09 cs.CV 79%

PhysFlow: Unleashing the Potential of Multi-modal Foundation Models and Video Diffusion for 4D Dynamic Physical Scene Simulation

Zhuoman Liu, Weicai Ye, Yan Luximon, Pengfei Wan, Di Zhang

机构 * The Hong Kong Polytechnic University(香港理工大学) Kuaishou Technology(快手科技)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025. Homepage: https://zhuomanliu.github.io/PhysFlow/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04306 2025-05-08 cs.CV 79%

MoDE: Mixture of Diffusion Experts for Any Occluded Face Recognition

Qiannan Fan, Zhuoyang Li, Jitong Li, Chenyang Cao

机构 * State Grid Tianjin Economic Research Institute(国家电网天津经济研究院) College of Intelligence and Computing(智能与计算学院) Tianjin University(天津大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 8 pages,7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06675 2025-05-08 cs.CV 79%

Probability Density Geodesics in Image Diffusion Latent Space

Qingtao Yu, Jaskirat Singh, Zhaoyuan Yang, Peter Henry Tu, Jing Zhang, Hongdong Li, Richard Hartley, Dylan Campbell

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.06057 2025-05-08 cs.CV 79%

Illumination and Shadows in Head Rotation: experiments with Denoising Diffusion Models

Andrea Asperti, Gabriele Colasuonno, Antonio Guerra

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03507 2025-05-07 cs.CV 79%

Modality-Guided Dynamic Graph Fusion and Temporal Diffusion for Self-Supervised RGB-T Tracking

Shenglan Li, Rui Yao, Yong Zhou, Hancheng Zhu, Kunyang Sun, Bing Liu, Zhiwen Shao, Jiaqi Zhao

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by the 34th International Joint Conference on Artificial Intelligence (IJCAI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03261 2025-05-07 cs.CV eess.IV 79%

DiffVQA: Video Quality Assessment Using Diffusion Feature Extractor

Wei-Ting Chen, Yu-Jiet Vong, Yi-Tsung Lee, Sy-Yen Kuo, Qiang Gao, Sizhuo Ma, Jian Wang

机构 * National Taiwan University(国立台湾大学) Snap Inc.(Snap公司) Microsoft(微软公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03097 2025-05-07 cs.CV 79%

Not All Parameters Matter: Masking Diffusion Models for Enhancing Generation Ability

Lei Wang, Senmao Li, Fei Yang, Jianye Wang, Ziheng Zhang, Yuhan Liu, Yaxing Wang, Jian Yang

机构 * PCA Lab, VCIP, College of Computer Science, Nankai University(PCA实验室、VCIP、计算机科学学院、南开大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00334 2025-05-06 cs.CV cs.LG 79%

Quaternion Wavelet-Conditioned Diffusion Models for Image Super-Resolution

Luigi Sigillo, Christian Bianchi, Aurelio Uncini, Danilo Comminiello

机构 * Dept. Information Engineering, Electronics and Telecommunications (DIET), Sapienza University of Rome(信息工程、电子与电信系,罗马萨皮恩扎大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted for presentation at IJCNN 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00968 2025-05-06 cs.CV cs.LG 79%

CoDe: Blockwise Control for Denoising Diffusion Models

Anuj Singh, Sayak Mukherjee, Ahmad Beirami, Hadi Jamali-Rad

机构 * Delft University of Technology(代尔夫特理工大学) Shell Global Solutions International B.V.(壳牌全球解决方案国际有限公司) Massachusetts Institute of Technology(麻省理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Transactions on Machine Learning Research, 2025. ISSN: 2835-8856

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10821 2025-05-06 cs.CV 79%

Tex4D: Zero-shot 4D Scene Texturing with Video Diffusion Models

Jingzhi Bao, Xueting Li, Ming-Hsuan Yang

机构 * CUHK-Shenzhen(香港中文大学(深圳)) NVIDIA(英伟达) UC Merced(加州大学默塞德分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://tex4d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.11388 2025-05-06 eess.IV cs.CV 79%

Physics-informed Deep Diffusion MRI Reconstruction with Synthetic Data: Break Training Data Bottleneck in Artificial Intelligence

Chen Qian, Haoyu Zhang, Yuncheng Gao, Mingyang Han, Zi Wang, Dan Ruan, Yu Shen, Yaping Wu, Yirong Zhou, Chengyan Wang, Boyu Jiang, Ran Tao, Zhigang Wu, Jiazheng Wang, Liuhong Zhu, Yi Guo, Taishan Kang, Jianzhong Lin, Tao Gong, Chen Yang, Guoqiang Fei, Meijin Lin, Di Guo, Jianjun Zhou, Meiyun Wang, Xiaobo Qu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.00998 2025-05-05 cs.CV 79%

Part-aware Shape Generation with Latent 3D Diffusion of Neural Voxel Fields

Yuhang Huang, SHilong Zou, Xinwang Liu, Kai Xu

机构 * School of Computer, National University of Defense Technology(计算机学院,国防科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This paper is accepted by TVCG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00687 2025-05-02 eess.IV cs.CV 79%

GuideSR: Rethinking Guidance for One-Step High-Fidelity Diffusion-Based Super-Resolution

Aditya Arora, Zhengzhong Tu, Yufei Wang, Ruizheng Bai, Jian Wang, Sizhuo Ma

机构 * TU Darmstadt(图宾根大学) Texas A&M University(德克萨斯A&M大学) Snap Inc.(Snap公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08215 2025-05-02 cs.CV cs.AI 79%

LT3SD: Latent Trees for 3D Scene Diffusion

Quan Meng, Lei Li, Matthias Nießner, Angela Dai

机构 * Technical University of Munich(慕尼黑技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://quan-meng.github.io/projects/lt3sd/ Video: https://youtu.be/AJ5sG9VyjGA

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00426 2025-05-02 cs.CV 79%

Leveraging Pretrained Diffusion Models for Zero-Shot Part Assembly

Ruiyuan Zhang, Qi Wang, Jiaxiang Liu, Yu Zhang, Yuchi Huo, Chao Wu

机构 * Zhejiang University(浙江大学) North China Electric Power University(华北电力大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 12 figures, Accepted by IJCAI-2025

Journal ref IJCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.03048 2025-05-02 cs.CV 79%

Latte: Latent Diffusion Transformer for Video Generation

Xin Ma, Yaohui Wang, Xinyuan Chen, Gengyun Jia, Ziwei Liu, Yuan-Fang Li, Cunjian Chen, Yu Qiao

机构 * Department of Data Science & AI, Faculty of Information Technology, Monash University(数据科学与人工智能系,信息科技学院,莫纳什大学) Shanghai AI Laboratory(上海人工智能实验室) Nanjing University of Posts and Telecommunications(南京邮电大学) S-Lab, Nanyang Technological University(南洋理工大学S实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by Transactions on Machine Learning Research 2025; Project Page: https://maxin-cn.github.io/latte_project

详情

展开后加载摘要…

URL PDF HTML 收藏