arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70196 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70196 篇

2409.02638 2025-11-17 cs.CV 79%

MADiff: Motion-Aware Mamba Diffusion Models for Hand Trajectory Prediction on Egocentric Videos

Junyi Ma, Xieyuanli Chen, Wentao Bao, Jingyi Xu, Hesheng Wang

机构 * IRMV Lab, School of Automation and Intelligent Sensing, Shanghai Jiao Tong University and State Key Laboratory of Avionics Integration and Aviation System-of-Systems Synthesis, Shanghai Key Laboratory of Navigation and Location Based Services(IRMV实验室,自动化与智能感知学院,上海交通大学,航空集成与航空系统-of-Systems综合国家重点实验室,导航与定位服务重点实验室) College of Intelligence Science and Technology, National University of Defense Technology(智能科学与技术学院,国防科技大学) Department of Electronic Engineering, Shanghai Jiao Tong University(电子工程学院,上海交通大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to TPAMI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.04370 2025-11-17 cs.CV 79%

Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos

Junyi Ma, Jingyi Xu, Xieyuanli Chen, Hesheng Wang

机构 * IRMV Lab, the Department of Automation, Shanghai Jiao Tong University, Shanghai 200240, China(IRMV实验室,自动化系,上海交通大学,上海200240,中国) College of Intelligence Science and Technology, National University of Defense Technology, Changsha 410073, China(智能科学与技术学院,国防科技大学,长沙410073,中国) Department of Electronic Engineering, Shanghai Jiao Tong University, Shanghai 200240, China(电子工程系,上海交通大学,上海200240,中国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to IROS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10391 2025-11-14 cs.CV 79%

GrounDiff: Diffusion-Based Ground Surface Generation from Digital Surface Models

Oussema Dhaouadi, Johannes Meier, Jacques Kaiser, Daniel Cremers

机构 * DeepScenario TU Munich(慕尼黑工业大学) Munich Center for Machine Learning(慕尼黑机器学习中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09965 2025-11-14 cs.CV 79%

Equivariant Sampling for Improving Diffusion Model-based Image Restoration

Chenxu Wu, Qingpeng Kong, Peiang Zhao, Wendi Yang, Wenxin Ma, Fenghe Tang, Zihang Jiang, S. Kevin Zhou

机构 * School of Biomedical Engineering, Division of Life Sciences and Medicine, USTC(生物医学工程学院,生命科学与医学系,中国科学技术大学) MIRACLE Center, Suzhou Institute for Advance Research, USTC(MIRACLE中心,苏州市先进研究院,中国科学技术大学) Key Laboratory of Intelligent Information Processing of CAS, ICT, CAS(中国科学院智能信息处理重点实验室,ICT,中国科学院) State Key Laboratory of Precision and Intelligent Chemistry, USTC(中国科学技术大学精密与智能化学国家重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 12 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09962 2025-11-14 cs.LG cs.AI cs.MM 79%

AI-Integrated Decision Support System for Real-Time Market Growth Forecasting and Multi-Source Content Diffusion Analytics

Ziqing Yin, Xuanjing Chen, Xi Zhang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.11343 2025-11-14 cs.CV stat.AP 79%

Latent Knowledge-Guided Video Diffusion for Scientific Phenomena Generation from a Single Initial Frame

Qinglong Cao, Xirui Li, Ding Wang, Chao Ma, Yuntian Chen, Xiaokang Yang

机构 * Eitech(艾tech)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09502 2025-11-13 cs.CV cs.AI 79%

DreamPose3D: Hallucinative Diffusion with Prompt Learning for 3D Human Pose Estimation

Jerrin Bright, Yuhao Chen, John S. Zelek

机构 * Vision and Image Processing Lab(视觉与图像处理实验室) University of Waterloo(滑铁卢大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09184 2025-11-13 cs.CV 79%

DBINDS -- Can Initial Noise from Diffusion Model Inversion Help Reveal AI-Generated Videos?

Yanlin Wu, Xiaogang Yuan, Dezhi An

机构 * School of Cyber Security, Gansu University of Political Science and Law(政治法律科学学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Preprint. Submitted to IEEE Transactions on Dependable and Secure Computing (TDSC) on 16 September 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08823 2025-11-13 cs.CV cs.AI 79%

DT-NVS: Diffusion Transformers for Novel View Synthesis

Wonbong Jang, Jonathan Tremblay, Lourdes Agapito

机构 * UCL(伦敦大学学院) NVIDIA(英伟达)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07496 2025-11-13 cs.CV cs.AI cs.LG stat.ML 79%

Laplacian Score Sharpening for Mitigating Hallucination in Diffusion Models

Barath Chandran. C, Srinivas Anumasa, Dianbo Liu

机构 * Indian Institute of Technology, Roorkee(印度理工学院罗尔基分校) National University of Singapore(新加坡国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08258 2025-11-12 cs.CV 79%

Top2Ground: A Height-Aware Dual Conditioning Diffusion Model for Robust Aerial-to-Ground View Generation

Jae Joong Lee, Bedrich Benes

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08173 2025-11-12 cs.CV 79%

VLMDiff: Leveraging Vision-Language Models for Multi-Class Anomaly Detection with Diffusion

Samet Hicsonmez, Abd El Rahman Shabayek, Djamila Aouada

机构 * University of Luxembourg(卢森堡大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07976 2025-11-12 cs.CV cs.AI 79%

Morphing Through Time: Diffusion-Based Bridging of Temporal Gaps for Robust Alignment in Change Detection

Seyedehanita Madani, Vishal M. Patel

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 9 pages, 5 figures. To appear in WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07816 2025-11-12 cs.CV 79%

Cancer-Net PCa-MultiSeg: Multimodal Enhancement of Prostate Cancer Lesion Segmentation Using Synthetic Correlated Diffusion Imaging

Jarett Dewbury, Chi-en Amy Tai, Alexander Wong

机构 * Systems Design Engineering University of Waterloo(水力工程系统设计系大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at ML4H 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05873 2025-11-12 eess.IV cs.AI cs.CV cs.RO 79%

EndoIR: Degradation-Agnostic All-in-One Endoscopic Image Restoration via Noise-Aware Routing Diffusion

Tong Chen, Xinyu Ma, Long Bai, Wenyang Wang, Yue Sun, Luping Zhou

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21528 2025-11-12 cs.CV cs.AI cs.LG 79%

A Unified and Fast-Sampling Diffusion Bridge Framework via Stochastic Optimal Control

Mokai Pan, Kaizhen Zhu, Yuexin Ma, Yanwei Fu, Jingyi Yu, Jingya Wang, Ye Shi

机构 * School of Information Science and Technology, ShanghaiTech University(信息科学与技术学院,上海科技大学) School of Data Science, Fudan University(数据科学学院,复旦大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05870 2025-11-12 cs.CV cs.AI eess.IV 79%

FaSDiff: Balancing Perception and Semantics in Face Compression via Stable Diffusion Priors

Yimin Zhou, Yichong Xia, Bin Chen, Mingyao Hong, Jiawei Li, Zhi Wang, Yaowei Wang

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) Research Center of Artificial Intelligence, Pengcheng Laboratory(人工智能研究中心,鹏城实验室) Harbin Institute of Technology(哈尔滨工业大学) Huawei Manufacturing(华为制造)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18334 2025-11-12 cs.CV 79%

DODA: Adapting Object Detectors to Dynamic Agricultural Environments in Real-Time with Diffusion

Shuai Xiang, Pieter M. Blok, James Burridge, Haozhou Wang, Wei Guo

机构 * Graduate School of Agricultural and Life Sciences(农业与生命科学研究生院) The University of Tokyo(东京大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments WACV2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07103 2025-11-11 cs.CV cs.AI 79%

GEWDiff: Geometric Enhanced Wavelet-based Diffusion Model for Hyperspectral Image Super-resolution

Sirui Wang, Jiang He, Natàlia Blasco Andreo, Xiao Xiang Zhu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This manuscript has been accepted for publication in AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07067 2025-11-11 cs.CV 79%

RaLD: Generating High-Resolution 3D Radar Point Clouds with Latent Diffusion

Ruijie Zhang, Bixin Zeng, Shengpeng Wang, Fuhui Zhou, Wei Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06948 2025-11-11 cs.CV 79%

PADM: A Physics-aware Diffusion Model for Attenuation Correction

Trung Kien Pham, Hoang Minh Vu, Anh Duc Chu, Dac Thai Nguyen, Trung Thanh Nguyen, Thao Nguyen Truong, Mai Hong Son, Thanh Trung Nguyen, Phi Le Nguyen

机构 * AI4LIFE, Hanoi University of Science and Technology, Vietnam(AI4LIFE,越南河内科学技术大学) Nagoya Univeristy, Japan(名古屋大学) National Institute of Advanced Industrial Science and Technology, Japan(日本国家先进工业科学和技术研究院) Military Central Hospital, Vietnam(越南108中央军医院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06823 2025-11-11 cs.CV 79%

Integrating Reweighted Least Squares with Plug-and-Play Diffusion Priors for Noisy Image Restoration

Ji Li, Chao Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05308 2025-11-11 cs.CV cs.AI cs.LG 79%

Rethinking Metrics and Diffusion Architecture for 3D Point Cloud Generation

Matteo Bastico, David Ryckelynck, Laurent Corté, Yannick Tillier, Etienne Decencière

机构 * Mines Paris, Université PSL(Mines Paris,Université PSL) Centre des Matériaux (MAT), UMR7633 CNRS(Centre des Matériaux(MAT),UMR7633 CNRS) Centre de Morphologie Mathématique (CMM)(Centre de Morphologie Mathématique(CMM)) Centre de Mise en Forme des Matériaux (CEMEF), UMR7635 CNRS(Centre de Mise en Forme des Matériaux(CEMEF),UMR7635 CNRS)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This paper has been accepted at International Conference on 3D Vision (3DV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17383 2025-11-11 cs.LG cs.CV cs.CY 79%

The Evolving Nature of Latent Spaces: From GANs to Diffusion

Ludovica Schaerf

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Presented and published at Ethics and Aesthetics of Artificial Intelligence Conference (EA-AI'25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08009 2025-11-11 cs.CV cs.AI cs.LG 79%

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Xun Huang, Zhengqi Li, Guande He, Mingyuan Zhou, Eli Shechtman

机构 * Adobe Research(Adobe研究院) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments NeurIPS 2025 spotlight. Project website: http://self-forcing.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18167 2025-11-11 physics.med-ph cs.CV physics.bio-ph 79%

Scattering approach to diffusion quantifies axonal damage in brain injury

Ali Abdollahzadeh, Ricardo Coronado-Leija, Hong-Hsi Lee, Alejandra Sierra, Els Fieremans, Dmitry S. Novikov

机构 * Center for Biomedical Imaging, Department of Radiology, New York University School of Medicine, New York, NY, USA(纽约大学医学学院放射科生物医学成像中心) A.I. Virtanen Institute for Molecular Sciences, University of Eastern Finland, Kuopio, Finland(东部芬兰大学A.I. Virtanen分子科学研究所) Athinoula A. Martinos Center for Biomedical Imaging, Department of Radiology, Massachusetts General Hospital, Harvard Medical School, Boston, MA, USA(马萨诸塞州总医院哈佛医学院生物医学成像中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Nat Commun 16, 9808 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02261 2025-11-11 cs.CV 79%

Diffusion Implicit Policy for Unpaired Scene-aware Motion Synthesis

Jingyu Gong, Chong Zhang, Fengqi Liu, Ke Fan, Qianyu Zhou, Xin Tan, Zhizhong Zhang, Yuan Xie

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06422 2025-11-11 cs.CV 79%

DiffusionUavLoc: Visually Prompted Diffusion for Cross-View UAV Localization

Tao Liu, Kan Ren, Qian Chen

机构 * School of Electronic and Optical Engineering, Nanjing University of Science and Technology(电子与光学工程学院,南京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06310 2025-11-11 cs.CV 79%

Adaptive 3D Reconstruction via Diffusion Priors and Forward Curvature-Matching Likelihood Updates

Seunghyeok Shin, Dabin Kim, Hongki Lim

机构 * Department of Electrical and Computer Engineering, Inha University(电子与计算机工程系,延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06245 2025-11-11 cs.CV 79%

Gait Recognition via Collaborating Discriminative and Generative Diffusion Models

Haijun Xiong, Bin Feng, Bang Wang, Xinggang Wang, Wenyu Liu

机构 * School of EIC, Huazhong University of Science & Technology(电子信息学院,华中科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages, 4figures

详情

展开后加载摘要…

URL PDF HTML 收藏