arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86872 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2506.08009 2025-11-11 cs.CV cs.AI cs.LG 79%

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Xun Huang, Zhengqi Li, Guande He, Mingyuan Zhou, Eli Shechtman

机构 * Adobe Research(Adobe研究院) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments NeurIPS 2025 spotlight. Project website: http://self-forcing.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18167 2025-11-11 physics.med-ph cs.CV physics.bio-ph 79%

Scattering approach to diffusion quantifies axonal damage in brain injury

Ali Abdollahzadeh, Ricardo Coronado-Leija, Hong-Hsi Lee, Alejandra Sierra, Els Fieremans, Dmitry S. Novikov

机构 * Center for Biomedical Imaging, Department of Radiology, New York University School of Medicine, New York, NY, USA(纽约大学医学学院放射科生物医学成像中心) A.I. Virtanen Institute for Molecular Sciences, University of Eastern Finland, Kuopio, Finland(东部芬兰大学A.I. Virtanen分子科学研究所) Athinoula A. Martinos Center for Biomedical Imaging, Department of Radiology, Massachusetts General Hospital, Harvard Medical School, Boston, MA, USA(马萨诸塞州总医院哈佛医学院生物医学成像中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Nat Commun 16, 9808 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02261 2025-11-11 cs.CV 79%

Diffusion Implicit Policy for Unpaired Scene-aware Motion Synthesis

Jingyu Gong, Chong Zhang, Fengqi Liu, Ke Fan, Qianyu Zhou, Xin Tan, Zhizhong Zhang, Yuan Xie

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06422 2025-11-11 cs.CV 79%

DiffusionUavLoc: Visually Prompted Diffusion for Cross-View UAV Localization

Tao Liu, Kan Ren, Qian Chen

机构 * School of Electronic and Optical Engineering, Nanjing University of Science and Technology(电子与光学工程学院,南京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06310 2025-11-11 cs.CV 79%

Adaptive 3D Reconstruction via Diffusion Priors and Forward Curvature-Matching Likelihood Updates

Seunghyeok Shin, Dabin Kim, Hongki Lim

机构 * Department of Electrical and Computer Engineering, Inha University(电子与计算机工程系,延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06245 2025-11-11 cs.CV 79%

Gait Recognition via Collaborating Discriminative and Generative Diffusion Models

Haijun Xiong, Bin Feng, Bang Wang, Xinggang Wang, Wenyu Liu

机构 * School of EIC, Huazhong University of Science & Technology(电子信息学院,华中科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages, 4figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06019 2025-11-11 cs.CV cs.AI 79%

MiVID: Multi-Strategic Self-Supervision for Video Frame Interpolation using Diffusion Model

Priyansh Srivastava, Romit Chatterjee, Abir Sen, Aradhana Behura, Ratnakar Dash

机构 * School of Computer Engineering, KIIT Deemed to be University(计算机工程学院,KIIT被认定大学) Department of Computer Science and Engineering, National Institute of Technology(计算机科学与工程系,国家理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05989 2025-11-11 cs.CV 79%

A Dual-Mode ViT-Conditioned Diffusion Framework with an Adaptive Conditioning Bridge for Breast Cancer Segmentation

Prateek Singh, Moumita Dholey, P. K. Vinod

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 5 pages, 2 figures, 3 tables, submitted to ISBI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05571 2025-11-11 cs.CV cs.AI 79%

C3-Diff: Super-resolving Spatial Transcriptomics via Cross-modal Cross-content Contrastive Diffusion Modelling

Xiaofei Wang, Stephen Price, Chao Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03044 2025-11-11 cs.CV 79%

DCDB: Dynamic Conditional Dual Diffusion Bridge for Ill-posed Multi-Tasks

Chengjie Huang, Jiafeng Yan, Jing Li, Lu Bai

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments The article contains factual errors

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04090 2025-11-11 cs.CV 79%

Bridging Diffusion Models and 3D Representations: A 3D Consistent Super-Resolution Framework

Yi-Ting Chen, Ting-Hsuan Liao, Pengsheng Guo, Alexander Schwing, Jia-Bin Huang

机构 * University of Maryland, College Park(马里兰大学 College Park 分校) Carnegie Mellon University(卡内基梅隆大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICCV 2025. Project website: https://consistent3dsr.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02240 2025-11-11 cs.CV cs.AI 79%

Forecasting When to Forecast: Accelerating Diffusion Models with Confidence-Gated Taylor

Xiaoliu Guan, Lielin Jiang, Hanqi Chen, Xu Zhang, Jiaxing Yan, Guanzhong Wang, Yi Liu, Zetao Zhang, Yu Wu

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) PaddlePaddle Team, Baidu Inc(百度公司PaddlePaddle团队) International Joint Innovation Center, The Electromagnetics Academy at Zhejiang University(浙江大学电磁学学院国际联合创新中心) Yunnan Key Laboratory of Media Convergence(云南媒体融合重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18584 2025-11-11 cs.CV 79%

Unleashing Diffusion Transformers for Visual Correspondence by Modulating Massive Activations

Chaofan Gan, Yuanpeng Tu, Xi Chen, Tieyuan Chen, Yuxi Li, Mehrtash Harandi, Weiyao Lin

机构 * Shanghai Jiao Tong University(上海交通大学) Monash University(墨尔本大学) The University of Hong Kong(香港大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments NeurIPS 2025, code: https://github.com/ganchaofan0000/DiTF

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05007 2025-11-11 cs.CV cs.LG 79%

SVDQuant: Absorbing Outliers by Low-Rank Components for 4-Bit Diffusion Models

Muyang Li, Yujun Lin, Zhekai Zhang, Tianle Cai, Xiuyu Li, Junxian Guo, Enze Xie, Chenlin Meng, Jun-Yan Zhu, Song Han

机构 * MIT(麻省理工学院) NVIDIA(NVIDIA公司) CMU(卡内基梅隆大学) Princeton(普林斯顿大学) UC Berkeley(加州大学伯克利分校) SJTU(上海交通大学) Pika Labs(Pika实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICLR 2025 Spotlight Quantization Library: https://github.com/mit-han-lab/deepcompressor Inference Engine: https://github.com/mit-han-lab/nunchaku Website: https://hanlab.mit.edu/projects/svdquant Demo: https://demo.nunchaku.tech/ Blog: https://hanlab.mit.edu/blog/svdquant

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.19604 2025-11-11 eess.IV cs.CV 79%

X-Diffusion: Generating Detailed 3D MRI Volumes From a Single Image Using Cross-Sectional Diffusion Models

Emmanuelle Bourigault, Abdullah Hamdi, Amir Jamaludin

机构 * Visual Geometry Group, University of Oxford(牛津大学视觉几何组)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments accepted at ICCV 2025 GAIA workshop https://era-ai-biomed.github.io/GAIA/ , project website: https://emmanuelleb985.github.io/XDiffusion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04963 2025-11-10 cs.CV cs.AI 79%

Pattern-Aware Diffusion Synthesis of fMRI/dMRI with Tissue and Microstructural Refinement

Xiongri Shen, Jiaqi Wang, Yi Zhong, Zhenxi Song, Leilei Zhao, Yichen Wei, Lingyan Liang, Shuqiang Wang, Baiying Lei, Demao Deng, Zhiguo Zhang

机构 * Department of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术系,哈尔滨工业大学) School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology(智能科学与工程学院,人工智能学院,哈尔滨工业大学) School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical, Measurements and Ultrasound Imaging, Shenzhen University Medical School, Shenzhen University, Shenzhen, China(生物医学工程学院,医学超声关键技术工程实验室,广东生物医学、测量与超声成像重点实验室,深圳大学医学院,深圳大学,深圳,中国) Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院) Department of Radiology, The People’s Hospital of Guangxi Zhuang Autonomous Region, Guangxi Academy of Medical Sciences(放射科,广西壮族自治区人民医院,广西医学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04871 2025-11-10 cs.CV stat.AP 79%

Clinical-ComBAT: a diffusion-weighted MRI harmonization method for clinical applications

Gabriel Girard, Manon Edde, Félix Dumais, Yoan David, Matthieu Dumont, Guillaume Theaud, Jean-Christophe Houde, Arnaud Boré, Maxime Descoteaux, Pierre-Marc Jodoin

机构 * Videos & Images Theory and Analytics Lab (VITAL)(视频与图像理论与分析实验室) Department of Computer Science(计算机科学系) Université de Sherbrooke(Sherbrooke大学) Sherbrooke Connectivity Imaging Lab (SCIL)(Sherbrooke连接成像实验室) Imeka Solutions inc(Imeka Solutions公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 39 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17013 2025-11-10 cs.LG cs.CV 79%

When Are Concepts Erased From Diffusion Models?

Kevin Lu, Nicky Kriplani, Rohit Gandikota, Minh Pham, David Bau, Chinmay Hegde, Niv Cohen

机构 * Northeastern University(东北大学) New York University(纽约大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 2025. Our code, data, and results are available at https://unerasing.baulab.info/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11772 2025-11-06 cs.CV cs.LG 79%

CLIP Meets Diffusion: A Synergistic Approach to Anomaly Detection

Byeongchan Lee, John Won, Seunghyun Lee, Jinwoo Shin

机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国先进科学技术研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at TMLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10036 2025-11-06 cs.GR cs.CL 79%

Token Perturbation Guidance for Diffusion Models

Javad Rajabi, Soroush Mehraban, Seyedmorteza Sadat, Babak Taati

机构 * University of Toronto(多伦多大学) Vector Institute for Artificial Intelligence(人工智能向量研究所) KITE Research Institute(KITE研究机构) ETH Zürich(苏黎世联邦理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

Comments Accepted at NeurIPS 2025. Project page: https://github.com/TaatiTeam/Token-Perturbation-Guidance

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03219 2025-11-06 cs.CV 79%

Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation

Pengyu Jie, Wanquan Liu, Rui He, Yihui Wen, Deyu Meng, Chenqiang Gao

机构 * School of Intelligent Engineering, Sun Yat-sen University (Shenzhen Campus)(中山大学智能工程学院) Department of Otolaryngology, The First Affiliated Hospital of Sun Yat-sen University(中山大学附属第一医院耳鼻喉科) School of Mathematics and Statistics, Xi’an Jiaotong University(西安交通大学数学与统计学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10687 2025-11-06 cs.CV 79%

Stable Part Diffusion 4D: Multi-View RGB and Kinematic Parts Video Generation

Hao Zhang, Chun-Han Yao, Simon Donné, Narendra Ahuja, Varun Jampani

机构 * Stability AI University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Page: https://stablepartdiffusion4d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15737 2025-11-06 cs.LG cs.CV 79%

Probability Density from Latent Diffusion Models for Out-of-Distribution Detection

Joonas Järve, Karl Kaspar Haavel, Meelis Kull

机构 * Institute of Computer Science, University of Tartu, Estonia(计算机科学研究所,塔尔图大学,爱沙尼亚)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ECAI 2025

Journal ref Frontiers in Artificial Intelligence and Applications 413 (ECAI 2025) 5027 - 5034

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09054 2025-11-06 eess.IV cs.GR 79%

NeurOp-Diff:Continuous Remote Sensing Image Super-Resolution via Neural Operator Diffusion

Zihao Xu, Yuzhi Tang, Bowen Xu, Qingquan Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02478 2025-11-05 cs.MM cs.AI 79%

Wireless Video Semantic Communication with Decoupled Diffusion Multi-frame Compensation

Bingyan Xie, Yongpeng Wu, Yuxuan Shi, Biqian Feng, Wenjun Zhang, Jihong Park, Tony Quek

机构 * Department of Electronic Engineering, Shanghai Jiao Tong University(电子工程系,上海交通大学) School of Cyber and Engineering, Shanghai Jiao Tong University(网络与工程学院,上海交通大学) ISTD Pillar, Singapore University of Technology of Design(新加坡设计大学技术支柱)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.12973 2025-11-05 eess.IV cs.CV cs.LG q-bio.QM 79%

Cross-modal Diffusion Modelling for Super-resolved Spatial Transcriptomics

Xiaofei Wang, Xingxu Huang, Stephen J. Price, Chao Li

机构 * Department of Clinical Neurosciences, University of Cambridge, UK(剑桥大学临床神经科学系) Department of Applied Mathematics and Theoretical Physics, University of Cambridge, UK(剑桥大学应用数学与理论物理系) School of Science and Engineering, University of Dundee, UK(邓迪大学科学与工程学院) School of Medicine, University of Dundee, UK(邓迪大学医学院) Zhejiang Lab, China(浙江实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01795 2025-11-04 cs.LG cs.AI cs.CV cs.RO stat.ML 79%

Fractional Diffusion Bridge Models

Gabriel Nobis, Maximilian Springenberg, Arina Belova, Rembert Daems, Christoph Knochenhauer, Manfred Opper, Tolga Birdal, Wojciech Samek

机构 * Fraunhofer HHI(弗劳恩霍夫研究所) Ghent University–imec FlandersMake–MIRO(根特大学–imec 荷兰弗拉芒Make–MIRO) Technical University of Munich(慕尼黑技术大学) Technical University of Berlin(柏林技术大学) University of Potsdam(波茨坦大学) University of Birmingham(伯明翰大学) Imperial College London(伦敦帝国理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments To appear in NeurIPS 2025 proceedings. This version includes post-camera-ready revisions

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01704 2025-11-04 cs.CV 79%

Learnable Fractional Reaction-Diffusion Dynamics for Under-Display ToF Imaging and Beyond

Xin Qiao, Matteo Poggi, Xing Wei, Pengchao Deng, Yanhui Zhou, Stefano Mattoccia

机构 * Xi’an Jiaotong University(西安交通大学) University of Bologna(博洛尼亚大学) Anyang Institute of Technology(安阳职业技术学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01307 2025-11-04 cs.CV cs.AI 79%

Perturb a Model, Not an Image: Towards Robust Privacy Protection via Anti-Personalized Diffusion Models

Tae-Young Lee, Juwon Seo, Jong Hwan Ko, Gyeong-Moon Park

机构 * Korea University(韩国大学) Kyung Hee University(庆熙大学) Sungkyunkwan University(成均馆大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 26 pages, 9 figures, 16 tables, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13030 2025-11-04 cs.CV 79%

WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild

Morris Alper, David Novotny, Filippos Kokkinos, Hadar Averbuch-Elor, Tom Monnier

机构 * Tel Aviv University(特拉维夫大学) Meta AI Cornell University(康奈尔大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 2025. Project page: https://wildcat3d.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏