arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70196 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70196 篇

2511.06019 2025-11-11 cs.CV cs.AI 79%

MiVID: Multi-Strategic Self-Supervision for Video Frame Interpolation using Diffusion Model

Priyansh Srivastava, Romit Chatterjee, Abir Sen, Aradhana Behura, Ratnakar Dash

机构 * School of Computer Engineering, KIIT Deemed to be University(计算机工程学院,KIIT被认定大学) Department of Computer Science and Engineering, National Institute of Technology(计算机科学与工程系,国家理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05989 2025-11-11 cs.CV 79%

A Dual-Mode ViT-Conditioned Diffusion Framework with an Adaptive Conditioning Bridge for Breast Cancer Segmentation

Prateek Singh, Moumita Dholey, P. K. Vinod

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 5 pages, 2 figures, 3 tables, submitted to ISBI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05571 2025-11-11 cs.CV cs.AI 79%

C3-Diff: Super-resolving Spatial Transcriptomics via Cross-modal Cross-content Contrastive Diffusion Modelling

Xiaofei Wang, Stephen Price, Chao Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03044 2025-11-11 cs.CV 79%

DCDB: Dynamic Conditional Dual Diffusion Bridge for Ill-posed Multi-Tasks

Chengjie Huang, Jiafeng Yan, Jing Li, Lu Bai

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments The article contains factual errors

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04090 2025-11-11 cs.CV 79%

Bridging Diffusion Models and 3D Representations: A 3D Consistent Super-Resolution Framework

Yi-Ting Chen, Ting-Hsuan Liao, Pengsheng Guo, Alexander Schwing, Jia-Bin Huang

机构 * University of Maryland, College Park(马里兰大学 College Park 分校) Carnegie Mellon University(卡内基梅隆大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICCV 2025. Project website: https://consistent3dsr.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02240 2025-11-11 cs.CV cs.AI 79%

Forecasting When to Forecast: Accelerating Diffusion Models with Confidence-Gated Taylor

Xiaoliu Guan, Lielin Jiang, Hanqi Chen, Xu Zhang, Jiaxing Yan, Guanzhong Wang, Yi Liu, Zetao Zhang, Yu Wu

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) PaddlePaddle Team, Baidu Inc(百度公司PaddlePaddle团队) International Joint Innovation Center, The Electromagnetics Academy at Zhejiang University(浙江大学电磁学学院国际联合创新中心) Yunnan Key Laboratory of Media Convergence(云南媒体融合重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18584 2025-11-11 cs.CV 79%

Unleashing Diffusion Transformers for Visual Correspondence by Modulating Massive Activations

Chaofan Gan, Yuanpeng Tu, Xi Chen, Tieyuan Chen, Yuxi Li, Mehrtash Harandi, Weiyao Lin

机构 * Shanghai Jiao Tong University(上海交通大学) Monash University(墨尔本大学) The University of Hong Kong(香港大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments NeurIPS 2025, code: https://github.com/ganchaofan0000/DiTF

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05007 2025-11-11 cs.CV cs.LG 79%

SVDQuant: Absorbing Outliers by Low-Rank Components for 4-Bit Diffusion Models

Muyang Li, Yujun Lin, Zhekai Zhang, Tianle Cai, Xiuyu Li, Junxian Guo, Enze Xie, Chenlin Meng, Jun-Yan Zhu, Song Han

机构 * MIT(麻省理工学院) NVIDIA(NVIDIA公司) CMU(卡内基梅隆大学) Princeton(普林斯顿大学) UC Berkeley(加州大学伯克利分校) SJTU(上海交通大学) Pika Labs(Pika实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICLR 2025 Spotlight Quantization Library: https://github.com/mit-han-lab/deepcompressor Inference Engine: https://github.com/mit-han-lab/nunchaku Website: https://hanlab.mit.edu/projects/svdquant Demo: https://demo.nunchaku.tech/ Blog: https://hanlab.mit.edu/blog/svdquant

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.19604 2025-11-11 eess.IV cs.CV 79%

X-Diffusion: Generating Detailed 3D MRI Volumes From a Single Image Using Cross-Sectional Diffusion Models

Emmanuelle Bourigault, Abdullah Hamdi, Amir Jamaludin

机构 * Visual Geometry Group, University of Oxford(牛津大学视觉几何组)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments accepted at ICCV 2025 GAIA workshop https://era-ai-biomed.github.io/GAIA/ , project website: https://emmanuelleb985.github.io/XDiffusion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04963 2025-11-10 cs.CV cs.AI 79%

Pattern-Aware Diffusion Synthesis of fMRI/dMRI with Tissue and Microstructural Refinement

Xiongri Shen, Jiaqi Wang, Yi Zhong, Zhenxi Song, Leilei Zhao, Yichen Wei, Lingyan Liang, Shuqiang Wang, Baiying Lei, Demao Deng, Zhiguo Zhang

机构 * Department of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术系,哈尔滨工业大学) School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology(智能科学与工程学院,人工智能学院,哈尔滨工业大学) School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical, Measurements and Ultrasound Imaging, Shenzhen University Medical School, Shenzhen University, Shenzhen, China(生物医学工程学院,医学超声关键技术工程实验室,广东生物医学、测量与超声成像重点实验室,深圳大学医学院,深圳大学,深圳,中国) Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院) Department of Radiology, The People’s Hospital of Guangxi Zhuang Autonomous Region, Guangxi Academy of Medical Sciences(放射科,广西壮族自治区人民医院,广西医学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04871 2025-11-10 cs.CV stat.AP 79%

Clinical-ComBAT: a diffusion-weighted MRI harmonization method for clinical applications

Gabriel Girard, Manon Edde, Félix Dumais, Yoan David, Matthieu Dumont, Guillaume Theaud, Jean-Christophe Houde, Arnaud Boré, Maxime Descoteaux, Pierre-Marc Jodoin

机构 * Videos & Images Theory and Analytics Lab (VITAL)(视频与图像理论与分析实验室) Department of Computer Science(计算机科学系) Université de Sherbrooke(Sherbrooke大学) Sherbrooke Connectivity Imaging Lab (SCIL)(Sherbrooke连接成像实验室) Imeka Solutions inc(Imeka Solutions公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 39 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17013 2025-11-10 cs.LG cs.CV 79%

When Are Concepts Erased From Diffusion Models?

Kevin Lu, Nicky Kriplani, Rohit Gandikota, Minh Pham, David Bau, Chinmay Hegde, Niv Cohen

机构 * Northeastern University(东北大学) New York University(纽约大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 2025. Our code, data, and results are available at https://unerasing.baulab.info/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11772 2025-11-06 cs.CV cs.LG 79%

CLIP Meets Diffusion: A Synergistic Approach to Anomaly Detection

Byeongchan Lee, John Won, Seunghyun Lee, Jinwoo Shin

机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国先进科学技术研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at TMLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10036 2025-11-06 cs.GR cs.CL 79%

Token Perturbation Guidance for Diffusion Models

Javad Rajabi, Soroush Mehraban, Seyedmorteza Sadat, Babak Taati

机构 * University of Toronto(多伦多大学) Vector Institute for Artificial Intelligence(人工智能向量研究所) KITE Research Institute(KITE研究机构) ETH Zürich(苏黎世联邦理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

Comments Accepted at NeurIPS 2025. Project page: https://github.com/TaatiTeam/Token-Perturbation-Guidance

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03219 2025-11-06 cs.CV 79%

Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation

Pengyu Jie, Wanquan Liu, Rui He, Yihui Wen, Deyu Meng, Chenqiang Gao

机构 * School of Intelligent Engineering, Sun Yat-sen University (Shenzhen Campus)(中山大学智能工程学院) Department of Otolaryngology, The First Affiliated Hospital of Sun Yat-sen University(中山大学附属第一医院耳鼻喉科) School of Mathematics and Statistics, Xi’an Jiaotong University(西安交通大学数学与统计学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10687 2025-11-06 cs.CV 79%

Stable Part Diffusion 4D: Multi-View RGB and Kinematic Parts Video Generation

Hao Zhang, Chun-Han Yao, Simon Donné, Narendra Ahuja, Varun Jampani

机构 * Stability AI University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Page: https://stablepartdiffusion4d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15737 2025-11-06 cs.LG cs.CV 79%

Probability Density from Latent Diffusion Models for Out-of-Distribution Detection

Joonas Järve, Karl Kaspar Haavel, Meelis Kull

机构 * Institute of Computer Science, University of Tartu, Estonia(计算机科学研究所,塔尔图大学,爱沙尼亚)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ECAI 2025

Journal ref Frontiers in Artificial Intelligence and Applications 413 (ECAI 2025) 5027 - 5034

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09054 2025-11-06 eess.IV cs.GR 79%

NeurOp-Diff:Continuous Remote Sensing Image Super-Resolution via Neural Operator Diffusion

Zihao Xu, Yuzhi Tang, Bowen Xu, Qingquan Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02478 2025-11-05 cs.MM cs.AI 79%

Wireless Video Semantic Communication with Decoupled Diffusion Multi-frame Compensation

Bingyan Xie, Yongpeng Wu, Yuxuan Shi, Biqian Feng, Wenjun Zhang, Jihong Park, Tony Quek

机构 * Department of Electronic Engineering, Shanghai Jiao Tong University(电子工程系,上海交通大学) School of Cyber and Engineering, Shanghai Jiao Tong University(网络与工程学院,上海交通大学) ISTD Pillar, Singapore University of Technology of Design(新加坡设计大学技术支柱)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.12973 2025-11-05 eess.IV cs.CV cs.LG q-bio.QM 79%

Cross-modal Diffusion Modelling for Super-resolved Spatial Transcriptomics

Xiaofei Wang, Xingxu Huang, Stephen J. Price, Chao Li

机构 * Department of Clinical Neurosciences, University of Cambridge, UK(剑桥大学临床神经科学系) Department of Applied Mathematics and Theoretical Physics, University of Cambridge, UK(剑桥大学应用数学与理论物理系) School of Science and Engineering, University of Dundee, UK(邓迪大学科学与工程学院) School of Medicine, University of Dundee, UK(邓迪大学医学院) Zhejiang Lab, China(浙江实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01795 2025-11-04 cs.LG cs.AI cs.CV cs.RO stat.ML 79%

Fractional Diffusion Bridge Models

Gabriel Nobis, Maximilian Springenberg, Arina Belova, Rembert Daems, Christoph Knochenhauer, Manfred Opper, Tolga Birdal, Wojciech Samek

机构 * Fraunhofer HHI(弗劳恩霍夫研究所) Ghent University–imec FlandersMake–MIRO(根特大学–imec 荷兰弗拉芒Make–MIRO) Technical University of Munich(慕尼黑技术大学) Technical University of Berlin(柏林技术大学) University of Potsdam(波茨坦大学) University of Birmingham(伯明翰大学) Imperial College London(伦敦帝国理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments To appear in NeurIPS 2025 proceedings. This version includes post-camera-ready revisions

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01704 2025-11-04 cs.CV 79%

Learnable Fractional Reaction-Diffusion Dynamics for Under-Display ToF Imaging and Beyond

Xin Qiao, Matteo Poggi, Xing Wei, Pengchao Deng, Yanhui Zhou, Stefano Mattoccia

机构 * Xi’an Jiaotong University(西安交通大学) University of Bologna(博洛尼亚大学) Anyang Institute of Technology(安阳职业技术学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01307 2025-11-04 cs.CV cs.AI 79%

Perturb a Model, Not an Image: Towards Robust Privacy Protection via Anti-Personalized Diffusion Models

Tae-Young Lee, Juwon Seo, Jong Hwan Ko, Gyeong-Moon Park

机构 * Korea University(韩国大学) Kyung Hee University(庆熙大学) Sungkyunkwan University(成均馆大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 26 pages, 9 figures, 16 tables, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13030 2025-11-04 cs.CV 79%

WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild

Morris Alper, David Novotny, Filippos Kokkinos, Hadar Averbuch-Elor, Tom Monnier

机构 * Tel Aviv University(特拉维夫大学) Meta AI Cornell University(康奈尔大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 2025. Project page: https://wildcat3d.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.18932 2025-11-04 cs.CV 79%

ReviveDiff: A Universal Diffusion Model for Restoring Images in Adverse Weather Conditions

Wenfeng Huang, Guoan Xu, Wenjing Jia, Stuart Perry, Guangwei Gao

机构 * Faculty of Engineering and Information Technology, University of Technology Sydney(工程与信息技术学院,技术悉尼大学) PCA Lab, Key Lab of Intelligent Perception and Systems for High-Dimensional Information of Ministry of Education, School of Computer Science and Engineering, Nanjing University of Science and Technology(智能感知与高维信息系统教育部重点实验室,计算机科学与工程学院,南京理工大学) Key Laboratory of Artificial Intelligence, Ministry of Education(人工智能教育部重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00858 2025-11-04 cs.CV cs.AI cs.LG 79%

Occlusion-Aware Diffusion Model for Pedestrian Intention Prediction

Yu Liu, Zhijie Liu, Zedong Yang, You-Fu Li, He Kong

机构 * Shenzhen Key Laboratory of Control Theory and Intelligent Systems, Southern University of Science and Technology (SUSTech)(控制理论与智能系统深圳重点实验室,南方科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This manuscript has been accepted to the IEEE Transactions on Intelligent Transportation Systems as a regular paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00815 2025-11-04 cs.CV 79%

TA-LSDiff:Topology-Aware Diffusion Guided by a Level Set Energy for Pancreas Segmentation

Yue Gou, Fanghui Song, Yuming Xing, Shengzhu Shi, Zhichang Guo, Boying Wu

机构 * Department of Computational Mathematics, School of Mathematics, Harbin Institute of Technology(计算数学系,数学学院,哈尔滨工业大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00344 2025-11-04 cs.CV 79%

Federated Dialogue-Semantic Diffusion for Emotion Recognition under Incomplete Modalities

Xihang Qiu, Jiarong Cheng, Yuhao Fang, Wanpeng Zhang, Yao Lu, Ye Zhang, Chun Li

机构 * Shenzhen MSU-BIT University(深圳MSU-BIT大学) Beijing Institude of Technology(北京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24770 2025-11-04 eess.IV cs.AI cs.CV 79%

DMVFC: Deep Learning Based Functionally Consistent Tractography Fiber Clustering Using Multimodal Diffusion MRI and Functional MRI

Bocheng Guo, Jin Wang, Yijie Li, Junyi Wang, Mingyu Gao, Puming Feng, Yuqian Chen, Jarrett Rushmore, Nikos Makris, Yogesh Rathi, Lauren J O'Donnell, Fan Zhang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01161 2025-11-04 cs.LG cs.CV 79%

Split Gibbs Discrete Diffusion Posterior Sampling

Wenda Chu, Zihui Wu, Yifan Chen, Yang Song, Yisong Yue

机构 * California Institute of Technology(加州理工学院) New York University(纽约大学) OpenAI

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏