arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2507.00687 2025-07-02 cs.LG cs.CV 79%

Diffusion Classifier Guidance for Non-robust Classifiers

Philipp Vaeth, Dibyanshu Kumar, Benjamin Paassen, Magda Gregorová

机构 * Center for Artificial Intelligence and Robotics(人工智能与机器人中心) Technical University of Applied Sciences Würzburg-Schweinfurt(应用科学大学魏玛-施韦因富特分校) Bielefeld University(比勒菲尔德大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at ECML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12747 2025-07-02 cs.CV cs.AI 79%

Unleashing Diffusion and State Space Models for Medical Image Segmentation

Rong Wu, Ziqi Chen, Liming Zhong, Heng Li, Hai Shu

机构 * Department of Biostatistics, School of Global Public Health, New York University(生物统计学系、全球公共卫生学院、纽约大学) School of Statistics, KLATASDS-MOE, East China Normal University(统计学系、KLATASDS-MOE、华东师范大学) School of Biomedical Engineering, Southern Medical University(生物医学工程学院、南方医科大学) Faculty of Biomedical Engineering, Shenzhen University of Advanced Technology(生物医学工程学院、深圳先进技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10819 2025-07-02 cs.CV cs.LG 79%

GAUDA: Generative Adaptive Uncertainty-guided Diffusion-based Augmentation for Surgical Segmentation

Yannik Frisch, Christina Bornberg, Moritz Fuchs, Anirban Mukhopadhyay

机构 * Technical University Darmstadt(德累斯顿技术大学) University Medical Center Mainz(马因茨大学医学中心) University of Girona(加泰罗尼亚大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04814 2025-07-02 cs.CV cs.LG 79%

Lifelong Learning of Video Diffusion Models From a Single Video Stream

Jason Yoo, Yingchen He, Saeid Naderiparizi, Dylan Green, Gido M. van de Ven, Geoff Pleiss, Frank Wood

机构 * University of British Columbia(不列颠哥伦比亚大学) KU Leuven(鲁文大学) Vector Institute(向量研究所) Amii(阿米研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Video samples are available here: https://drive.google.com/drive/folders/1CsmWqug-CS7I6NwGDvHsEN9FqN2QzspN

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.02587 2025-07-02 eess.IV cs.CV cs.LG 79%

Synthesising Rare Cataract Surgery Samples with Guided Diffusion Models

Yannik Frisch, Moritz Fuchs, Antoine Sanner, Felix Anton Ucar, Marius Frenzel, Joana Wasielica-Poslednik, Adrian Gericke, Felix Mathias Wagner, Thomas Dratsch, Anirban Mukhopadhyay

机构 * Technical University Darmstadt(德累斯顿技术大学) Universitätsmedizin der Johannes Gutenberg-Universität Mainz(莱茵-美因大学医学中心) Uniklinik Köln(科隆大学医院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.24113 2025-07-01 cs.CV 79%

Epona: Autoregressive Diffusion World Model for Autonomous Driving

Kaiwen Zhang, Zhenyu Tang, Xiaotao Hu, Xingang Pan, Xiaoyang Guo, Yuan Liu, Jingwei Huang, Li Yuan, Qian Zhang, Xiao-Xiao Long, Xun Cao, Wei Yin

机构 * Horizon Robotics Tsinghua University(清华大学) Peking University(北京大学) Nanjing University(南京大学) The Hong Kong University of Science and Technology(香港科学与技术大学) Nanyang Technological University(南洋理工大学) Tencent Hunyuan(腾讯文心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV2025, Project Page: https://kevin-thu.github.io/Epona/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20073 2025-07-01 eess.IV cs.CV cs.LG physics.med-ph physics.optics 79%

Pixel super-resolved virtual staining of label-free tissue using diffusion models

Yijie Zhang, Luzhe Huang, Nir Pillar, Yuzhu Li, Hanlong Chen, Aydogan Ozcan

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 39 Pages, 7 Figures

Journal ref Nature Communications (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23858 2025-07-01 cs.CV 79%

VMoBA: Mixture-of-Block Attention for Video Diffusion Models

Jianzong Wu, Liang Hou, Haotian Yang, Xin Tao, Ye Tian, Pengfei Wan, Di Zhang, Yunhai Tong

机构 * Peking University(北京大学) Kling Team, Kuaishou Technology(快手科技 Kling 团队)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Code is at https://github.com/KwaiVGI/VMoBA

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23701 2025-07-01 eess.IV cs.CV 79%

MDPG: Multi-domain Diffusion Prior Guidance for MRI Reconstruction

Lingtong Zhang, Mengdie Song, Xiaohan Hao, Huayu Mai, Bensheng Qiu

机构 * School of Information Science and Technology(信息科学与技术学院) University of Science and Technology of China(中国科学技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accept by MICCAI2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23606 2025-07-01 cs.CV 79%

SG-LDM: Semantic-Guided LiDAR Generation via Latent-Aligned Diffusion

Zhengkang Xiang, Zizhao Li, Amir Khodabandeh, Kourosh Khoshelham

机构 * The University of Melbourne(墨尔本大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23566 2025-07-01 cs.CV cs.LG 79%

Metadata, Wavelet, and Time Aware Diffusion Models for Satellite Image Super Resolution

Luigi Sigillo, Renato Giamba, Danilo Comminiello

机构 * Dept. of Information Engineering, Electronics, and Telecomm., Sapienza University of Rome(信息工程、电子与电信系,罗马萨皮恩扎大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICLR 2025 Workshop on Machine Learning for Remote Sensing (ML4RS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23513 2025-07-01 cs.CV 79%

ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models

Zixun Fang, Kai Zhu, Zhiheng Liu, Yu Liu, Wei Zhai, Yang Cao, Zheng-Jun Zha

机构 * USTC(中国科学技术大学) TongYi Lab(通义实验室) HKU(香港大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments https://becauseimbatman0.github.io/ViewPoint

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23184 2025-07-01 eess.IV cs.AI cs.CV 79%

Score-based Diffusion Model for Unpaired Virtual Histology Staining

Anran Liu, Xiaofei Wang, Jing Cai, Chao Li

机构 * Department of Health Technology and Informatics, Hong Kong Polytechnic University, China(香港理工大学健康科技与信息学系) Department of Clinical Neurosciences, University of Cambridge, UK(剑桥大学临床神经科学系) School of Science and Engineering, University of Dundee, UK(邓迪大学科学与工程学院) Department of Applied Mathematics and Theoretical Physics, University of Cambridge, UK(剑桥大学应用数学与理论物理系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 11 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22882 2025-07-01 eess.IV cs.CV cs.LG 79%

CA-Diff: Collaborative Anatomy Diffusion for Brain Tissue Segmentation

Qilong Xing, Zikai Song, Yuteng Ye, Yuke Chen, Youjia Zhang, Na Feng, Junqing Yu, Wei Yang

机构 * School of Computer Science and Technology(计算机科学与技术学院) Huazhong University of Science and Technology(华中科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICME 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22753 2025-07-01 cs.CV 79%

Degradation-Modeled Multipath Diffusion for Tunable Metalens Photography

Jianing Zhang, Jiayi Zhu, Feiyu Ji, Xiaokang Yang, Xiaoyun Yuan

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21042 2025-07-01 cs.CV 79%

Boosting Domain Generalized and Adaptive Detection with Diffusion Models: Fitness, Generalization, and Transferability

Boyong He, Yuxiang Ji, Zhuoyue Tan, Liaoni Wu

机构 * Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能学院) School of Aerospace Engineering, Xiamen University(厦门大学航空航天工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20832 2025-07-01 cs.CV cs.AI 79%

Leveraging Vision-Language Models to Select Trustworthy Super-Resolution Samples Generated by Diffusion Models

Cansu Korkmaz, Ahmet Murat Tekalp, Zafer Dogan

机构 * Department of Electrical and Electronics Engineering and KUIS AI Center, Koç University(电气与电子工程系和KUIS人工智能中心,科克大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages, 9 figures, 5 tables, accepted to IEEE Transactions on Circuits and Systems for Video Technology

Journal ref IEEE Transactions on Circuits and Systems for Video Technology 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09185 2025-07-01 cs.CV 79%

Incomplete Multi-view Clustering via Diffusion Contrastive Generation

Yuanyang Zhang, Yijie Lin, Weiqing Yan, Li Yao, Xinhang Wan, Guangyuan Li, Chao Zhang, Guanzhou Ke, Jie Xu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12080 2025-07-01 cs.CV 79%

HumanGif: Single-View Human Diffusion with Generative Prior

Shoukang Hu, Takuya Narihira, Kazumi Fukuda, Ryosuke Sawata, Takashi Shibuya, Yuki Mitsufuji

机构 * Sony AI(索尼人工智能实验室) Sony Group Corporation(索尼集团)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://skhu101.github.io/HumanGif/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.04444 2025-07-01 cs.CV 79%

Disentangled Diffusion-Based 3D Human Pose Estimation with Hierarchical Spatial and Temporal Denoiser

Qingyuan Cai, Xuecai Hu, Saihui Hou, Li Yao, Yongzhen Huang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by AAAI24

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.15952 2025-07-01 cs.LG cs.CV 79%

Improving Robustness and Reliability in Medical Image Classification with Latent-Guided Diffusion and Nested-Ensembles

Xing Shen, Hengguan Huang, Brennan Nichyporuk, Tal Arbel

机构 * McGill University(麦吉尔大学) Centre for Intelligent Machines(智能机器中心) Mila – Quebec AI Institute(魁北克人工智能研究所) University of Copenhagen(哥本哈根大学) Department of Public Health, Section for Health Data Science and AI(公共卫生系,健康数据科学与人工智能部门)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to IEEE Transactions on Medical Imaging, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22509 2025-07-01 cs.CV cs.AI 79%

FreeDNA: Endowing Domain Adaptation of Diffusion-Based Dense Prediction with Training-Free Domain Noise Alignment

Hang Xu, Jie Huang, Linjiang Huang, Dong Li, Yidi Liu, Feng Zhao

机构 * University of Science and Technology of China(中国科学技术大学) Beihang University(北航大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22298 2025-06-30 cs.CV 79%

OutDreamer: Video Outpainting with a Diffusion Transformer

Linhao Zhong, Fan Li, Yi Huang, Jianzhuang Liu, Renjing Pei, Fenglong Song

机构 * Shanghai Jiao Tong University(上海交通大学) Huawei Noah’s Ark Lab(华为诺亚实验室) Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22012 2025-06-30 eess.IV cs.CV 79%

Noise-Inspired Diffusion Model for Generalizable Low-Dose CT Reconstruction

Qi Gao, Zhihao Chen, Dong Zeng, Junping Zhang, Jianhua Ma, Hongming Shan

机构 * Institute of Science and Technology for Brain-inspired Intelligence(脑启发智能科学与技术研究院) Fudan University(复旦大学) School of Biomedical Engineering(生物医学工程学院) Southern Medical University(南方医科大学) School of Computer Science(计算机科学学院) Xi’an Jiaotong University(西安交通大学) School of Life Science and Technology(生命科学与技术学院) MOE Frontiers Center for Brain Science(教育部脑科学前沿中心) Key Laboratory of Computational Neuroscience and Brain-Inspired Intelligence (Ministry of Education)(教育部计算神经科学与脑启发智能重点实验室) State Key Laboratory of Brain Function and Disorders(脑功能与疾病国家重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted for publication in Medical Image Analysis, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21748 2025-06-30 physics.optics cs.CV cs.LG 79%

Inverse Design of Diffractive Metasurfaces Using Diffusion Models

Liav Hen, Erez Yosef, Dan Raviv, Raja Giryes, Jacob Scheuer

机构 * School of Electrical and Computer Engineering(电气与计算机工程学院) Tel-Aviv University(特拉维夫大学) The Center for Nanosciences and Nanotechnology(纳米科学与技术中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21722 2025-06-30 cs.CV cs.AI 79%

Elucidating and Endowing the Diffusion Training Paradigm for General Image Restoration

Xin Lu, Xueyang Fu, Jie Xiao, Zihao Fan, Yurui Zhu, Zheng-Jun Zha

机构 * MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, School of Information Science and Technology, University of Science and Technology of China(脑启发智能感知与认知联合实验室,信息科学与技术学院,中国科学技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07381 2025-06-30 cs.CV 79%

Spatial Degradation-Aware and Temporal Consistent Diffusion Model for Compressed Video Super-Resolution

Hongyu An, Xinfeng Zhang, Shijie Zhao, Li Zhang, Ruiqin Xiong

机构 * School of Computer Science and Technology, University of Chinese Academy of Sciences(中国科学院大学计算机科学与技术学院) ByteDance Inc.(字节跳动公司) Institute of Digital Media, School of Electronic Engineering and Computer Science, Peking University(北京大学电子工程与计算机科学学院数字媒体研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21287 2025-06-27 cs.CV 79%

HieraSurg: Hierarchy-Aware Diffusion Model for Surgical Video Generation

Diego Biagini, Nassir Navab, Azade Farshad

机构 * Chair for Computer Aided Medical Procedures (CAMP), TU Munich, Germany(计算机辅助医疗程序研究所(CAMP),慕尼黑技术大学,德国) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19391 2025-06-27 cs.CV 79%

Generate the Forest before the Trees -- A Hierarchical Diffusion model for Climate Downscaling

Declan J. Curran, Sanaa Hobeichi, Hira Saleem, Hao Xue, Flora D. Salim

机构 * School of Computer Science and Engineering, University of New South Wales(计算机科学与工程学院,新南威尔士大学) ARC Centre of Excellence for the Weather of the 21 st Century and Climate Change Research Centre, University of New South Wales(21世纪天气卓越研究中心及气候变化研究中心,新南威尔士大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09229 2025-06-26 cs.CV 79%

Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models

Sungwon Hwang, Hyojin Jang, Kinam Kim, Minho Park, Jaegul Choo

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://crepavideo.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏