arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2508.07146 2025-08-12 cs.CV cs.AI 79%

Intention-Aware Diffusion Model for Pedestrian Trajectory Prediction

Yu Liu, Zhijie Liu, Xiao Ren, You-Fu Li, He Kong

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07006 2025-08-12 eess.IV cs.CV 79%

Spatio-Temporal Conditional Diffusion Models for Forecasting Future Multiple Sclerosis Lesion Masks Conditioned on Treatments

Gian Mario Favero, Ge Ya Luo, Nima Fathi, Justin Szeto, Douglas L. Arnold, Brennan Nichyporuk, Chris Pal, Tal Arbel

机构 * McGill University(麦吉尔大学) Mila – Quebec AI Institute(魁北克人工智能研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to MICCAI 2025 (LMID Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03295 2025-08-12 cs.CV 79%

CPKD: Clinical Prior Knowledge-Constrained Diffusion Models for Surgical Phase Recognition in Endoscopic Submucosal Dissection

Xiangning Zhang, Jinnan Chen, Qingwei Zhang, Yaqi Wang, Chengfeng Zhou, Xiaobo Li, Dahong Qian

机构 * School of Biomedical Engineering(生物医学工程学院) Division of Gastroenterology and Hepatology, Shanghai Institute of Digestive Disease, NHC Key Laboratory of Digestive Diseases, Renji Hospital(消化内科与肝病科、上海消化疾病研究所、国家消化疾病临床医学研究中心、仁济医院) College of Media Engineering(媒体工程学院) Aier Institute of Digital Ophthalmology and Visual Science(数字眼科与视觉科学研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19914 2025-08-12 cs.CV 79%

Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models

Sangwon Baik, Hyeonwoo Kim, Hanbyul Joo

机构 * Seoul National University(首尔国立大学) RLWRLD

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://tlb-miss.github.io/oor/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08333 2025-08-12 cs.CV 79%

DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models

Hyeonwoo Kim, Sangwon Baik, Hanbyul Joo

机构 * Seoul National University(首尔国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://snuvclab.github.io/david/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01440 2025-08-12 cs.CV 79%

BadPatch: Diffusion-Based Generation of Physical Adversarial Patches

Zhixiang Wang, Xingjun Ma, Yu-Gang Jiang

机构 * Shanghai Key Lab of Intell. Info. Processing, School of CS, Fudan University(上海智能信息处理关键实验室,计算机科学学院,复旦大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Code available at: https://github.com/Wwangb/BadPatch

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12777 2025-08-12 cs.CV cs.CL cs.CR cs.LG 79%

Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts

Hongcheng Gao, Tianyu Pang, Chao Du, Taihang Hu, Zhijie Deng, Min Lin

机构 * Sea AI Lab, Singapore(新加坡Sea AI实验室) University of Chinese Academy of Sciences(中国科学院大学) Shanghai Jiao Tong University(上海交通大学) Nankai University(南开大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07128 2025-08-12 cs.CV cs.AI 79%

Perceptual Evaluation of GANs and Diffusion Models for Generating X-rays

Gregory Schuit, Denis Parra, Cecilia Besa

机构 * Pontificia Universidad Católica de Chile(智利天主教大学) iHealth - Instituto Milenio en Ingeniería e Inteligencia Artificial para la Salud(iHealth - 毫米级工程与人工智能健康研究所) CENIA – Centro Nacional de Inteligencia Artificial(国家人工智能中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to the Workshop on Human-AI Collaboration at MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02056 2025-08-12 cs.CV 79%

StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion

Haoxin Yang, Weihong Chen, Xuemiao Xu, Cheng Xu, Peng Xiao, Cuifeng Sun, Shaoyu Huang, Shengfeng He

机构 * School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院) Centre for Smart Health, The Hong Kong Polytechnic University(香港理工大学智能健康中心) Cloud Computing Center, Chinese Academy of Sciences(中国科学院云计算中心) Guangzhou Yichuang Information Technology Co., Ltd.(广州亿创信息技术有限公司) School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算机与信息系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14769 2025-08-12 cs.CV cs.RO 79%

CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion

Jiahua Ma, Yiran Qin, Yixiong Li, Xuanqi Liao, Yulan Guo, Ruimao Zhang

机构 * Sun Yat-sen University(中山大学) CUHK(SZ)(香港中文大学(深圳)) Oxford(牛津大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05855 2025-08-12 cs.RO cs.CV 79%

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Junjie Wen, Yichen Zhu, Jinming Li, Zhibin Tang, Chaomin Shen, Feifei Feng

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments The webpage is at https://dex-vla.github.io/. DexVLA is accepted by CoRL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06494 2025-08-11 cs.CV 79%

LightSwitch: Multi-view Relighting with Material-guided Diffusion

Yehonathan Litman, Fernando De la Torre, Shubham Tulsiani

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025, Project page & Code: https://yehonathanlitman.github.io/light_switch/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06327 2025-08-11 cs.CV 79%

Can Diffusion Models Bridge the Domain Gap in Cardiac MR Imaging?

Xin Ci Wong, Duygu Sarikaya, Kieran Zucker, Marc De Kamps, Nishant Ravikumar

机构 * Centre for Doctoral Training in AI for Medical Diagnosis and Care(人工智能医学诊断与护理博士培训中心) School of Computing, University of Leeds(莱斯特大学计算机学院) Leeds Cancer Centre, St James’s University Hospital, Leeds, UK(利兹癌症中心,圣詹姆斯大学医院,利兹,英国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICONIP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06160 2025-08-11 cs.CV 79%

Fewer Denoising Steps or Cheaper Per-Step Inference: Towards Compute-Optimal Diffusion Model Deployment

Zhenbang Du, Yonggan Fu, Lifu Wang, Jiayi Qian, Xiao Luo, Yingyan, Lin

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06139 2025-08-11 cs.CV 79%

DiffCap: Diffusion-based Real-time Human Motion Capture using Sparse IMUs and a Monocular Camera

Shaohua Pan, Xinyu Yi, Yan Zhou, Weihua Jian, Yuan Zhang, Pengfei Wan, Feng Xu

机构 * Tsinghua University(清华大学) Kuaishou Technology(快手科技)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06014 2025-08-11 cs.CV 79%

ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors

Minsu Kim, Subin Jeon, In Cho, Mijin Yoo, Seon Joo Kim

机构 * Yonsei University(延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 6 Figures, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06001 2025-08-11 cs.DC cs.CV 79%

KnapFormer: An Online Load Balancer for Efficient Diffusion Transformers Training

Kai Zhang, Peng Wang, Sai Bi, Jianming Zhang, Yuanjun Xiong

机构 * Adobe(Adobe公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Code is available at https://github.com/Kai-46/KnapFormer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.03687 2025-08-11 cs.CV cs.LG 79%

Conditional Diffusion Models are Medical Image Classifiers that Provide Explainability and Uncertainty for Free

Gian Mario Favero, Parham Saremi, Emily Kaczmarek, Brennan Nichyporuk, Tal Arbel

机构 * McGill University(麦吉尔大学) Mila – Quebec AI Institute(魁北克人工智能研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted for publication at MIDL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18350 2025-08-11 cs.CV cs.AI 79%

TryOffDiff: Virtual-Try-Off via High-Fidelity Garment Reconstruction using Diffusion Models

Riza Velioglu, Petra Bevandic, Robin Chan, Barbara Hammer

机构 * Machine Learning Group, CITEC Bielefeld University Germany(比勒菲尔德大学机器学习小组、CITEC) Institute of Mathematics Technical University Berlin Germany(柏林技术大学数学研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at BMVC'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05060 2025-08-08 cs.CV 79%

DualMat: PBR Material Estimation via Coherent Dual-Path Diffusion

Yifeng Huang, Zhang Chen, Yi Xu, Minh Hoai, Zhong Li

机构 * Stony Brook University(石溪大学) Meta Goertek Alpha Labs(Goertek Alpha实验室) The University of Adelaide(阿德莱德大学) Apple Inc.(苹果公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02903 2025-08-08 cs.CV eess.IV stat.ML 79%

RDDPM: Robust Denoising Diffusion Probabilistic Model for Unsupervised Anomaly Segmentation

Mehrdad Moradi, Kamran Paynabar

机构 * H. Milton Stewart School of Industrial and Systems Engineering, Georgia Institute of Technology(H. Milton Stewart工业与系统工程学院,佐治亚理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 5 figures. Accepted to the ICCV 2025 Workshop on Vision-based Industrial InspectiON (VISION)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01603 2025-08-08 cs.CV 79%

DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation

Yue-Jiang Dong, Wang Zhao, Jiale Xu, Ying Shan, Song-Hai Zhang

机构 * Tsinghua University(清华大学) ARC Lab, Tencent PCG(腾讯PCG实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICCV 2025; Project Homepage: https://yuejiangdong.github.io/depthsync

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12108 2025-08-08 cs.CV cs.AI 79%

EarthSynth: Generating Informative Earth Observation with Diffusion Models

Jiancheng Pan, Shiye Lei, Yuqian Fu, Jiahao Li, Yanxing Liu, Yuze Sun, Xiao He, Long Peng, Xiaomeng Huang, Bo Zhao

机构 * Tsinghua University(清华大学) University of Sydney(悉尼大学) University of Chinese Academy of Sciences(中国科学院大学) Wuhan University(武汉大学) University of Science and Technology of China(中国科学技术大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 25 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15451 2025-08-08 cs.CV 79%

MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space

Lixing Xiao, Shunlin Lu, Huaijin Pi, Ke Fan, Liang Pan, Yueer Zhou, Ziyong Feng, Xiaowei Zhou, Sida Peng, Jingbo Wang

机构 * Zhejiang University(浙江大学) The Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)) The University of Hong Kong(香港大学) Shanghai Jiao Tong University(上海交通大学) DeepGlint Shanghai AI Laboratory(上海人工智能实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025. Project Page: https://zju3dv.github.io/MotionStreamer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09514 2025-08-08 cs.CV 79%

CM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images

Bin Hu, Chenqiang Gao, Shurui Liu, Junjie Guo, Fang Chen, Fangcen Liu, Junwei Han

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15894 2025-08-08 cs.CV 79%

RIFLEx: A Free Lunch for Length Extrapolation in Video Diffusion Transformers

Min Zhao, Guande He, Yixiao Chen, Hongzhou Zhu, Chongxuan Li, Jun Zhu

机构 * Dept. of Comp. Sci. \& Tech., BNRist Center, THU-Bosch ML Center, Tsinghua University. The University of Texas at Austin. Gaoling School of Artificial Intelligence Renmin University of China Beijing, China. Beijing Key Laboratory of Research on Large Models Engineering Research Center of Next-Generation Intelligent Search Pazhou Laboratory (Huangpu)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICML 2025. Project page: https://riflex-video.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04565 2025-08-07 cs.CV 79%

TAlignDiff: Automatic Tooth Alignment assisted by Diffusion-based Transformation Learning

Yunbi Liu, Enqi Tang, Shiyu Li, Lei Ma, Juncheng Li, Shu Lou, Yongchu Pan, Qingshan Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Submitted to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04467 2025-08-07 cs.CV 79%

4DVD: Cascaded Dense-view Video Diffusion Model for High-quality 4D Content Generation

Shuzhou Yang, Xiaodong Cun, Xiaoyu Li, Yaowei Li, Jian Zhang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04233 2025-08-07 cs.CV 79%

DocVCE: Diffusion-based Visual Counterfactual Explanations for Document Image Classification

Saifullah Saifullah, Stefan Agne, Andreas Dengel, Sheraz Ahmed

机构 * Smarte Daten and Wissensdienste (SDS), Deutsches Forschungszentrum für Künstliche Intelligenz GmbH (DFKI)(智能数据与知识服务(SDS),德国人工智能研究中心(DFKI)) Department of Computer Science, RPTU Kaiserslautern-Landau(计算机科学系,凯撒斯劳滕-兰道大学) DeepReader GmbH

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04229 2025-08-07 cs.CV 79%

Intention Enhanced Diffusion Model for Multimodal Pedestrian Trajectory Prediction

Yu Liu, Zhijie Liu, Xiao Ren, You-Fu Li, He Kong

机构 * Guangdong Provincial Key Laboratory of Fully Actuated System Control Theory and Technology, the Southern University of Science and Technology, Shenzhen 518055, China(广东省全自动化系统控制理论与技术重点实验室,南方科技大学,深圳518055,中国) Department of Mechanical Engineering, City University of Hong Kong, Hong Kong SAR, China(香港城市大学机械工程系,香港特别行政区,中国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments To be presented at the 28th IEEE International Conference on Intelligent Transportation Systems (ITSC), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏