arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86872 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2411.08402 2025-08-26 cs.CV 79%

V2X-R: Cooperative LiDAR-4D Radar Fusion with Denoising Diffusion for 3D Object Detection

Xun Huang, Jinlong Wang, Qiming Xia, Siheng Chen, Bisheng Yang, Xin Li, Cheng Wang, Chenglu Wen

机构 * Fujian Key Laboratory of Sensing and Computing for Smart Cities, Xiamen University, China(福建智能城市感知与计算重点实验室,厦门大学) Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University, China(多媒体可信感知与高效计算重点实验室,中国教育部,厦门大学) Zhongguancun Academy(中关村学院) Shanghai Jiao Tong University(上海交通大学) Wuhan University(武汉大学) Texas A&M University(德克萨斯A&M大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.02374 2025-08-26 cs.CV 79%

PromptRR: Diffusion Models as Prompt Generators for Single Image Reflection Removal

Tao Wang, Wanglong Lu, Kaihao Zhang, Tong Lu, Ming-Hsuan Yang

机构 * State Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学) Department of Computer Science at Memorial University of Newfoundland(纪念大学计算机科学系) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) School of Engineering, University of California at Merced(加州大学默塞德分校工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16655 2025-08-26 cs.LG cs.CV 79%

A Laplace diffusion-based transformer model for heart rate forecasting within daily activity context

Andrei Mateescu, Ioana Hadarau, Ionut Anghel, Tudor Cioara, Ovidiu Anchidin, Ancuta Nemes

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12835 2025-08-26 cs.CV 79%

DiffS-NOCS: 3D Point Cloud Reconstruction through Coloring Sketches to NOCS Maps Using Diffusion Models

Di Kong, Qianhui Wan

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Tsinghua University(清华大学) Zhongguancun Academy(中关村学院) Beijing Normal University(北京师范大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16577 2025-08-25 cs.CV cs.AI 79%

MV-RAG: Retrieval Augmented Multiview Diffusion

Yosef Dayani, Omer Benishu, Sagie Benaim

机构 * Hebrew University of Jerusalem(耶路撒冷希伯来大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://yosefdayani.github.io/MV-RAG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15972 2025-08-25 cs.RO cs.CV 79%

UnPose: Uncertainty-Guided Diffusion Priors for Zero-Shot Pose Estimation

Zhaodong Jiang, Ashish Sinha, Tongtong Cao, Yuan Ren, Bingbing Liu, Binbin Xu

机构 * Huawei Noah’s Ark Lab, Canada(华为诺亚实验室,加拿大) University of Toronto, Canada(多伦多大学,加拿大)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Published at the Conference on Robot Learning (CoRL) 2025. For more details please visit https://frankzhaodong.github.io/UnPose

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15641 2025-08-22 cs.CV 79%

When and What: Diffusion-Grounded VideoLLM with Entity Aware Segmentation for Long Video Understanding

Pengcheng Fang, Yuxia Chen, Rui Guo

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15236 2025-08-22 eess.IV cs.CV 79%

Pathology-Informed Latent Diffusion Model for Anomaly Detection in Lymph Node Metastasis

Jiamu Wang, Keunho Byeon, Jinsol Song, Anh Nguyen, Sangjeong Ahn, Sung Hak Lee, Jin Tae Kwak

机构 * School of Electrical Engineering, Korea University, Seoul 02841, Korea(韩国大学电子工程学院) Department of Pathology, Korea University Anam Hospital and Department of Biomedical Informatics, Korea University College of Medicine, Seoul 02841, Korea(韩国大学医学院病理学系) Department of Hospital Pathology, Seoul St. Mary’s Hospital, College of Medicine, The Catholic University of Korea, Seoul 06591, Korea(韩国天主大学医学院圣玛丽医院医院病理学系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15233 2025-08-22 cs.CV cs.LG 79%

Pretrained Diffusion Models Are Inherently Skipped-Step Samplers

Wenju Xu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14930 2025-08-22 cs.GR 79%

Hybrelighter: Combining Deep Anisotropic Diffusion and Scene Reconstruction for On-device Real-time Relighting in Mixed Reality

Hanwen Zhao, John Akers, Baback Elmieh, Ira Kemelmacher-Shlizerman

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08170 2025-08-22 cs.CV 79%

ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction

Chaojun Ni, Guosheng Zhao, Xiaofeng Wang, Zheng Zhu, Wenkang Qin, Xinze Chen, Guanghong Jia, Guan Huang, Wenjun Mei

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14558 2025-08-22 cs.CV cs.RO 79%

SuperPC: A Single Diffusion Model for Point Cloud Completion, Upsampling, Denoising, and Colorization

Yi Du, Zhipeng Zhao, Shaoshu Su, Sharath Golluri, Haoze Zheng, Runmao Yao, Chen Wang

机构 * Spatial AI & Robotics (SAIR) Lab, University at Buffalo(空间人工智能与机器人实验室,布法罗大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14811 2025-08-21 cs.CV 79%

Tinker: Diffusion's Gift to 3D--Multi-View Consistent Editing From Sparse Inputs without Per-Scene Optimization

Canyu Zhao, Xiaoman Li, Tianjian Feng, Zhiyue Zhao, Hao Chen, Chunhua Shen

机构 * Zhejiang University(浙江大学) Zhejiang University of Technology(浙江工业大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project webpage: https://aim-uofa.github.io/Tinker

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14717 2025-08-21 cs.CV 79%

GSFix3D: Diffusion-Guided Repair of Novel Views in Gaussian Splatting

Jiaxin Wei, Stefan Leutenegger, Simon Schaefer

机构 * Technical University of Munich(慕尼黑技术大学) ETH Zurich(苏黎世联邦理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14437 2025-08-21 cs.CV 79%

FOCUS: Frequency-Optimized Conditioning of DiffUSion Models for mitigating catastrophic forgetting during Test-Time Adaptation

Gabriel Tjio, Jie Zhang, Xulei Yang, Yun Xing, Nhat Chung, Xiaofeng Cao, Ivor W. Tsang, Chee Keong Kwoh, Qing Guo

机构 * Centre for Frontier AI Research (CFAR)(前沿人工智能研究中心) Agency for Science, Technology and Research (A*STAR)(科技研究局) College of Computing and Data Science(计算与数据科学学院) Nanyang Technological University(南洋理工大学) Institute of Infocomm Research (I2R)(信息通信研究所) School of Computer Science and Technology(计算机科学与技术学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14431 2025-08-21 cs.CV 79%

HyperDiff: Hypergraph Guided Diffusion Model for 3D Human Pose Estimation

Bing Han, Yuhua Huang, Pan Gao

机构 * Nanjing University of Aeronautics and Astronautics(南京航空航天大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14413 2025-08-21 cs.LG cs.CV 79%

Disentanglement in T-space for Faster and Distributed Training of Diffusion Models with Fewer Latent-states

Samarth Gupta, Raghudeep Gadde, Rui Chen, Aleix M. Martinez

机构 * Amazon(亚马逊)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14364 2025-08-21 physics.med-ph cs.CV 79%

Physics-Constrained Diffusion Reconstruction with Posterior Correction for Quantitative and Fast PET Imaging

Yucun Hou, Fenglin Zhan, Chenxi Li, Ziquan Yuan, Haoyu Lu, Yue Chen, Yihao Chen, Kexin Wang, Runze Liao, Haoqi Wen, Ganxi Du, Jiaru Ni, Taoran Chen, Jinyue Zhang, Jigang Yang, Jianyong Jiang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14122 2025-08-21 eess.IV cs.CV cs.LG q-bio.TO 79%

3D Cardiac Anatomy Generation Using Mesh Latent Diffusion Models

Jolanta Mozyrska, Marcel Beetz, Luke Melas-Kyriazi, Vicente Grau, Abhirup Banerjee, Alfonso Bueno-Orovio

机构 * Department of Computer Science, University of Oxford(计算机科学系,牛津大学) Institute of Biomedical Engineering, Department of Engineering Science, University of Oxford(生物医学工程研究所,工程科学系,牛津大学) Visual Geometry Group, Department of Engineering Science, University of Oxford(视觉几何组,工程科学系,牛津大学) Division of Cardiovascular Medicine, Radcliffe Department of Medicine, University of Oxford(心血管医学部,拉德克利夫医学系,牛津大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11579 2025-08-21 cs.CV cs.LG 79%

SketchDNN: Joint Continuous-Discrete Diffusion for CAD Sketch Generation

Sathvik Chereddy, John Femiani

机构 * Department of Computer Science, Miami-Oxford University, Oxford OH, USA(计算机科学系,迈阿密-牛津大学,牛津 OH,美国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 17 pages, 63 figures, Proceedings of the 42nd International Conference on Machine Learning (ICML2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13584 2025-08-20 cs.CV 79%

Temporal-Conditional Referring Video Object Segmentation with Noise-Free Text-to-Video Diffusion Model

Ruixin Zhang, Jiaqing Fan, Yifan Liao, Qian Qiao, Fanzhang Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 11 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09578 2025-08-20 cs.CV cs.LG 79%

Unsupervised Anomaly Detection Using Diffusion Trend Analysis for Display Inspection

Eunwoo Kim, Un Yang, Cheol Lae Roh, Stefano Ermon

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Published in the SID Digest of Technical Papers 2025 (Volume 56, Issue 1)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.20077 2025-08-20 cs.CV 79%

HouseCrafter: Lifting Floorplans to 3D Scenes with 2D Diffusion Model

Hieu T. Nguyen, Yiwen Chen, Vikram Voleti, Varun Jampani, Huaizu Jiang

机构 * Northeastern University(东北大学) Stability AI

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13300 2025-08-20 cs.CV cs.AI 79%

GaitCrafter: Diffusion Model for Biometric Preserving Gait Synthesis

Sirshapan Mitra, Yogesh S. Rawat

机构 * CRCV, University of Central Florida(CRCV,中央佛罗里达大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13091 2025-08-19 cs.CV 79%

DMS:Diffusion-Based Multi-Baseline Stereo Generation for Improving Self-Supervised Depth Estimation

Zihua Liu, Yizhou Li, Songyan Zhang, Masatoshi Okutomi

机构 * Institute of Science Tokyo, Japan(东京科学研究所) Sony Semiconductor Solutions Group, Japan(日本索尼半导体解决方案集团) Nanyang Technological University, Singapore(南洋理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12784 2025-08-19 cs.CV 79%

Leveraging Diffusion Models for Stylization using Multiple Style Images

Dan Ruta, Abdelaziz Djelouah, Raphael Ortiz, Christopher Schroers

机构 * DisneyResearch|Studios(迪士尼研究|工作室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12663 2025-08-19 cs.CV 79%

Stable Diffusion-Based Approach for Human De-Occlusion

Seung Young Noh, Ju Yong Chang

机构 * Kwangwoon University(韩国成均馆大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12603 2025-08-19 cs.CV 79%

ViLaD: A Large Vision Language Diffusion Framework for End-to-End Autonomous Driving

Can Cui, Yupeng Zhou, Juntong Peng, Sung-Yeon Park, Zichong Yang, Prashanth Sankaranarayanan, Jiaru Zhang, Ruqi Zhang, Ziran Wang

机构 * College of Engineering, Purdue University(普渡大学工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12148 2025-08-19 cs.CV cs.AI 79%

Demystifying Foreground-Background Memorization in Diffusion Models

Jimmy Z. Di, Yiwei Lu, Yaoliang Yu, Gautam Kamath, Adam Dziedzic, Franziska Boenisch

机构 * Cheriton School of Computer Science, University of Waterloo and Vector Institute(查尔顿计算机科学学院,滑铁卢大学和向量研究所) University of Ottawa(渥太华大学) CISPA Helmholtz Center for Information Security(信息安全赫尔姆霍茨中心CISPA)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12084 2025-08-19 cs.CV cs.AI 79%

Generic Event Boundary Detection via Denoising Diffusion

Jaejun Hwang, Dayoung Gong, Manjin Kim, Minsu Cho

机构 * Pohang University of Science and Technology (POSTECH)(釜山科学技术大学) GenGenAI

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏