arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2508.17017 2025-08-26 cs.CV 79%

Dual Orthogonal Guidance for Robust Diffusion-based Handwritten Text Generation

Konstantina Nikolaidou, George Retsinas, Giorgos Sfikas, Silvia Cascianelli, Rita Cucchiara, Marcus Liwicki

机构 * Luleå University of Technology(卢莱大学) National Technical University of Athens(雅典技术大学) University of West Attica(西阿提卡大学) University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16956 2025-08-26 cs.CV 79%

RPD-Diff: Region-Adaptive Physics-Guided Diffusion Model for Visibility Enhancement under Dense and Non-Uniform Haze

Ruicheng Zhang, Puxin Yan, Zeyu Zhang, Yicheng Chang, Hongyi Chen, Zhi Jin

机构 * School of Intelligent Systems Engineering, Shenzhen Campus of Sun Yat-sen University(中山大学智能系统工程学院) Guangdong Provincial Key Laboratory of Fire Science and Intelligent Emergency Technology(广东省火灾科学与智能应急技术重点实验室) The Australian National University(澳大利亚国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16930 2025-08-26 eess.AS cs.CV cs.SD 79%

HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation

Sizhe Shan, Qiulin Li, Yutao Cui, Miles Yang, Yuehai Wang, Qun Yang, Jin Zhou, Zhao Zhong

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07903 2025-08-26 eess.IV cs.AI cs.CV 79%

Diffusing the Blind Spot: Uterine MRI Synthesis with Diffusion Models

Johanna P. Müller, Anika Knupfer, Pedro Blöss, Edoardo Berardi Vittur, Bernhard Kainz, Jana Hutter

机构 * Imperial College London, London, UK(伦敦帝国理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at MICCAI CAPI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10072 2025-08-26 cs.CV 79%

Frequency Regulation for Exposure Bias Mitigation in Diffusion Models

Meng Yu, Kun Zhan

机构 * School of Information Science and Engineering(信息科学与工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ACM Multimedia 2025 accepted!

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13087 2025-08-26 cs.CV cs.LG 79%

Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation

Akshay Krishnan, Xinchen Yan, Vincent Casser, Abhijit Kundu

机构 * Google DeepMind(谷歌DeepMind) Georgia Institute of Technology(佐治亚理工学院) Waymo

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICCV 2025. Project webpage: https://orchid3d.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08402 2025-08-26 cs.CV 79%

V2X-R: Cooperative LiDAR-4D Radar Fusion with Denoising Diffusion for 3D Object Detection

Xun Huang, Jinlong Wang, Qiming Xia, Siheng Chen, Bisheng Yang, Xin Li, Cheng Wang, Chenglu Wen

机构 * Fujian Key Laboratory of Sensing and Computing for Smart Cities, Xiamen University, China(福建智能城市感知与计算重点实验室,厦门大学) Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University, China(多媒体可信感知与高效计算重点实验室,中国教育部,厦门大学) Zhongguancun Academy(中关村学院) Shanghai Jiao Tong University(上海交通大学) Wuhan University(武汉大学) Texas A&M University(德克萨斯A&M大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by CVPR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.02374 2025-08-26 cs.CV 79%

PromptRR: Diffusion Models as Prompt Generators for Single Image Reflection Removal

Tao Wang, Wanglong Lu, Kaihao Zhang, Tong Lu, Ming-Hsuan Yang

机构 * State Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学) Department of Computer Science at Memorial University of Newfoundland(纪念大学计算机科学系) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) School of Engineering, University of California at Merced(加州大学默塞德分校工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16655 2025-08-26 cs.LG cs.CV 79%

A Laplace diffusion-based transformer model for heart rate forecasting within daily activity context

Andrei Mateescu, Ioana Hadarau, Ionut Anghel, Tudor Cioara, Ovidiu Anchidin, Ancuta Nemes

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12835 2025-08-26 cs.CV 79%

DiffS-NOCS: 3D Point Cloud Reconstruction through Coloring Sketches to NOCS Maps Using Diffusion Models

Di Kong, Qianhui Wan

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Tsinghua University(清华大学) Zhongguancun Academy(中关村学院) Beijing Normal University(北京师范大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16577 2025-08-25 cs.CV cs.AI 79%

MV-RAG: Retrieval Augmented Multiview Diffusion

Yosef Dayani, Omer Benishu, Sagie Benaim

机构 * Hebrew University of Jerusalem(耶路撒冷希伯来大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://yosefdayani.github.io/MV-RAG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15972 2025-08-25 cs.RO cs.CV 79%

UnPose: Uncertainty-Guided Diffusion Priors for Zero-Shot Pose Estimation

Zhaodong Jiang, Ashish Sinha, Tongtong Cao, Yuan Ren, Bingbing Liu, Binbin Xu

机构 * Huawei Noah’s Ark Lab, Canada(华为诺亚实验室,加拿大) University of Toronto, Canada(多伦多大学,加拿大)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Published at the Conference on Robot Learning (CoRL) 2025. For more details please visit https://frankzhaodong.github.io/UnPose

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15641 2025-08-22 cs.CV 79%

When and What: Diffusion-Grounded VideoLLM with Entity Aware Segmentation for Long Video Understanding

Pengcheng Fang, Yuxia Chen, Rui Guo

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15236 2025-08-22 eess.IV cs.CV 79%

Pathology-Informed Latent Diffusion Model for Anomaly Detection in Lymph Node Metastasis

Jiamu Wang, Keunho Byeon, Jinsol Song, Anh Nguyen, Sangjeong Ahn, Sung Hak Lee, Jin Tae Kwak

机构 * School of Electrical Engineering, Korea University, Seoul 02841, Korea(韩国大学电子工程学院) Department of Pathology, Korea University Anam Hospital and Department of Biomedical Informatics, Korea University College of Medicine, Seoul 02841, Korea(韩国大学医学院病理学系) Department of Hospital Pathology, Seoul St. Mary’s Hospital, College of Medicine, The Catholic University of Korea, Seoul 06591, Korea(韩国天主大学医学院圣玛丽医院医院病理学系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15233 2025-08-22 cs.CV cs.LG 79%

Pretrained Diffusion Models Are Inherently Skipped-Step Samplers

Wenju Xu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14930 2025-08-22 cs.GR 79%

Hybrelighter: Combining Deep Anisotropic Diffusion and Scene Reconstruction for On-device Real-time Relighting in Mixed Reality

Hanwen Zhao, John Akers, Baback Elmieh, Ira Kemelmacher-Shlizerman

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08170 2025-08-22 cs.CV 79%

ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction

Chaojun Ni, Guosheng Zhao, Xiaofeng Wang, Zheng Zhu, Wenkang Qin, Xinze Chen, Guanghong Jia, Guan Huang, Wenjun Mei

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14558 2025-08-22 cs.CV cs.RO 79%

SuperPC: A Single Diffusion Model for Point Cloud Completion, Upsampling, Denoising, and Colorization

Yi Du, Zhipeng Zhao, Shaoshu Su, Sharath Golluri, Haoze Zheng, Runmao Yao, Chen Wang

机构 * Spatial AI & Robotics (SAIR) Lab, University at Buffalo(空间人工智能与机器人实验室,布法罗大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14811 2025-08-21 cs.CV 79%

Tinker: Diffusion's Gift to 3D--Multi-View Consistent Editing From Sparse Inputs without Per-Scene Optimization

Canyu Zhao, Xiaoman Li, Tianjian Feng, Zhiyue Zhao, Hao Chen, Chunhua Shen

机构 * Zhejiang University(浙江大学) Zhejiang University of Technology(浙江工业大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project webpage: https://aim-uofa.github.io/Tinker

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14717 2025-08-21 cs.CV 79%

GSFix3D: Diffusion-Guided Repair of Novel Views in Gaussian Splatting

Jiaxin Wei, Stefan Leutenegger, Simon Schaefer

机构 * Technical University of Munich(慕尼黑技术大学) ETH Zurich(苏黎世联邦理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14437 2025-08-21 cs.CV 79%

FOCUS: Frequency-Optimized Conditioning of DiffUSion Models for mitigating catastrophic forgetting during Test-Time Adaptation

Gabriel Tjio, Jie Zhang, Xulei Yang, Yun Xing, Nhat Chung, Xiaofeng Cao, Ivor W. Tsang, Chee Keong Kwoh, Qing Guo

机构 * Centre for Frontier AI Research (CFAR)(前沿人工智能研究中心) Agency for Science, Technology and Research (A*STAR)(科技研究局) College of Computing and Data Science(计算与数据科学学院) Nanyang Technological University(南洋理工大学) Institute of Infocomm Research (I2R)(信息通信研究所) School of Computer Science and Technology(计算机科学与技术学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14431 2025-08-21 cs.CV 79%

HyperDiff: Hypergraph Guided Diffusion Model for 3D Human Pose Estimation

Bing Han, Yuhua Huang, Pan Gao

机构 * Nanjing University of Aeronautics and Astronautics(南京航空航天大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14413 2025-08-21 cs.LG cs.CV 79%

Disentanglement in T-space for Faster and Distributed Training of Diffusion Models with Fewer Latent-states

Samarth Gupta, Raghudeep Gadde, Rui Chen, Aleix M. Martinez

机构 * Amazon(亚马逊)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14364 2025-08-21 physics.med-ph cs.CV 79%

Physics-Constrained Diffusion Reconstruction with Posterior Correction for Quantitative and Fast PET Imaging

Yucun Hou, Fenglin Zhan, Chenxi Li, Ziquan Yuan, Haoyu Lu, Yue Chen, Yihao Chen, Kexin Wang, Runze Liao, Haoqi Wen, Ganxi Du, Jiaru Ni, Taoran Chen, Jinyue Zhang, Jigang Yang, Jianyong Jiang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14122 2025-08-21 eess.IV cs.CV cs.LG q-bio.TO 79%

3D Cardiac Anatomy Generation Using Mesh Latent Diffusion Models

Jolanta Mozyrska, Marcel Beetz, Luke Melas-Kyriazi, Vicente Grau, Abhirup Banerjee, Alfonso Bueno-Orovio

机构 * Department of Computer Science, University of Oxford(计算机科学系,牛津大学) Institute of Biomedical Engineering, Department of Engineering Science, University of Oxford(生物医学工程研究所,工程科学系,牛津大学) Visual Geometry Group, Department of Engineering Science, University of Oxford(视觉几何组,工程科学系,牛津大学) Division of Cardiovascular Medicine, Radcliffe Department of Medicine, University of Oxford(心血管医学部,拉德克利夫医学系,牛津大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11579 2025-08-21 cs.CV cs.LG 79%

SketchDNN: Joint Continuous-Discrete Diffusion for CAD Sketch Generation

Sathvik Chereddy, John Femiani

机构 * Department of Computer Science, Miami-Oxford University, Oxford OH, USA(计算机科学系,迈阿密-牛津大学,牛津 OH,美国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 17 pages, 63 figures, Proceedings of the 42nd International Conference on Machine Learning (ICML2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13584 2025-08-20 cs.CV 79%

Temporal-Conditional Referring Video Object Segmentation with Noise-Free Text-to-Video Diffusion Model

Ruixin Zhang, Jiaqing Fan, Yifan Liao, Qian Qiao, Fanzhang Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 11 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09578 2025-08-20 cs.CV cs.LG 79%

Unsupervised Anomaly Detection Using Diffusion Trend Analysis for Display Inspection

Eunwoo Kim, Un Yang, Cheol Lae Roh, Stefano Ermon

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Published in the SID Digest of Technical Papers 2025 (Volume 56, Issue 1)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.20077 2025-08-20 cs.CV 79%

HouseCrafter: Lifting Floorplans to 3D Scenes with 2D Diffusion Model

Hieu T. Nguyen, Yiwen Chen, Vikram Voleti, Varun Jampani, Huaizu Jiang

机构 * Northeastern University(东北大学) Stability AI

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13300 2025-08-20 cs.CV cs.AI 79%

GaitCrafter: Diffusion Model for Biometric Preserving Gait Synthesis

Sirshapan Mitra, Yogesh S. Rawat

机构 * CRCV, University of Central Florida(CRCV,中央佛罗里达大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏