arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70196 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70196 篇

2509.01837 2025-09-09 cs.CV 79%

PractiLight: Practical Light Control Using Foundational Diffusion Models

Yotam Erel, Rishabh Dabral, Vladislav Golyanik, Amit H. Bermano, Christian Theobalt

机构 * Tel Aviv University(特拉维夫大学) Max Planck Institute for Informatics(马克斯·普朗克信息研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://yoterel.github.io/PractiLight-project-page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12103 2025-09-09 cs.CV cs.CY 79%

DeepShade: Enable Shade Simulation by Text-conditioned Image Generation

Longchao Da, Xiangrui Liu, Mithun Shivakoti, Thirulogasankar Pranav Kutralingam, Yezhou Yang, Hua Wei

机构 * Arizona State University(亚利桑那州立大学)

专题命中 扩散模型 :image generation(title);diffusion(abstract);分类 cs.CV

Comments 7pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19757 2025-09-09 cs.RO cs.CV 79%

Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy

Zhi Hou, Tianyi Zhang, Yuwen Xiong, Haonan Duan, Hengjun Pu, Ronglei Tong, Chengyang Zhao, Xizhou Zhu, Yu Qiao, Jifeng Dai, Yuntao Chen

机构 * Shanghai AI Lab(上海人工智能实验室) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) MMLab, The Chinese University of Hong Kong(香港中文大学MMLab) Peking University(北京大学) SenseTime Research(商汤科技研究院) Tsinghua University(清华大学) HKISI, CAS(中国科学院香港中文大学研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Preprint; https://robodita.github.io; To appear in ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16897 2025-09-08 eess.IV cs.CV physics.med-ph 79%

Generating Synthetic Contrast-Enhanced Chest CT Images from Non-Contrast Scans Using Slice-Consistent Brownian Bridge Diffusion Network

Pouya Shiri, Xin Yi, Neel P. Mistry, Samaneh Javadinia, Mohammad Chegini, Seok-Bum Ko, Amirali Baniasadi, Scott J. Adams

机构 * Department of Medical Imaging, University of Saskatchewan(医学成像系,萨斯喀彻温大学) Department of Electrical and Computer Engineering, University of Saskatchewan(电气与计算机工程系,萨斯喀彻温大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04193 2025-09-05 cs.CV cs.LG 79%

DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval

Ruohong Yang, Peng Hu, Yunfan Li, Xi Peng

专题命中 扩散模型 :diffusion(title);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03794 2025-09-05 cs.CV 79%

Fitting Image Diffusion Models on Video Datasets

Juhun Lee, Simon S. Woo

机构 * Dept. of Artificial Intelligence(人工智能系) Sungkyungwan University(顺天妇女大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV25 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20789 2025-09-05 cs.CV cs.LG 79%

Integrating Intermediate Layer Optimization and Projected Gradient Descent for Solving Inverse Problems with Diffusion Models

Yang Zheng, Wen Li, Zhaoqiang Liu

机构 * uestc(电子科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09885 2025-09-05 cs.SD cs.CV eess.AS 79%

Separate to Collaborate: Dual-Stream Diffusion Model for Coordinated Piano Hand Motion Synthesis

Zihao Liu, Mingwen Ou, Zunnan Xu, Jiaqi Huang, Haonan Han, Ronghui Li, Xiu Li

机构 * Tsinghua University(清华大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 7 figures, Accepted to ACMMM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14891 2025-09-05 cs.CV cs.AI 79%

CoDiff: Conditional Diffusion Model for Collaborative 3D Object Detection

Zhe Huang, Shuo Wang, Yongcai Wang, Lei Wang

机构 * School of future Transportation, Chang’an University(未来交通学院,长安大学) School of information, Renmin University of China(信息学院,中国人民大学) Hangzhou International Innovation Institute, Beihang University(杭州国际创新研究院,北航) University of Wollongong(沃林戈大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01919 2025-09-04 cs.CV cs.PF 79%

A Diffusion-Based Framework for Configurable and Realistic Multi-Storage Trace Generation

Seohyun Kim, Junyoung Lee, Jongho Park, Jinhyung Koo, Sungjin Lee, Yeseong Kim

机构 * DGIST POSTECH

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03324 2025-09-04 cs.CV 79%

InfraDiffusion: zero-shot depth map restoration with diffusion models and prompted segmentation from sparse infrastructure point clouds

Yixiong Jing, Cheng Zhang, Haibing Wu, Guangming Wang, Olaf Wysocki, Brian Sheil

机构 * Construction Engineering, University of Cambridge(剑桥大学建筑工程系) College of Civil Engineering, Hunan University(湖南大学土木工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03267 2025-09-04 cs.CV 79%

SynBT: High-quality Tumor Synthesis for Breast Tumor Segmentation by 3D Diffusion Model

Hongxu Yang, Edina Timko, Levente Lippenszky, Vanda Czipczer, Lehel Ferenczi

机构 * Science \& Technology Org. AI \& ML, GE HealthCare, Eindhoven, Netherlands Science \& Technology Org. AI \& ML, GE HealthCare, Budapest, Hungary

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by MICCAI 2025 Deep-Breath Workshop. Supported by IHI SYNTHIA project

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03141 2025-09-04 cs.CV cs.LG 79%

Temporally-Aware Diffusion Model for Brain Progression Modelling with Bidirectional Temporal Regularisation

Mattia Litrico, Francesco Guarnera, Mario Valerio Giuffrida, Daniele Ravì, Sebastiano Battiato

机构 * University of Catania(卡塔尼亚大学) University of Nottingham(诺丁汉大学) University of Messina(梅萨尼亚大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03114 2025-09-04 cs.CV 79%

Towards Realistic Hand-Object Interaction with Gravity-Field Based Diffusion Bridge

Miao Xu, Xiangyu Zhu, Xusheng Liang, Zidu Wang, Jinlin Wu, Zhen Lei

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Centre for Artificial Intelligence and Robotics, Hong Kong Institute of Science(香港科学院人工智能与机器人中心) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) China Mobile Financial Technology Co., Ltd.(中国移动金融科技有限公司) School of Computer Science and Engineering, the Faculty of Innovation Engineering, M.U.S.T(澳门科技大学计算机科学与工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06656 2025-09-04 cs.CV cs.LG 79%

Enhancing Diffusion Model Stability for Image Restoration via Gradient Management

Hongjie Wu, Mingqin Zhang, Linchao He, Ji-Zhe Zhou, Jiancheng Lv

机构 * College of Computer Science, Sichuan University(四川大学计算机学院) National Key Laboratory of Fundamental Science on Synthetic Vision, Sichuan University(合成视觉基础科学国家重点实验室(四川大学)) Engineering Research Center of Machine Learning(机器学习工程研究中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ACM Multimedia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02357 2025-09-03 cs.CV 79%

Category-Aware 3D Object Composition with Disentangled Texture and Shape Multi-view Diffusion

Zeren Xiong, Zikun Chen, Zedong Zhang, Xiang Li, Ying Tai, Jian Yang, Jun Li

机构 * Nanjing University of Science and Technology(南京理工大学) Nankai University(南开大学) Nanjing University(南京大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ACM Multimedia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02099 2025-09-03 cs.CV 79%

A Data-Centric Approach to Pedestrian Attribute Recognition: Synthetic Augmentation via Prompt-driven Diffusion Models

Alejandro Alonso, Sawaiz A. Chaudhry, Juan C. SanMiguel, Álvaro García-Martín, Pablo Ayuso-Albizu, Pablo Carballeira

机构 * Autonomous University of Madrid(马德里自主大学) University of Bordeaux(波尔多大学) Pázmány Péter Catholic University(帕兹曼·彼得天主教大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Paper Acepted at AVSS 2025 conference. Best paper award

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01864 2025-09-03 cs.CV 79%

Latent Gene Diffusion for Spatial Transcriptomics Completion

Paula Cárdenas, Leonardo Manrique, Daniela Vega, Daniela Ruiz, Pablo Arbeláez

机构 * Center for Research and Formation in Artificial Intelligence(人工智能研究与培训中心) Universidad de los Andes(andes大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 8 figures. Accepted to CVAMD Workshop, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01177 2025-09-03 cs.CV cs.AI cs.HC eess.SP 79%

DynaMind: Reconstructing Dynamic Visual Scenes from EEG by Aligning Temporal Dynamics and Multimodal Semantics to Guided Diffusion

Junxiang Liu, Junming Lin, Jiangtong Li, Jie Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 14 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22459 2025-09-03 cs.CV 79%

Exploiting Diffusion Prior for Task-driven Image Restoration

Jaeha Kim, Junghun Oh, Kyoung Mu Lee

机构 * Dept. of ECE&ASRI(电子工程与先进科学研究院部) IPAI(人工智能研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICCV 2025. Code is available at https://github.com/JaehaKim97/EDTR

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21541 2025-09-03 cs.CV cs.AI 79%

DiffDecompose: Layer-Wise Decomposition of Alpha-Composited Images via Diffusion Transformers

Zitong Wang, Hang Zhao, Qianyu Zhou, Xuequan Lu, Xiangtai Li, Yiren Song

机构 * Jilin University(吉林大学) University of Western Australia(西澳大学) Nanyang Technological University(南洋理工大学) National University of Singapore(新加坡国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12470 2025-09-03 cs.CV 79%

SC-Diff: 3D Shape Completion with Latent Diffusion Models

Simon Schaefer, Juan D. Galvis, Xingxing Zuo, Stefan Leutengger

机构 * Technical University of Munich(慕尼黑技术大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) ETH Zurich(苏黎世联邦理工学院) MBZUAI(穆桑比克大学人工智能研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00843 2025-09-03 cs.CV cs.AI 79%

Look Beyond: Two-Stage Scene View Generation via Panorama and Video Diffusion

Xueyang Kang, Zhengkang Xiang, Zezheng Zhang, Kourosh Khoshelham

机构 * University of Melbourne(墨尔本大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 26 pages, 30 figures, 2025 ACM Multimedia

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00395 2025-09-03 cs.CV 79%

Double-Constraint Diffusion Model with Nuclear Regularization for Ultra-low-dose PET Reconstruction

Mengxiao Geng, Ran Hong, Bingxuan Li, Qiegen Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00065 2025-09-03 cs.RO cs.CV 79%

Hybrid Perception and Equivariant Diffusion for Robust Multi-Node Rebar Tying

Zhitao Wang, Yirong Xiong, Roberto Horowitz, Yanke Wang, Yuxing Han

机构 * Tsinghua University, Shenzhen International Graduate School, Shenzhen, China(清华大学深圳国际研究生院,深圳,中国) Department of Mechanical Engineering, University of California, Berkeley(加州大学伯克利分校机械工程系) Hong Kong Center for Construction Robotics, The Hong Kong University of Science and Technology(香港科技大学建设机器人中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by The IEEE International Conference on Automation Science and Engineering (CASE) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11999 2025-09-03 cs.RO cs.CV cs.SY eess.SY 79%

Diffusion Dynamics Models with Generative State Estimation for Cloth Manipulation

Tongxuan Tian, Haoyang Li, Bo Ai, Xiaodi Yuan, Zhiao Huang, Hao Su

机构 * University of California San Diego(加州大学圣地亚哥分校) Hillbot Inc(Hillbot公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CoRL 2025. Project website: https://uniclothdiff.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21542 2025-09-01 cs.CV cs.AI cs.RO 79%

Complete Gaussian Splats from a Single Image with Denoising Diffusion Models

Ziwei Liao, Mohamed Sayed, Steven L. Waslander, Sara Vicente, Daniyar Turmukhambetov, Michael Firman

机构 * University of Toronto(多伦多大学) Niantic Spatial

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Main paper: 11 pages; Supplementary materials: 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21257 2025-09-01 cs.CV 79%

PHD: Personalized 3D Human Body Fitting with Point Diffusion

Hsuan-I Ho, Chen Guo, Po-Chen Wu, Ivan Shugurov, Chengcheng Tang, Abhay Mittal, Sizhe An, Manuel Kaufmann, Linguang Zhang

机构 * Department of Computer Science, ETH Zürich(苏黎世联邦理工学院计算机科学系) ETH AI Center, ETH Zürich(苏黎世联邦理工学院人工智能中心) Reality Labs, Meta(Meta现实实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025, 19 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10257 2025-09-01 cs.CV cs.AI cs.LG 79%

Guiding a diffusion model using sliding windows

Nikolas Adaloglou, Tim Kaiser, Damir Iagudin, Markus Kollmann

机构 * Heinrich Heine University Düsseldorf(海因里希-海涅大学杜伊斯堡)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at BMVC 2025. 30 pages, 16 figures in total, including appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03548 2025-09-01 cs.CV 79%

HiDiff: Hybrid Diffusion Framework for Medical Image Segmentation

Tao Chen, Chenhui Wang, Zhihao Chen, Yiming Lei, Hongming Shan

机构 * Institute of Science and Technology for Brain-inspired Intelligence, Fudan University(脑启发智能科学与技术研究院,复旦大学) Shanghai Key Lab of Intelligent Information Processing, School of Computer Science, Fudan University(上海智能信息处理重点实验室,复旦大学计算机学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by IEEE Transactions on Medical Imaging 2024

Journal ref IEEE TRANSACTIONS ON MEDICAL IMAGING, VOL. 43, NO. 10, OCTOBER 2024

详情

展开后加载摘要…

URL PDF HTML 收藏