arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2504.11154 2025-04-16 cs.CV eess.IV 79%

SAR-to-RGB Translation with Latent Diffusion for Earth Observation

Kaan Aydin, Joelle Hanna, Damian Borth

机构 * University of St. Gallen(圣加仑大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11034 2025-04-16 cs.CV 79%

Defending Against Frequency-Based Attacks with Diffusion Models

Fatemeh Amerehi, Patrick Healy

机构 * University of Limerick(利默里克大学) Computer Science and Information Systems(计算机科学与信息系统学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), 5th Workshop on Adversarial Machine Learning in Computer Vision: Foundation Models + X

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20502 2025-04-16 cs.CV 79%

ARLON: Boosting Diffusion Transformers with Autoregressive Models for Long Video Generation

Zongyi Li, Shujie Hu, Shujie Liu, Long Zhou, Jeongsoo Choi, Lingwei Meng, Xun Guo, Jinyu Li, Hefei Ling, Furu Wei

机构 * Huazhong University of Science and Technology(华中科技大学) The Chinese University of Hong Kong(香港中文大学) Microsoft Corporation(微软公司) KAIST(韩国科学技术院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at ICLR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10317 2025-04-15 cs.CV 79%

Analysis of Attention in Video Diffusion Transformers

Yuxin Wen, Jim Wu, Ajay Jain, Tom Goldstein, Ashwinee Panda

机构 * University of Maryland(马里兰大学) GenmoAI

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10278 2025-04-15 cs.CV 79%

DiffMOD: Progressive Diffusion Point Denoising for Moving Object Detection in Remote Sensing

Jinyue Zhang, Xiangrong Zhang, Zhongjian Huang, Tianyang Zhang, Yifei Jiang, Licheng Jiao

机构 * Xidian University(西安电子科技大学) Inspur Software Co., Ltd.(浪潮软件股份有限公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 9 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10041 2025-04-15 cs.RO cs.CV 79%

Prior Does Matter: Visual Navigation via Denoising Diffusion Bridge Models

Hao Ren, Yiming Zeng, Zetong Bi, Zhaoliang Wan, Junlong Huang, Hui Cheng

机构 * School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院) School of Intelligent Systems Engineering, Sun Yat-sen University(中山大学智能系统工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10003 2025-04-15 cs.RO cs.CV 79%

NaviDiffusor: Cost-Guided Diffusion Model for Visual Navigation

Yiming Zeng, Hao Ren, Shuhang Wang, Junlong Huang, Hui Cheng

机构 * School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院) School of Intelligent Systems Engineering, Sun Yat-sen University(中山大学智能系统工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref ICRA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09202 2025-04-15 cs.CV 79%

From Visual Explanations to Counterfactual Explanations with Latent Diffusion

Tung Luu, Nam Le, Duc Le, Bac Le

机构 * Faculty of Information Technology, University of Science, VNU-HCM(越南胡志明市国家大学科学学院信息技术学院) Vietnam National University, Ho Chi Minh City(越南胡志明市国家大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV)

Journal ref Proceedings of the Winter Conference on Applications of Computer Vision (WACV), 2025, pp. 420-429

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.24037 2025-04-15 cs.CV 79%

TPC: Test-time Procrustes Calibration for Diffusion-based Human Image Animation

Sunjae Yoon, Gwanhyeong Koo, Younghwan Lee, Chang D. Yoo

机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院(KAIST))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 24 pages, 16 figures, NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06044 2025-04-15 cs.CV 79%

FRAG: Frequency Adapting Group for Diffusion Video Editing

Sunjae Yoon, Gwanhyeong Koo, Geonwoo Kim, Chang D. Yoo

机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 16 pages, 16 figures, ICML 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01867 2025-04-15 cs.CV 79%

MoLA: Motion Generation and Editing with Latent Diffusion Enhanced by Adversarial Training

Kengo Uchida, Takashi Shibuya, Yuhta Takida, Naoki Murata, Julian Tanke, Shusuke Takahashi, Yuki Mitsufuji

机构 * Sony AI(索尼人工智能) Sony Group Corporation(索尼集团公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025 HuMoGen Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08591 2025-04-14 cs.CV 79%

ZipIR: Latent Pyramid Diffusion Transformer for High-Resolution Image Restoration

Yongsheng Yu, Haitian Zheng, Zhifei Zhang, Jianming Zhang, Yuqian Zhou, Connelly Barnes, Yuchen Liu, Wei Xiong, Zhe Lin, Jiebo Luo

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08348 2025-04-14 cs.CV 79%

Geometric Consistency Refinement for Single Image Novel View Synthesis via Test-Time Adaptation of Diffusion Models

Josef Bengtson, David Nilsson, Fredrik Kahl

机构 * Chalmers University of Technology(查尔姆斯理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to CVPR 2025 EDGE Workshop. Project page: https://gc-ref.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08344 2025-04-14 cs.CV 79%

EasyGenNet: An Efficient Framework for Audio-Driven Gesture Video Generation Based on Diffusion Model

Renda Li, Xiaohua Qi, Qiang Ling, Jun Yu, Ziyi Chen, Peng Chang, Mei HanJing Xiao

机构 * University of Science and Technology of China(中国科学技术大学) PAII Inc.(PAII公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08291 2025-04-14 cs.CV 79%

DreamFuse: Adaptive Image Fusion with Diffusion Transformer

Junjia Huang, Pengxiang Yan, Jiyang Liu, Jie Wu, Zhao Wang, Yitong Wang, Liang Lin, Guanbin Li

机构 * Sun Yat-sen University(中山大学) ByteDance Intelligent Creation(字节跳动智能创作) Peng Cheng Laboratory(鹏城实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17459 2025-04-14 cs.CV cs.AI 79%

WF-VAE: Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Model

Zongjian Li, Bin Lin, Yang Ye, Liuhan Chen, Xinhua Cheng, Shenghai Yuan, Li Yuan

机构 * Shenzhen Graduate School, Peking University(北京大学深圳研究生院) Peng Cheng Laboratory(鹏城实验室) Rabbitpre Intelligence(睿智科技)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 8 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.07977 2025-04-14 cs.CV cs.AI cs.LG 79%

Dynamic Attention-Guided Diffusion for Image Super-Resolution

Brian B. Moser, Stanislav Frolov, Federico Raue, Sebastian Palacio, Andreas Dengel

机构 * German Research Center for Artificial Intelligence(德国人工智能研究中心) RPTU Kaiserslautern-Landau(凯泽斯劳滕-兰道莱茵-普法尔茨理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Brian B. Moser and Stanislav Frolov contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07560 2025-04-11 eess.IV cs.CV cs.LG 79%

PhaseGen: A Diffusion-Based Approach for Complex-Valued MRI Data Generation

Moritz Rempe, Fabian Hörst, Helmut Becker, Marco Schlimbach, Lukas Rotkopf, Kevin Kröninger, Jens Kleesiek

机构 * Institute for AI in Medicine (IKIM)(人工智能医学研究所) University Hospital Essen(埃森大学医院) Cancer Research Center Cologne Essen (CCCE)(科隆埃森癌症研究中心) University Medicine Essen(埃森大学医学部) Technical University Dortmund(多特蒙德工业大学) German Cancer Research Center (DKFZ)(德国癌症研究中心) German Cancer Consortium (DKTK)(德国癌症联盟) University of Duisburg-Essen(杜伊斯堡-埃森大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07308 2025-04-11 eess.IV cs.CV 79%

MoEDiff-SR: Mixture of Experts-Guided Diffusion Model for Region-Adaptive MRI Super-Resolution

Zhe Wang, Yuhua Ru, Aladine Chetouani, Fang Chen, Fabian Bauer, Liping Zhang, Didier Hans, Rachid Jennane, Mohamed Jarraya, Yung Hsin Chen

机构 * Massachusetts General Hospital(麻省总医院) Harvard Medical School(哈佛医学院) Jiangsu Institute of Hematology(江苏血液学研究所) The First Affiliated Hospital of Soochow University(苏州大学附属第一医院) University Sorbonne Paris Nord(巴黎北索邦大学) Henan University of Chinese Medicine(河南中医药大学) German Cancer Research Center(德国癌症研究中心) Athinoula A. Martinos Centre for Biomedical Imaging(阿西诺拉·A.马蒂诺斯生物医学成像中心) Geneva University Hospital(日内瓦大学医院) IDP Institute(IDP研究所) University of Orleans(奥尔良大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15139 2025-04-11 cs.CV cs.RO 79%

DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving

Bencheng Liao, Shaoyu Chen, Haoran Yin, Bo Jiang, Cheng Wang, Sixu Yan, Xinbang Zhang, Xiangyu Li, Ying Zhang, Qian Zhang, Xinggang Wang

机构 * Institute of Artificial Intelligence, Huazhong University of Science & Technology(华中科技大学人工智能研究院) School of EIC, Huazhong University of Science & Technology(华中科技大学电子信息与通信学院) Horizon Robotics(地平线机器人科技公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to CVPR 2025 as Highlight. Code & demo & model are available at https://github.com/hustvl/DiffusionDrive

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12816 2025-04-11 cs.LG cs.CV eess.IV 79%

Neural Approximate Mirror Maps for Constrained Diffusion Models

Berthy T. Feng, Ricardo Baptista, Katherine L. Bouman

机构 * California Institute of Technology(加州理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07008 2025-04-10 cs.CV 79%

Latent Diffusion U-Net Representations Contain Positional Embeddings and Anomalies

Jonas Loos, Lorenz Linhardt

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICLR 2025 Workshop on Deep Generative Models: Theory, Principle, and Efficacy

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06950 2025-04-10 cs.CV 79%

PathSegDiff: Pathology Segmentation using Diffusion model representations

Sachin Kumar Danisetty, Alexandros Graikos, Srikar Yellapragada, Dimitris Samaras

机构 * Stony Brook University(石溪大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05741 2025-04-10 cs.CV cs.AI 79%

DDT: Decoupled Diffusion Transformer

Shuai Wang, Zhi Tian, Weilin Huang, Limin Wang

机构 * Nanjing University(南京大学) ByteDance Seed Vision(字节跳动种子视觉)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments sota on ImageNet256 and ImageNet512

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.24354 2025-04-10 cs.LG cs.AI cs.CL cs.CV 79%

ORAL: Prompting Your Large-Scale LoRAs via Conditional Recurrent Diffusion

Rana Muhammad Shahroz Khan, Dongwen Tang, Pingzhi Li, Kai Wang, Tianlong Chen

机构 * The University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) National University of Singapore(新加坡国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17550 2025-04-10 cs.LG cs.MM cs.SD eess.AS 79%

A Simple but Strong Baseline for Sounding Video Generation: Effective Adaptation of Audio and Video Diffusion Models for Joint Generation

Masato Ishii, Akio Hayakawa, Takashi Shibuya, Yuki Mitsufuji

机构 * Sony AI(索尼人工智能) Sony Group Corp.(索尼集团公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

Comments IJCNN 2025. The source code is available: https://github.com/SonyResearch/SVG_baseline

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.13055 2025-04-10 cs.CV 79%

Atlas Gaussians Diffusion for 3D Generation

Haitao Yang, Yuan Dong, Hanwen Jiang, Dejia Xu, Georgios Pavlakos, Qixing Huang

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校) Alibaba Group(阿里巴巴集团)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Published at ICLR 2025 (Spotlight). Project page: https://yanghtr.github.io/projects/atlas_gaussians

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05468 2025-04-09 cs.CV 79%

Studying Image Diffusion Features for Zero-Shot Video Object Segmentation

Thanos Delatolas, Vicky Kalogeiton, Dim P. Papadopoulos

机构 * Technical University of Denmark(丹麦技术大学) Pioneer Center for AI(先锋人工智能中心) Ecole Polytechnique(巴黎综合理工学院) CNRS(法国国家科学研究中心) Institut Polytechnique de Paris(巴黎理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to CVPRW2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05271 2025-04-08 cs.CV cs.LG 79%

AnomalousNet: A Hybrid Approach with Attention U-Nets and Change Point Detection for Accurate Characterization of Anomalous Diffusion in Video Data

Yusef Ahsini, Marc Escoto, J. Alberto Conejero

机构 * Instituto Universitario de Matemática Pura y Aplicada (IUMPA), Universitat Politècnica de València(瓦伦西亚理工大学纯数学与应用数学大学研究所) Centro de Investigación en Gestión e Ingeniería de Producción (CIGIP), Universitat Politècnica de València(瓦伦西亚理工大学生产管理与工程研究中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 20 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05135 2025-04-08 cs.CV 79%

DA2Diff: Exploring Degradation-aware Adaptive Diffusion Priors for All-in-One Weather Restoration

Jiamei Xiong, Xuefeng Yan, Yongzhen Wang, Wei Zhao, Xiao-Ping Zhang, Mingqiang Wei

机构 * School of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大学计算机科学与技术学院) Collaborative Innovation Center of Novel Software Technology and Industrialization(新型软件技术与产业化协同创新中心) College of Computer Science and Technology, Anhui University of Technology(安徽工业大学计算机科学与技术学院) Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏