arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-17 至 2025-10-17 共收录 39 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 39 篇

2510.14526 2025-10-17 cs.CV cs.LG 89%

Noise Projection: Closing the Prompt-Agnostic Gap Behind Text-to-Image Misalignment in Diffusion Models

Yunze Tong, Didi Zhu, Zijing Hu, Jinluan Yang, Ziyu Zhao

机构 * Zhejiang University(浙江大学)

专题命中 扩散模型 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Appendix will be appended soon

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14981 2025-10-17 cs.CV cs.AI 88%

Coupled Diffusion Sampling for Training-Free Multi-View Image Editing

Hadi Alzayer, Yunzhi Zhang, Chen Geng, Jia-Bin Huang, Jiajun Wu

机构 * Stanford University(斯坦福大学) University of Maryland, College Park(马里兰大学学院公园分校)

专题命中 扩散模型 :diffusion(title,abstract);image editing(title,abstract);分类 cs.CV

Comments Project page: https://coupled-diffusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18033 2025-10-17 cs.CV 83%

OmnimatteZero: Fast Training-free Omnimatte with Pre-trained Video Diffusion Models

Dvir Samuel, Matan Levy, Nir Darshan, Gal Chechik, Rami Ben-Ari

机构 * Bar-Ilan University \& OriginAI Israel The Hebrew University of Jerusalem Israel Bar-Ilan University \& NVIDIA Research Israel Bar-Ilan University \& OriginAI The Hebrew University of Jerusalem Bar-Ilan University \& NVIDIA Research

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted to SIGGRAPH ASIA 2025. Project Page: https://dvirsamuel.github.io/omnimattezero.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14427 2025-10-17 cs.MM cs.CV 81%

Deep Compositional Phase Diffusion for Long Motion Sequence Generation

Ho Yin Au, Jie Chen, Junkun Jiang, Jingyu Xiang

机构 * Hong Kong Baptist University(香港 Baptist 大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Accepted by NeurIPS 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14962 2025-10-17 cs.CV 79%

RainDiff: End-to-end Precipitation Nowcasting Via Token-wise Attention Diffusion

Thao Nguyen, Jiaqi Ma, Fahad Shahbaz Khan, Souhaib Ben Taieb, Salman Khan

机构 * Mohamed Bin Zayed University of AI(穆罕默德·本·扎耶德人工智能大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14634 2025-10-17 cs.CV 79%

SteeringTTA: Guiding Diffusion Trajectories for Robust Test-Time-Adaptation

Jihyun Yu, Yoojin Oh, Wonho Bae, Mingyu Kim, Junhyug Noh

机构 * Department of Artificial Intelligence, Ewha Womans University, Seoul, Republic of Korea(人工智能系,世ultan大学,韩国首尔) Department of Computer Science, University of British Columbia, Vancouver, Canada(计算机科学系,不列颠哥伦比亚大学,加拿大温哥华)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14314 2025-10-17 cs.CV 79%

A Multi-domain Image Translative Diffusion StyleGAN for Iris Presentation Attack Detection

Shivangi Yadav, Arun Ross

机构 * Michigan State University(密歇根州立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07944 2025-10-17 cs.CV 79%

CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving

Tianrui Zhang, Yichen Liu, Zilin Guo, Yuxin Guo, Jingcheng Ni, Chenjing Ding, Dan Xu, Lewei Lu, Zehuan Wu

机构 * Sensetime Research(商汤科技研究院) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10569 2025-10-17 cs.CR cs.AI cs.MM 79%

MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models

Leyi Pan, Sheng Guan, Zheyu Fu, Luyang Si, Huan Wang, Zian Wang, Hanqian Li, Xuming Hu, Irwin King, Philip S. Yu, Aiwei Liu, Lijie Wen

机构 * Tsinghua University(清华大学) Beijing University of Posts and Telecommunications(北京邮电大学) The Chinese University of Hong Kong(香港中文大学) University of Illinois at Chicago(伊利诺伊大学香槟分校) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

Comments 23 pages, 13 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09789 2025-10-17 cs.CV cs.AI 79%

On Equivariance and Fast Sampling in Video Diffusion Models Trained with Warped Noise

Chao Liu, Arash Vahdat

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15008 2025-10-17 cs.CV 79%

HuGDiffusion: Generalizable Single-Image Human Rendering via 3D Gaussian Diffusion

Yingzhi Tang, Qijian Zhang, Junhui Hou

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14961 2025-10-17 cs.LG cs.CL 78%

Efficient Parallel Samplers for Recurrent-Depth Models and Their Connection to Diffusion Language Models

Jonas Geiping, Xinyu Yang, Guinan Su

机构 * ELLIS Institute Tübingen & Max-Planck Institute for Intelligent Systems, Tübingen AI Center(图宾根ELLIS研究所及图宾根马克斯·普朗克智能系统研究所AI中心)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Code can be found at https://github.com/seal-rg/recurrent-pretraining

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14467 2025-10-17 cs.RO 78%

Restoring Noisy Demonstration for Imitation Learning With Diffusion Models

Shang-Fu Chen, Co Yong, Shao-Hua Sun

机构 * Graduate Institute of Communication Engineering, National Taiwan University(国立台湾大学通信工程研究所) Data Science Degree Program, National Taiwan University and Academia Sinica(国立台湾大学数据科学学士学位计划) Department of Electrical Engineering, National Taiwan University(国立台湾大学电子工程系)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Published in IEEE Transactions on Neural Networks and Learning Systems (TNNLS)

Journal ref IEEE Transactions on Neural Networks and Learning Systems (TNNLS), pp. 1-13, Sept. 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14269 2025-10-17 cs.LG stat.ML 78%

Nonparametric Data Attribution for Diffusion Models

Yutian Zhao, Chao Du, Xiaosen Zheng, Tianyu Pang, Min Lin

机构 * Sea AI Lab, Singapore(新加坡海人工智能实验室) Department of Mathematics, National University of Singapore(新加坡国立大学数学系) Singapore Management University(新加坡管理大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14114 2025-10-17 cs.LG 78%

Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey

Yazid Janati, Alain Durmus, Jimmy Olsson, Eric Moulines

机构 * Institute of Foundation Models Paris, MBZUAI(基础模型研究所巴黎,MBZUAI) KTH Royal Institute of Technology(皇家理工学院) Ecole polytechnique(巴黎高等理工学院) MBZUAI

专题命中 扩散模型 :diffusion(title,abstract)

Journal ref Philosophical Transactions A, 383(2299), 20240331 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14111 2025-10-17 cs.NI 78%

DiffLoc: Diffusion Model-Based High-Precision Positioning for 6G Networks

Taekyun Lee, Tommaso Balercia, Heasung Kim, Hyeji Kim, Jeffrey G. Andrews

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14094 2025-10-17 cs.LG 78%

Neural Network approximation power on homogeneous and heterogeneous reaction-diffusion equations

Haotian Feng

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12853 2025-10-17 astro-ph.IM math-ph math.MP 78%

Diffusion models for polarimetric reconstruction of circumstellar environments

Quentin Villegas, Laurence Denneulin, Simon Prunet, André Ferrari, Nelly Pustelnik, Éric Thiébaut, Julian Tachella, Maud Langlois

专题命中 扩散模型 :diffusion(title,abstract)

Comments in French language. GRETSI 2025 -- XXXe Colloque sur le Traitement du Signal et des Images, Aug 2025, Strasboug, France

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12210 2025-10-17 eess.AS cs.CL cs.LG 78%

DiSTAR: Diffusion over a Scalable Token Autoregressive Representation for Speech Generation

Yakun Song, Xiaobin Zhuang, Jiawei Chen, Zhikang Niu, Guanrou Yang, Chenpeng Du, Dongya Jia, Zhuo Chen, Yuping Wang, Yuxuan Wang, Xie Chen

机构 * X-LANCE Lab, School of Computer Science, Shanghai Jiao Tong University(X-LANCE实验室,计算机科学学院,上海交通大学) ByteDance Inc.(字节跳动公司)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09383 2025-10-17 cs.RO cs.LG 78%

Real-Time Adaptive Motion Planning via Point Cloud-Guided, Energy-Based Diffusion and Potential Fields

Wondmgezahu Teshome, Kian Behzad, Octavia Camps, Michael Everett, Milad Siami, Mario Sznaier

机构 * ECE Department, Northeastern University(东北大学电子与计算机工程系)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted to IEEE RA-L 2025

Journal ref IEEE Robotics and Automation Letters 10 (2025) 9160-9167

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00746 2025-10-17 hep-th hep-lat hep-ph 78%

Probing Quarkonium Diffusion in a Magnetized Quark-Gluon Plasma

Siddhi Swarupa Jena, Arpan Bhattacharjee, David Dudal, Subhash Mahapatra

专题命中 扩散模型 :diffusion(title,abstract)

Comments 40 pages, 12 figures, typos corrected, references added. Published version

Journal ref Phys. Rev. D 112 (2025) 086010

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08403 2025-10-17 cs.LG cs.AI stat.ML 78%

ConDiSim: Conditional Diffusion Models for Simulation Based Inference

Mayank Nautiyal, Andreas Hellander, Prashant Singh

机构 * Science for Life Laboratory, Uppsala University(生命科学实验室,乌普萨拉大学) Uppsala University(乌普萨拉大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03924 2025-10-17 cs.IT cs.LG eess.SP math.IT 78%

Generating High Dimensional User-Specific Wireless Channels using Diffusion Models

Taekyun Lee, Juseong Park, Hyeji Kim, Jeffrey G. Andrews

机构 * University of Texas at Austin(德克萨斯大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.16196 2025-10-17 math.PR 78%

Asymptotically unbiased approximation of the QSD of diffusion processes with a decreasing time step Euler scheme

Fabien Panloup, Julien Reygner

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09656 2025-10-17 cs.CV 74%

KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation

Xingrui Wang, Jiang Liu, Ze Wang, Xiaodong Yu, Jialian Wu, Ximeng Sun, Yusheng Su, Alan Yuille, Zicheng Liu, Emad Barsoum

专题命中 扩散模型 :diffusion(title);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13325 2025-10-17 physics.soc-ph cs.SI q-bio.PE 71%

A data-driven analysis of the impact of non-compliant individuals on epidemic diffusion in urban settings

Fabio Mazza, Marco Brambilla, Carlo Piccardi, Francesco Pierri

专题命中 扩散模型 :diffusion(title)

Comments 20 pages, 10 figures

Journal ref Royal Society Proceedings A, 2025, Volume 481, Issue 2324

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14230 2025-10-17 cs.CV 70%

LOTA: Bit-Planes Guided AI-Generated Image Detection

Hongsong Wang, Renxi Cheng, Yang Zhang, Chaolei Han, Jie Gui

机构 * School of Computer Science and Engineering, Southeast University, Nanjing 210096, China(东南大学计算机科学与工程学院) Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及其交叉应用重点实验室) School of Cyber Science and Engineering, Southeast University, Nanjing 210096, China(东南大学网络安全工程学院) School of Computer Science and Software Engineering, Shenzhen University, Shenzhen 518060, China(深圳大学计算机科学与软件工程学院) Purple Mountain Laboratories, Nanjing 210000, China(紫金山实验室) Engineering Research Center of Blockchain Application, Supervision And Management (Southeast University), Ministry of Education, China(区块链应用、监督与管理工程研究中心)

专题命中 扩散模型 :image generation(abstract);diffusion(abstract);分类 cs.CV

Comments Published in the ICCV2025, COde is https://github.com/hongsong-wang/LOTA

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23402 2025-10-17 cs.CV 57%

WorldSplat: Gaussian-Centric Feed-Forward 4D Scene Generation for Autonomous Driving

Ziyue Zhu, Zhanqian Wu, Zhenxin Zhu, Lijun Zhou, Haiyang Sun, Bing Wan, Kun Ma, Guang Chen, Hangjun Ye, Jin Xie, jian Yang

机构 * Nankai University(南开大学) Nanjing University, Suzhou(南京大学苏州校区)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01126 2025-10-17 cs.CV 57%

UniEgoMotion: A Unified Model for Egocentric Motion Reconstruction, Forecasting, and Generation

Chaitanya Patel, Hiroki Nakamura, Yuta Kyuragi, Kazuki Kozuka, Juan Carlos Niebles, Ehsan Adeli

机构 * Stanford University(斯坦福大学) Panasonic Holdings Corporation(松下控股公司) Panasonic R&D Company of America(美国松下研发公司)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments ICCV 2025. Project Page: https://chaitanya100100.github.io/UniEgoMotion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08209 2025-10-17 cs.CV cs.AI cs.LG 57%

Emergent Visual Grounding in Large Multimodal Models Without Grounding Supervision

Shengcao Cao, Liang-Yan Gui, Yu-Xiong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments ICCV 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏