arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-17 至 2025-10-17 共收录 56 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 2 篇

2405.17965 2025-10-17 cs.CV 83%

AttenCraft: Attention-guided Disentanglement of Multiple Concepts for Text-to-Image Customization

Junjie Shentu, Matthew Watson, Noura Al Moubayed

机构 * Department of Computer Science, Durham University(计算机科学系,达勒姆大学)

专题命中 文生图 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10623 2025-10-17 cs.LG cs.CV 70%

Flows and Diffusions on the Neural Manifold

Daniel Saragih, Deyu Cao, Tejas Balaji

机构 * Queen’s University and Vector Institute(女王大学和向量研究所) University of Tokyo(东京大学) University of Toronto(多伦多大学)

专题命中 文生图 :diffusion(abstract);image synthesis(abstract);分类 cs.CV

Comments 43 pages, 11 figures, 14 tables

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 图像编辑 2 篇

2506.11863 2025-10-17 cs.CV 79%

SphereDrag: Spherical Geometry-Aware Panoramic Image Editing

Zhiao Feng, Xuewei Li, Junjie Yang, Jingchao Li, Yuxin Peng, Xi Li

机构 * School of Electronic and Information Engineering, Shanghai DianJi University(电子信息学院,上海电力大学) Department of Sports Science, Zhejiang University(体育学院,浙江大学) College of Computer Science and Technology, Zhejiang University(计算机科学与技术学院,浙江大学)

专题命中 图像编辑 :image editing(title,abstract);分类 cs.CV

Comments Accepted by PRCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14648 2025-10-17 cs.CV cs.AI 57%

In-Context Learning with Unpaired Clips for Instruction-based Video Editing

Xinyao Liao, Xianfang Zeng, Ziye Song, Zhoujie Fu, Gang Yu, Guosheng Lin

机构 * Nanyang Technological University(南洋理工大学) StepFun

专题命中 图像编辑 :image editing(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 扩散模型 39 篇

2510.14526 2025-10-17 cs.CV cs.LG 89%

Noise Projection: Closing the Prompt-Agnostic Gap Behind Text-to-Image Misalignment in Diffusion Models

Yunze Tong, Didi Zhu, Zijing Hu, Jinluan Yang, Ziyu Zhao

机构 * Zhejiang University(浙江大学)

专题命中 扩散模型 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Appendix will be appended soon

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14981 2025-10-17 cs.CV cs.AI 88%

Coupled Diffusion Sampling for Training-Free Multi-View Image Editing

Hadi Alzayer, Yunzhi Zhang, Chen Geng, Jia-Bin Huang, Jiajun Wu

机构 * Stanford University(斯坦福大学) University of Maryland, College Park(马里兰大学学院公园分校)

专题命中 扩散模型 :diffusion(title,abstract);image editing(title,abstract);分类 cs.CV

Comments Project page: https://coupled-diffusion.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18033 2025-10-17 cs.CV 83%

OmnimatteZero: Fast Training-free Omnimatte with Pre-trained Video Diffusion Models

Dvir Samuel, Matan Levy, Nir Darshan, Gal Chechik, Rami Ben-Ari

机构 * Bar-Ilan University \& OriginAI Israel The Hebrew University of Jerusalem Israel Bar-Ilan University \& NVIDIA Research Israel Bar-Ilan University \& OriginAI The Hebrew University of Jerusalem Bar-Ilan University \& NVIDIA Research

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

Comments Accepted to SIGGRAPH ASIA 2025. Project Page: https://dvirsamuel.github.io/omnimattezero.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14427 2025-10-17 cs.MM cs.CV 81%

Deep Compositional Phase Diffusion for Long Motion Sequence Generation

Ho Yin Au, Jie Chen, Junkun Jiang, Jingyu Xiang

机构 * Hong Kong Baptist University(香港 Baptist 大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

Comments Accepted by NeurIPS 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14962 2025-10-17 cs.CV 79%

RainDiff: End-to-end Precipitation Nowcasting Via Token-wise Attention Diffusion

Thao Nguyen, Jiaqi Ma, Fahad Shahbaz Khan, Souhaib Ben Taieb, Salman Khan

机构 * Mohamed Bin Zayed University of AI(穆罕默德·本·扎耶德人工智能大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14634 2025-10-17 cs.CV 79%

SteeringTTA: Guiding Diffusion Trajectories for Robust Test-Time-Adaptation

Jihyun Yu, Yoojin Oh, Wonho Bae, Mingyu Kim, Junhyug Noh

机构 * Department of Artificial Intelligence, Ewha Womans University, Seoul, Republic of Korea(人工智能系,世ultan大学,韩国首尔) Department of Computer Science, University of British Columbia, Vancouver, Canada(计算机科学系,不列颠哥伦比亚大学,加拿大温哥华)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14314 2025-10-17 cs.CV 79%

A Multi-domain Image Translative Diffusion StyleGAN for Iris Presentation Attack Detection

Shivangi Yadav, Arun Ross

机构 * Michigan State University(密歇根州立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07944 2025-10-17 cs.CV 79%

CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving

Tianrui Zhang, Yichen Liu, Zilin Guo, Yuxin Guo, Jingcheng Ni, Chenjing Ding, Dan Xu, Lewei Lu, Zehuan Wu

机构 * Sensetime Research(商汤科技研究院) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10569 2025-10-17 cs.CR cs.AI cs.MM 79%

MarkDiffusion: An Open-Source Toolkit for Generative Watermarking of Latent Diffusion Models

Leyi Pan, Sheng Guan, Zheyu Fu, Luyang Si, Huan Wang, Zian Wang, Hanqian Li, Xuming Hu, Irwin King, Philip S. Yu, Aiwei Liu, Lijie Wen

机构 * Tsinghua University(清华大学) Beijing University of Posts and Telecommunications(北京邮电大学) The Chinese University of Hong Kong(香港中文大学) University of Illinois at Chicago(伊利诺伊大学香槟分校) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

Comments 23 pages, 13 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09789 2025-10-17 cs.CV cs.AI 79%

On Equivariance and Fast Sampling in Video Diffusion Models Trained with Warped Noise

Chao Liu, Arash Vahdat

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15008 2025-10-17 cs.CV 79%

HuGDiffusion: Generalizable Single-Image Human Rendering via 3D Gaussian Diffusion

Yingzhi Tang, Qijian Zhang, Junhui Hou

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14961 2025-10-17 cs.LG cs.CL 78%

Efficient Parallel Samplers for Recurrent-Depth Models and Their Connection to Diffusion Language Models

Jonas Geiping, Xinyu Yang, Guinan Su

机构 * ELLIS Institute Tübingen & Max-Planck Institute for Intelligent Systems, Tübingen AI Center(图宾根ELLIS研究所及图宾根马克斯·普朗克智能系统研究所AI中心)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Code can be found at https://github.com/seal-rg/recurrent-pretraining

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14467 2025-10-17 cs.RO 78%

Restoring Noisy Demonstration for Imitation Learning With Diffusion Models

Shang-Fu Chen, Co Yong, Shao-Hua Sun

机构 * Graduate Institute of Communication Engineering, National Taiwan University(国立台湾大学通信工程研究所) Data Science Degree Program, National Taiwan University and Academia Sinica(国立台湾大学数据科学学士学位计划) Department of Electrical Engineering, National Taiwan University(国立台湾大学电子工程系)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Published in IEEE Transactions on Neural Networks and Learning Systems (TNNLS)

Journal ref IEEE Transactions on Neural Networks and Learning Systems (TNNLS), pp. 1-13, Sept. 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14269 2025-10-17 cs.LG stat.ML 78%

Nonparametric Data Attribution for Diffusion Models

Yutian Zhao, Chao Du, Xiaosen Zheng, Tianyu Pang, Min Lin

机构 * Sea AI Lab, Singapore(新加坡海人工智能实验室) Department of Mathematics, National University of Singapore(新加坡国立大学数学系) Singapore Management University(新加坡管理大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14114 2025-10-17 cs.LG 78%

Briding Diffusion Posterior Sampling and Monte Carlo methods: a survey

Yazid Janati, Alain Durmus, Jimmy Olsson, Eric Moulines

机构 * Institute of Foundation Models Paris, MBZUAI(基础模型研究所巴黎,MBZUAI) KTH Royal Institute of Technology(皇家理工学院) Ecole polytechnique(巴黎高等理工学院) MBZUAI

专题命中 扩散模型 :diffusion(title,abstract)

Journal ref Philosophical Transactions A, 383(2299), 20240331 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14111 2025-10-17 cs.NI 78%

DiffLoc: Diffusion Model-Based High-Precision Positioning for 6G Networks

Taekyun Lee, Tommaso Balercia, Heasung Kim, Hyeji Kim, Jeffrey G. Andrews

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14094 2025-10-17 cs.LG 78%

Neural Network approximation power on homogeneous and heterogeneous reaction-diffusion equations

Haotian Feng

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12853 2025-10-17 astro-ph.IM math-ph math.MP 78%

Diffusion models for polarimetric reconstruction of circumstellar environments

Quentin Villegas, Laurence Denneulin, Simon Prunet, André Ferrari, Nelly Pustelnik, Éric Thiébaut, Julian Tachella, Maud Langlois

专题命中 扩散模型 :diffusion(title,abstract)

Comments in French language. GRETSI 2025 -- XXXe Colloque sur le Traitement du Signal et des Images, Aug 2025, Strasboug, France

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12210 2025-10-17 eess.AS cs.CL cs.LG 78%

DiSTAR: Diffusion over a Scalable Token Autoregressive Representation for Speech Generation

Yakun Song, Xiaobin Zhuang, Jiawei Chen, Zhikang Niu, Guanrou Yang, Chenpeng Du, Dongya Jia, Zhuo Chen, Yuping Wang, Yuxuan Wang, Xie Chen

机构 * X-LANCE Lab, School of Computer Science, Shanghai Jiao Tong University(X-LANCE实验室,计算机科学学院,上海交通大学) ByteDance Inc.(字节跳动公司)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09383 2025-10-17 cs.RO cs.LG 78%

Real-Time Adaptive Motion Planning via Point Cloud-Guided, Energy-Based Diffusion and Potential Fields

Wondmgezahu Teshome, Kian Behzad, Octavia Camps, Michael Everett, Milad Siami, Mario Sznaier

机构 * ECE Department, Northeastern University(东北大学电子与计算机工程系)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted to IEEE RA-L 2025

Journal ref IEEE Robotics and Automation Letters 10 (2025) 9160-9167

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00746 2025-10-17 hep-th hep-lat hep-ph 78%

Probing Quarkonium Diffusion in a Magnetized Quark-Gluon Plasma

Siddhi Swarupa Jena, Arpan Bhattacharjee, David Dudal, Subhash Mahapatra

专题命中 扩散模型 :diffusion(title,abstract)

Comments 40 pages, 12 figures, typos corrected, references added. Published version

Journal ref Phys. Rev. D 112 (2025) 086010

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08403 2025-10-17 cs.LG cs.AI stat.ML 78%

ConDiSim: Conditional Diffusion Models for Simulation Based Inference

Mayank Nautiyal, Andreas Hellander, Prashant Singh

机构 * Science for Life Laboratory, Uppsala University(生命科学实验室,乌普萨拉大学) Uppsala University(乌普萨拉大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03924 2025-10-17 cs.IT cs.LG eess.SP math.IT 78%

Generating High Dimensional User-Specific Wireless Channels using Diffusion Models

Taekyun Lee, Juseong Park, Hyeji Kim, Jeffrey G. Andrews

机构 * University of Texas at Austin(德克萨斯大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.16196 2025-10-17 math.PR 78%

Asymptotically unbiased approximation of the QSD of diffusion processes with a decreasing time step Euler scheme

Fabien Panloup, Julien Reygner

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09656 2025-10-17 cs.CV 74%

KeyVID: Keyframe-Aware Video Diffusion for Audio-Synchronized Visual Animation

Xingrui Wang, Jiang Liu, Ze Wang, Xiaodong Yu, Jialian Wu, Ximeng Sun, Yusheng Su, Alan Yuille, Zicheng Liu, Emad Barsoum

专题命中 扩散模型 :diffusion(title);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13325 2025-10-17 physics.soc-ph cs.SI q-bio.PE 71%

A data-driven analysis of the impact of non-compliant individuals on epidemic diffusion in urban settings

Fabio Mazza, Marco Brambilla, Carlo Piccardi, Francesco Pierri

专题命中 扩散模型 :diffusion(title)

Comments 20 pages, 10 figures

Journal ref Royal Society Proceedings A, 2025, Volume 481, Issue 2324

详情

展开后加载摘要…

URL PDF HTML 收藏