arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-11-19 至 2025-11-19 共收录 46 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 46 篇

2403.20105 2025-11-19 cs.CV 85%

FreeSeg-Diff: Training-Free Open-Vocabulary Segmentation with Diffusion Models

Barbara Toniella Corradini, Mustafa Shukor, Paul Couairon, Guillaume Couairon, Franco Scarselli, Matthieu Cord

机构 * DIISM University of Siena(DIISM锡耶纳大学) CNRS, ISIR Sorbonne University(CNRS,ISIR索邦大学) Inria, ARCHES Sorbonne University(Inria,ARCHES索邦大学) Valeo.ai Sorbonne University(Valeo.ai索邦大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);text-to-image(abstract);分类 cs.CV

Journal ref Proceedings of the 2025 International Joint Conference on Neural Networks (IJCNN 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14113 2025-11-19 cs.CV 83%

Coffee: Controllable Diffusion Fine-tuning

Ziyao Zeng, Jingcheng Ni, Ruyi Liu, Alex Wong

机构 * Yale University(耶鲁大学) Brown University(布朗大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13689 2025-11-19 cs.CL cs.CV 83%

Crossing Borders: A Multimodal Challenge for Indian Poetry Translation and Image Generation

Sofia Jamil, Kotla Sai Charan, Sriparna Saha, Koustava Goswami, Joseph K J

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09611 2025-11-19 cs.CV 83%

MMaDA-Parallel: Multimodal Large Diffusion Language Models for Thinking-Aware Editing and Generation

Ye Tian, Ling Yang, Jiongfan Yang, Anran Wang, Yu Tian, Jiani Zheng, Haochen Wang, Zhiyang Teng, Zhuochen Wang, Yinjie Wang, Yunhai Tong, Mengdi Wang, Xiangtai Li

机构 * Peking University(北京大学) ByteDance(字节跳动) Princeton University(普林斯顿大学) CASIA(中国科学院自动化研究所) The University of Chicago(芝加哥大学)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

Comments Project Page: https://tyfeld.github.io/mmadaparellel.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26324 2025-11-19 cs.LG cs.AI cs.DS math.ST stat.ML stat.TH 82%

Posterior Sampling by Combining Diffusion Models with Annealed Langevin Dynamics

Zhiyang Xun, Shivam Gupta, Eric Price

机构 * UT Austin(得克萨斯大学) Microsoft Research(微软研究院)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract)

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23752 2025-11-19 cs.GR cs.CV 81%

StrokeFusion: Vector Sketch Generation via Joint Stroke-UDF Encoding and Latent Sequence Diffusion

Jin Zhou, Yi Zhou, Hongliang Yang, Pengfei Xu, Hui Huang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14716 2025-11-19 cs.CV 79%

Diffusion As Self-Distillation: End-to-End Latent Diffusion In One Model

Xiyuan Wang, Muhan Zhang

机构 * Institute of Artificial Intelligence(人工智能研究院) Peking University(北京大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Tech Report. 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14481 2025-11-19 cs.CV 79%

Segmentation-Aware Latent Diffusion for Satellite Image Super-Resolution: Enabling Smallholder Farm Boundary Delineation

Aditi Agarwal, Anjali Jain, Nikita Saxena, Ishan Deshpande, Michal Kazmierski, Abigail Annkah, Nadav Sherman, Karthikeyan Shanmugam, Alok Talekar, Vaibhav Rajan

机构 * Google DeepMind(谷歌DeepMind) Google(谷歌) Google Research(谷歌研究)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12089 2025-11-19 cs.CV 79%

Playmate2: Training-Free Multi-Character Audio-Driven Animation via Diffusion Transformer with Reward Feedback

Xingpei Ma, Shenneng Huang, Jiaran Cai, Yuansheng Guan, Shen Zheng, Hanfeng Zhao, Qiang Zhang, Shunsi Zhang

机构 * Project lead & Corresponding Author(项目负责人及通讯作者)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01196 2025-11-19 cs.RO cs.AI cs.CV 79%

OG-VLA: Orthographic Image Generation for 3D-Aware Vision-Language Action Model

Ishika Singh, Ankit Goyal, Stan Birchfield, Dieter Fox, Animesh Garg, Valts Blukis

机构 * University of Southern California(美国南加州大学) NVIDIA(英伟达)

专题命中 扩散模型 :image generation(title);diffusion(abstract);分类 cs.CV

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14033 2025-11-19 cs.CV 79%

Flood-LDM: Generalizable Latent Diffusion Models for rapid and accurate zero-shot High-Resolution Flood Mapping

Sun Han Neo, Sachith Seneviratne, Herath Mudiyanselage Viraj Vidura Herath, Abhishek Saha, Sanka Rasnayaka, Lucy Amanda Marshall

机构 * Department of Computer Science, School of Computing, National University of Singapore(新加坡国立大学计算机科学系) Transport, Health and Urban Systems Research Lab, Melbourne School of Design, University of Melbourne(墨尔本大学设计学院交通、健康与城市系统研究实验室) School of Civil Engineering, Faculty of Engineering, University of Sydney(悉尼大学土木工程学院) Delft Institute of Applied Mathematics, Delft University of Technology(代尔夫特理工大学应用数学研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted for publication at the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13795 2025-11-19 cs.CV cs.AI cs.RO 79%

A Trajectory-free Crash Detection Framework with Generative Approach and Segment Map Diffusion

Weiying Shen, Hao Yu, Yu Dong, Pan Liu, Yu Han, Xin Wen

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments To be presented at TRB 2026 (TRBAM-26-01711) and a revised version will be submitted to Transportation Research Part C: Emerging Technologies

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04376 2025-11-19 cs.SD cs.AI cs.LG cs.MM eess.AS 79%

MusRec: Zero-Shot Text-to-Music Editing via Rectified Flow and Diffusion Transformers

Ali Boudaghi, Hadi Zare

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14871 2025-11-19 cs.LG cs.CV 79%

Squeezed Diffusion Models

Jyotirmai Singh, Samar Khanna, James Burgess

机构 * Stanford University(斯坦福大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 7 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21714 2025-11-19 cs.LG cs.CV 79%

ODE$_t$(ODE$_l$): Shortcutting the Time and the Length in Diffusion and Flow Models for Faster Sampling

Denis Gudovskiy, Wenzhao Zheng, Tomoyuki Okuno, Yohei Nakata, Kurt Keutzer

机构 * Panasonic AI Lab(松下人工智能实验室) UC Berkeley(加州大学伯克利分校) Panasonic DX-CPS(松下DX-CPS)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to WACV 2026. Preprint. Github page: github.com/gudovskiy/odelt

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14327 2025-11-19 eess.IV cs.AI cs.CV 79%

Autoregressive Image Diffusion: Generation of Image Sequence and Application in MRI

Guanxiong Luo, Shoujin Huang, Martin Uecker

机构 * University Medical Center Göttingen(哥廷根大学医学中心) Shenzhen Technology University(深圳科技大学) Graz University of Technology(格拉茨技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Advances in Neural Information Processing Systems 2024;37:129094-129119

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13745 2025-11-19 cs.LG cs.CV cs.IT math.IT math.ST stat.ML stat.TH 79%

Improved Sample Complexity Bounds for Diffusion Model Training

Shivam Gupta, Aditya Parulekar, Eric Price, Zhiyang Xun

机构 * UT Austin(德克萨斯大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Bugfix

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14667 2025-11-19 astro-ph.CO astro-ph.IM 78%

High-resolution weak lensing mass mapping from DES-Y3 data using diffusion-based prior

Supranta S. Boruah, Michael Jacob, Bhuvnesh Jain, Riya Maiya, Raghav Venkataramanan

专题命中 扩散模型 :diffusion(title,abstract)

Comments 6 pages, 3 figures, Accepted at Neurips Machine Learning and the Physical Sciences workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14559 2025-11-19 q-bio.BM cs.AI cs.LG q-bio.QM 78%

Apo2Mol: 3D Molecule Generation via Dynamic Pocket-Aware Diffusion Models

Xinzhe Zheng, Shiyu Jiang, Gustavo Seabra, Chenglong Li, Yanjun Li

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14543 2025-11-19 cs.LG cs.AI 78%

MissHDD: Hybrid Deterministic Diffusion for Hetrogeneous Incomplete Data Imputation

Youran Zhou, Mohamed Reda Bouadjenek, Sunil Aryal

机构 * Deakin University(德克萨斯大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14333 2025-11-19 math.ST stat.TH 78%

Akaike-type information criterion of SEM for jump-diffusion processes based on high-frequency data

Shogo Kusano, Masayuki Uchida

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04325 2025-11-19 cs.LG physics.flu-dyn 78%

FoilDiff: A Hybrid Transformer Backbone for Diffusion-based Modelling of 2D Airfoil Flow Fields

Kenechukwu Ogbuagu, Sepehr Maleki, Giuseppe Bruni, Senthil Krishnababu

机构 * Lincoln AI Lab(林肯人工智能实验室) School of Engineering and Physical Sciences(工程与物理科学学院) University of Lincoln(林肯大学) Siemens Energy(西门子能源)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18720 2025-11-19 cs.LG physics.ao-ph 78%

Appa: Bending Weather Dynamics with Latent Diffusion Models for Global Data Assimilation

Gérôme Andry, Sacha Lewin, François Rozet, Omer Rochman, Victor Mangeleer, Matthias Pirlet, Elise Faulx, Marilaure Grégoire, Gilles Louppe

机构 * University of Liège(利根大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14104 2025-11-19 eess.SP 78%

Lightweight Multi-task CNN for ECG Diagnosis with GRU-Diffusion

Lehuai Xu, Zirui Lu, Haoran Yang, Yina Zhou

专题命中 扩散模型 :diffusion(title,abstract)

Comments 15 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13937 2025-11-19 cs.LG cs.SI math.DS physics.soc-ph 78%

Complex-Weighted Convolutional Networks: Provable Expressiveness via Complex Diffusion

Cristina López Amado, Tassilo Schwarz, Yu Tian, Renaud Lambiotte

机构 * IST Austria(IST奥地利研究院) Center for Systems Biology Dresden(德累斯顿系统生物学中心) Mathematical bioPhysics Group, Max Planck Institute for Multidisciplinary Sciences(多学科科学研究所数学生物物理组) Max Planck Institute of Molecular Cell Biology and Genetics(分子细胞生物学和遗传学马克斯·普朗克研究所) Max Planck Institute for the Physics of Complex Systems(复杂系统物理马克斯·普朗克研究所) Dresden University of Technology(德累斯顿技术大学)

专题命中 扩散模型 :diffusion(title,abstract)

Comments 19 pages, 6 figures. Learning on Graphs Conference 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12597 2025-11-19 cs.IR 78%

MindRec: A Diffusion-driven Coarse-to-Fine Paradigm for Generative Recommendation

Mengyao Gao, Chongming Gao, Haoyan Liu, Qingpeng Cai, Peng Jiang, Jiajia Chen, Shuai Yuan, Xiangnan He

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08252 2025-11-19 cs.SD eess.AS 78%

Melodia: Training-Free Music Editing Guided by Attention Probing in Diffusion Models

Yi Yang, Haowen Li, Tianxiang Li, Boyu Cao, Xiaohan Zhang, Liqun Chen, Qi Liu

专题命中 扩散模型 :diffusion(title,abstract)

Comments AAAI 2026 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10040 2025-11-19 math.AP cs.NA math.NA 71%

Drift-diffusion equations with saturation

José Antonio Carrillo, Alejandro Fernández-Jiménez, David Gómez-Castro

专题命中 扩散模型 :diffusion(title)

Comments 52 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13442 2025-11-19 cs.CV cs.AI 70%

Unlocking the Forgery Detection Potential of Vanilla MLLMs: A Novel Training-Free Pipeline

Rui Zuo, Qinyue Tong, Zhe-Ming Lu, Ziqian Lu

机构 * Zhejiang University(浙江大学) Zhejiang Sci-Tech University(浙江科技学院)

专题命中 扩散模型 :image generation(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14719 2025-11-19 cs.CV cs.AI 57%

Zero-shot Synthetic Video Realism Enhancement via Structure-aware Denoising

Yifan Wang, Liya Ji, Zhanghan Ke, Harry Yang, Ser-Nam Lim, Qifeng Chen

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Everlyn AI University of Central Florida(中央佛罗里达大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Project Page: https://wyf0824.github.io/Video_Realism_Enhancement/

详情

展开后加载摘要…

URL PDF HTML 收藏