arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-13 至 2025-10-13 共收录 56 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 4 篇

2510.08994 2025-10-13 cs.CV 86%

Speculative Jacobi-Denoising Decoding for Accelerating Autoregressive Text-to-image Generation

Yao Teng, Fuyun Wang, Xian Liu, Zhekai Chen, Han Shi, Yu Wang, Zhenguo Li, Weiyang Liu, Difan Zou, Xihui Liu

机构 * The University of Hong Kong(香港大学) CUHK(香港中文大学) Huawei Noah’s Ark Lab(华为诺亚实验室) Tsinghua University(清华大学)

专题命中 文生图 :text-to-image(title,abstract);image generation(title);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.03815 2025-10-13 cs.CV cs.MM eess.IV 73%

T2IW: Joint Text to Image & Watermark Generation

An-An Liu, Guokai Zhang, Yuting Su, Ning Xu, Yongdong Zhang, Lanjun Wang

专题命中 文生图 :image generation(abstract);text-to-image(abstract);分类 cs.CV、cs.MM

Journal ref Machine Intelligence Research, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03355 2025-10-13 cs.LG cs.AI cs.CV 70%

Robustness in Both Domains: CLIP Needs a Robust Text Encoder

Elias Abad Rocamora, Christian Schlarmann, Naman Deep Singh, Yongtao Wu, Matthias Hein, Volkan Cevher

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments Accepted in NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08951 2025-10-13 eess.IV cs.CV 57%

FS-RWKV: Leveraging Frequency Spatial-Aware RWKV for 3T-to-7T MRI Translation

Yingtie Lei, Zimeng Li, Chi-Man Pun, Yupeng Liu, Xuhang Chen

机构 * Faculty of Science and Technology, University of Macau(澳门大学科学与技术学院) School of Electronic and Communication Engineering, Shenzhen Polytechnic University(深圳职业技术学院电子与通信工程学院) Department of Cardiology, Guangdong Provincial People’s Hospital (Guangdong Academy of Medical Sciences), Southern Medical University, Guangzhou, China(广东省人民医院心内科(广东省医学科学院)) Guangdong Cardiovascular Institute, Guangdong Provincial People’s Hospital, Guangdong Academy of Medical Sciences, Guangzhou, China(广东省心血管病研究院,广东省人民医院,广东省医学科学院,广州,中国)

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted by BIBM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 扩散模型 37 篇

2510.09094 2025-10-13 cs.CV 92%

Dense2MoE: Restructuring Diffusion Transformer to MoE for Efficient Text-to-Image Generation

Youwei Zheng, Yuxi Ren, Xin Xia, Xuefeng Xiao, Xiaohua Xie

机构 * Sun Yat-sen University(中山大学) ByteDance Intelligent Creation(字节跳动智能创作) ByteDance Seed Vision(字节跳动种子视觉) Guangdong Province Key Laboratory of Information Security Technology(广东省信息安全技术重点实验室) Pazhou Lab (Huangpu)(琶洲实验室(黄埔))

专题命中 扩散模型 :image generation(title,abstract);text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08625 2025-10-13 cs.CV 88%

Adjusting Initial Noise to Mitigate Memorization in Text-to-Image Diffusion Models

Hyeonggeun Han, Sehwan Kim, Hyungjun Joo, Sangwoo Hong, Jungwoo Lee

机构 * Seoul National University(首尔国立大学) CSE, Konkuk University(计算机科学与工程系,konkuk大学) Hodoo AI Labs(Hodoo AI实验室)

专题命中 扩散模型 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01428 2025-10-13 cs.CV eess.IV 83%

DiffMark: Diffusion-based Robust Watermark Against Deepfakes

Chen Sun, Haiyang Sun, Zhiqing Guo, Yunfeng Diao, Liejun Wang, Dan Ma, Gaobo Yang, Keqin Li

机构 * College of Computer Science and Technology, Xinjiang University(新疆大学计算机科学与技术学院) Silk Road Multilingual Cognitive Computing International Cooperation Joint Laboratory(丝绸之路多语种认知计算国际合作联合实验室) School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院) College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院) Department of Computer Science, State University of New York(纽约州立大学计算机科学系)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03016 2025-10-13 cs.LG cs.AI 82%

Learning Robust Diffusion Models from Imprecise Supervision

Dong-Dong Wu, Jiacheng Cui, Wei Wang, Zhiqiang Shen, Masashi Sugiyama

机构 * The University of Tokyo(东京大学) RIKEN AIP(日本理化学研究所人工智能研究中心) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.02851 2025-10-13 cs.CV cs.GR 81%

Human-VDM: Learning Single-Image 3D Human Gaussian Splatting from Video Diffusion Models

Zhibin Liu, Haoye Dong, Aviral Chharia, Hefeng Wu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments 14 Pages, 8 figures, Project page: https://human-vdm.github.io/Human-VDM/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11128 2025-10-13 cs.LG cs.CV 79%

What's Inside Your Diffusion Model? A Score-Based Riemannian Metric to Explore the Data Manifold

Simone Azeglio, Arianna Di Bernardo

机构 * Institut de la Vision & Laboratoire des Systèmes Perceptifs(视觉研究所及感知系统实验室) Sorbonne Université & École Normale Supérieure(索邦大学及巴黎高等师范学院) Group for Neural Theory(神经理论小组) École Normale Supérieure(巴黎高等师范学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03948 2025-10-13 cs.CV 79%

ProbRes: Probabilistic Jump Diffusion for Open-World Egocentric Activity Recognition

Sanjoy Kundu, Shanmukha Vellamcheti, Sathyanarayanan N. Aakur

机构 * CSSE Department, Auburn University(计算机科学与工程系,阿伯丁大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICCV 2025. 17 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09056 2025-10-13 cs.CV 79%

Lesion-Aware Post-Training of Latent Diffusion Models for Synthesizing Diffusion MRI from CT Perfusion

Junhyeok Lee, Hyunwoong Kim, Hyungjin Chung, Heeseong Eom, Joon Jang, Chul-Ho Sohn, Kyu Sung Choi

机构 * College of Medicine, Seoul National University, Seoul, Republic of Korea(首尔国立大学医学院) Department of Radiology, Seoul National University Hospital(首尔国立大学医院放射科) Department of Biomedical Sciences, Seoul National University(首尔国立大学生物医学科学系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments MICCAI 2025, Lecture Notes in Computer Science Vol. 15961

Journal ref Med Image Comput Comput Assist Interv. LNCS 15961, 282-291, Springer, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08669 2025-10-13 cs.LG cs.AI cs.CV 79%

FreqCa: Accelerating Diffusion Models via Frequency-Aware Caching

Jiacheng Liu, Peiliang Cai, Qinming Zhou, Yuqi Lin, Deyang Kong, Benhao Huang, Yupei Pan, Haowen Xu, Chang Zou, Junshu Tang, Shikang Zheng, Linfeng Zhang

机构 * EPIC Lab,STJU(EPIC实验室,STJU) Tencent Hunyuan(腾讯文言) SDU THU(清华大学) JLU(吉林大学) CMU(卡内基梅隆大学) UESTC SCUT(华南理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21787 2025-10-13 cs.CV cs.CL 79%

DeHate: A Stable Diffusion-based Multimodal Approach to Mitigate Hate Speech in Images

Dwip Dalal, Gautam Vashishtha, Anku Rani, Aishwarya Reganti, Parth Patwa, Mohd Sarique, Chandan Gupta, Keshav Nath, Viswanatha Reddy, Vinija Jain, Aman Chadha, Amitava Das, Amit Sheth, Asif Ekbal

机构 * MIT Media Lab, USA(麻省理工学院媒体实验室) Stanford University, USA(斯坦福大学) Amazon GenAI, USA(亚马逊生成人工智能) University of South Carolina, USA(南卡罗来纳大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Defactify 3 workshop at AAAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03517 2025-10-13 cs.CV 79%

DenseDPO: Fine-Grained Temporal Preference Optimization for Video Diffusion Models

Ziyi Wu, Anil Kag, Ivan Skorokhodov, Willi Menapace, Ashkan Mirzaei, Igor Gilitschenski, Sergey Tulyakov, Aliaksandr Siarohin

机构 * Snap Research University of Toronto(多伦多大学) Vector Institute(向量研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments NeurIPS 2025 Spotlight. Project page: https://snap-research.github.io/DenseDPO/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09544 2025-10-13 cs.CL 78%

Beyond Surface Reasoning: Unveiling the True Long Chain-of-Thought Capacity of Diffusion Large Language Models

Qiguang Chen, Hanjing Li, Libo Qin, Dengyun Peng, Jinhao Liu, Jiangyi Wang, Chengyue Wu, Xie Chen, Yantao Du, Wanxiang Che

机构 * LARG, Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology(LARG,社会计算与交互机器人研究中心,哈尔滨工业大学) School of Computer Science and Engineering, Central South University(计算机科学与工程学院,中南大学) The University of Hong Kong(香港大学) Shanghai Jiao Tong University(上海交通大学) ByteDance Seed (China)(字节跳动种子(中国))

专题命中 扩散模型 :diffusion(title,abstract)

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09479 2025-10-13 math-ph math.MP 78%

Modeling Protein Diffusion Across ER-Nuclear Envelope Junctions Reveals Efficient Transport via Simple Diffusion

Sara Merino-Aceituno, Carmela Moschella, Shotaro Otsuka, Christian Schmeiser, Julia Scholz

专题命中 扩散模型 :diffusion(title,abstract)

Comments 15 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09449 2025-10-13 math.NA cs.NA math.AP 78%

A posteriori analysis for nonlinear convection-diffusion systems

Andreas Dedner, Jan Giesselmann, Kiwoong Kwon, Tristan Pryer

专题命中 扩散模型 :diffusion(title,abstract)

Comments 33 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08744 2025-10-13 cs.LG cs.AI 78%

Graph Diffusion Transformers are In-Context Molecular Designers

Gang Liu, Jie Chen, Yihan Zhu, Michael Sun, Tengfei Luo, Nitesh V Chawla, Meng Jiang

机构 * University of Notre Dame(诺丁汉大学) MIT-IBM Watson AI Lab, IBM Research(麻省理工-IBM Watson人工智能实验室,IBM研究)

专题命中 扩散模型 :diffusion(title,abstract)

Comments 29 pages, 16 figures, 17 tables. Model available at: https://huggingface.co/liuganghuggingface/DemoDiff-0.7B

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08632 2025-10-13 cs.CL cs.LG 78%

Next Semantic Scale Prediction via Hierarchical Diffusion Language Models

Cai Zhou, Chenyu Wang, Dinghuai Zhang, Shangyuan Tong, Yifei Wang, Stephen Bates, Tommi Jaakkola

机构 * Massachusetts Institute of Technology(麻省理工学院) Microsoft Research(微软研究院) Mila - Quebec AI Institute(魁北克AI研究所)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08627 2025-10-13 cs.NE cs.DM 78%

A Denoising Diffusion-Based Evolutionary Algorithm Framework: Application to the Maximum Independent Set Problem

Joan Salvà Soler, Günther R. Raidl

专题命中 扩散模型 :diffusion(title,abstract)

Comments 11 pages, code available in https://github.com/jsalvasoler/difusco_ddea

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23450 2025-10-13 stat.ME cs.SI physics.soc-ph 78%

Understanding How Network Geometry Influences Diffusion Processes in Complex Networks: A Focus on Cryptocurrency Blockchains and Critical Infrastructure Networks

S M Mustaquim, Asim K. Dey, Abhijit Mandal

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06485 2025-10-13 cond-mat.mtrl-sci cs.AI cs.LG 78%

WyckoffDiff -- A Generative Diffusion Model for Crystal Symmetry

Filip Ekström Kelvinius, Oskar B. Andersson, Abhijith S. Parackal, Dong Qian, Rickard Armiento, Fredrik Lindsten

机构 * Department of Computer and Information Science (IDA), Linköping University, Sweden(计算机与信息科学系(IDA),_linköping大学,瑞典) Department of Physics, Chemistry and Biology (IFM), Linköping University, Sweden(物理、化学与生物学系(IFM),_linköping大学,瑞典)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted to ICML 2025, official PMLR proceedings can be found at https://proceedings.mlr.press/v267/ekstrom-kelvinius25a.html. Code is available online at https://github.com/httk/wyckoffdiff

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06379 2025-10-13 cs.LG cs.AI stat.ML 78%

Solving Linear-Gaussian Bayesian Inverse Problems with Decoupled Diffusion Sequential Monte Carlo

Filip Ekström Kelvinius, Zheng Zhao, Fredrik Lindsten

机构 * Department of Computer and Information Science (IDA), Linköping University, Sweden(计算机与信息科学系(IDA),_linköping大学,瑞典)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted to ICML 2025, official PMLR proceedings can be found at https://proceedings.mlr.press/v267/ekstrom-kelvinius25b.html. Code available at https://github.com/filipekstrm/ddsmc

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02680 2025-10-13 cs.CV eess.IV 77%

Solving Inverse Problems with FLAIR

Julius Erbach, Dominik Narnhofer, Andreas Dombos, Bernt Schiele, Jan Eric Lenssen, Konrad Schindler

机构 * ETH Zürich(苏黎世联邦理工学院) Max Planck Institute for Informatics(马克斯·普朗克信息研究所)

专题命中 扩散模型 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.07054 2025-10-13 math.NA cs.NA 71%

Second-order diffusion limit for the phonon transport equation-asymptotics and numerics

Anjali Nair, Qin Li, Weiran Sun

专题命中 扩散模型 :diffusion(title)

Journal ref Partial Differential Equations and Applications (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09438 2025-10-13 cs.CV 57%

Mono4DEditor: Text-Driven 4D Scene Editing from Monocular Video via Point-Level Localization of Language-Embedded Gaussians

Jin-Chuan Shi, Chengye Su, Jiajun Wang, Ariel Shamir, Miao Wang

机构 * Zhejiang University(浙江大学) State Key Laboratory of Virtual Reality Technology and Systems(虚拟现实技术与系统国家重点实验室) Beihang University(北京航空航天大学) Reichman University(Reichman大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments 19 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09314 2025-10-13 cs.CV 57%

RadioFlow: Efficient Radio Map Construction Framework with Flow Matching

Haozhe Jia, Wenshuo Chen, Xiucheng Wang, Nan Cheng, Hongbo Zhang, Kuimou Yu, Songning Lai, Nanjian Jia, Bowen Tian, Hongru Xiao, Yutao Yue

机构 * Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) State Key Laboratory of ISN and School of Telecommunications Engineering(信息与通信工程研究所和电信工程学院) Xidian University(西安电子科技大学) Peking University(北京大学) Beijing Technology and Business University(北京科技职业大学) College of Civil Engineering, Tongji University(同济大学土木工程学院) Institute of Deep Perception Technology, JITRI(深度感知技术研究所,JITRI)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09212 2025-10-13 cs.CV 57%

Stable Video Infinity: Infinite-Length Video Generation with Error Recycling

Wuyang Li, Wentao Pan, Po-Chien Luan, Yang Gao, Alexandre Alahi

机构 * EPFL(苏黎世联邦理工学院)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Project Page: https://stable-video-infinity.github.io/homepage/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09078 2025-10-13 cs.GR cs.LG 57%

MCMC: Bridging Rendering, Optimization and Generative AI

Gurprit Singh, Wenzel Jakob

机构 * Max Planck Institute for Informatics(马克斯·普朗克研究所信息学研究所) EPFL(瑞士联邦理工学院)

专题命中 扩散模型 :diffusion(abstract);分类 cs.GR

Comments SIGGRAPH Asia 2024 Courses. arXiv admin note: text overlap with arXiv:2208.11970 by other authors

Journal ref SIGGRAPH Asia 2024 Courses, Article No.: 8, Pages 1 - 27

详情

展开后加载摘要…

URL PDF HTML 收藏