arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-08-12 至 2025-08-12 共收录 78 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 78 篇

2505.22792 2025-08-12 cs.CV 92%

Rhetorical Text-to-Image Generation via Two-layer Diffusion Policy Optimization

Yuxi Zhang, Yueting Li, Xinyu Du, Sibo Wang

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) University of California, Berkeley(加州大学伯克利分校)

专题命中 扩散模型 :image generation(title,abstract);text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07519 2025-08-12 cs.CV 88%

Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing

Joonghyuk Shin, Alchan Hwang, Yujin Kim, Daneul Kim, Jaesik Park

机构 * Seoul National University(首尔国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image editing(title,abstract);分类 cs.CV

Comments ICCV 2025. Project webpage: https://joonghyuk.com/exploring-mmdit-web/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16726 2025-08-12 cs.CV cs.LG 87%

EDiT: Efficient Diffusion Transformers with Linear Compressed Attention

Philipp Becker, Abhinav Mehrotra, Ruchika Chavhan, Malcolm Chadwick, Luca Morreale, Mehdi Noroozi, Alberto Gil Ramos, Sourav Bhattacharya

机构 * Samsung, AI Center Cambridge(三星人工智能中心剑桥)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);text-to-image(abstract);image synthesis(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06134 2025-08-12 cs.CV 87%

X2I: Seamless Integration of Multimodal Understanding into Diffusion Transformer via Attention Distillation

Jian Ma, Qirong Peng, Xu Guo, Chen Chen, Haonan Lu, Zhenyu Yang

机构 * OPPO AI Center(OPPO人工智能中心) Tsinghua University(清华大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);text-to-image(abstract);image editing(abstract)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07747 2025-08-12 cs.CV 83%

Grouped Speculative Decoding for Autoregressive Image Generation

Junhyuk So, Juncheol Shin, Hyunho Kook, Eunhyeok Park

机构 * Department of Computer Science and Engineering, POSTECH(计算机科学与工程系,POSTECH) Graduate School of Artificial Intelligence, POSTECH(人工智能研究生院,POSTECH)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted to the ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07183 2025-08-12 cs.HC cs.AI cs.LG cs.MM 83%

Explainability-in-Action: Enabling Expressive Manipulation and Tacit Understanding by Bending Diffusion Models in ComfyUI

Ahmed M. Abuzuraiq, Philippe Pasquier

机构 * School of Interactive Arts and Technology(交互艺术与技术学院) Simon Fraser University(西蒙弗雷泽大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.MM

Comments In Proceedings of Explainable AI for the Arts Workshop 2025 (XAIxArts 2025) arXiv:2406.14485

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.04832 2025-08-12 cs.LG cs.AI math.OC 82%

Reward-Directed Score-Based Diffusion Models via q-Learning

Xuefeng Gao, Jiale Zha, Xun Yu Zhou

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07413 2025-08-12 cs.CV 81%

CLUE: Leveraging Low-Rank Adaptation to Capture Latent Uncovered Evidence for Image Forgery Localization

Youqi Wang, Shunquan Tan, Rongxuan Peng, Bin Li, Jiwu Huang

专题命中 扩散模型 :text-to-image(abstract);diffusion(abstract);image editing(abstract);image synthesis(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06923 2025-08-12 cs.CV cs.AI 80%

From Reusing to Forecasting: Accelerating Diffusion Models with TaylorSeers

Jiacheng Liu, Chang Zou, Yuanhuiyi Lyu, Junjie Chen, Linfeng Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Shandong University(山东大学) University of Electronic Science and Technology of China(电子科技大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 14 figures; Accepted by ICCV2025; Mainly focus on feature caching for diffusion transformers acceleration

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07700 2025-08-12 cs.CV 79%

Make Your MoVe: Make Your 3D Contents by Adapting Multi-View Diffusion Models to External Editing

Weitao Wang, Haoran Xu, Jun Meng, Haoqian Wang

机构 * Tsinghua University(清华大学) Zhejiang University(浙江大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07682 2025-08-12 eess.IV cs.CV 79%

DiffVC-OSD: One-Step Diffusion-based Perceptual Neural Video Compression Framework

Wenzhuo Ma, Zhenzhong Chen

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07557 2025-08-12 cs.CV 79%

Splat4D: Diffusion-Enhanced 4D Gaussian Splatting for Temporally and Spatially Consistent Content Creation

Minghao Yin, Yukang Cao, Songyou Peng, Kai Han

机构 * The University of Hong Kong(香港大学) Nanyang Technological University(南洋理工大学) Google DeepMind(谷歌DeepMind)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07346 2025-08-12 cs.CV 79%

SODiff: Semantic-Oriented Diffusion Model for JPEG Compression Artifacts Removal

Tingyu Yang, Jue Gong, Jinpei Guo, Wenbo Li, Yong Guo, Yulun Zhang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 7 pages, 5 figures. The code will be available at \url{https://github.com/frakenation/SODiff}

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07162 2025-08-12 cs.CV 79%

CoopDiff: Anticipating 3D Human-object Interactions via Contact-consistent Decoupled Diffusion

Xiaotong Lin, Tianming Liang, Jian-Fang Hu, Kun-Yu Lin, Yulei Kang, Chunwei Tian, Jianhuang Lai, Wei-Shi Zheng

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07146 2025-08-12 cs.CV cs.AI 79%

Intention-Aware Diffusion Model for Pedestrian Trajectory Prediction

Yu Liu, Zhijie Liu, Xiao Ren, You-Fu Li, He Kong

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07006 2025-08-12 eess.IV cs.CV 79%

Spatio-Temporal Conditional Diffusion Models for Forecasting Future Multiple Sclerosis Lesion Masks Conditioned on Treatments

Gian Mario Favero, Ge Ya Luo, Nima Fathi, Justin Szeto, Douglas L. Arnold, Brennan Nichyporuk, Chris Pal, Tal Arbel

机构 * McGill University(麦吉尔大学) Mila – Quebec AI Institute(魁北克人工智能研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to MICCAI 2025 (LMID Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03295 2025-08-12 cs.CV 79%

CPKD: Clinical Prior Knowledge-Constrained Diffusion Models for Surgical Phase Recognition in Endoscopic Submucosal Dissection

Xiangning Zhang, Jinnan Chen, Qingwei Zhang, Yaqi Wang, Chengfeng Zhou, Xiaobo Li, Dahong Qian

机构 * School of Biomedical Engineering(生物医学工程学院) Division of Gastroenterology and Hepatology, Shanghai Institute of Digestive Disease, NHC Key Laboratory of Digestive Diseases, Renji Hospital(消化内科与肝病科、上海消化疾病研究所、国家消化疾病临床医学研究中心、仁济医院) College of Media Engineering(媒体工程学院) Aier Institute of Digital Ophthalmology and Visual Science(数字眼科与视觉科学研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19914 2025-08-12 cs.CV 79%

Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models

Sangwon Baik, Hyeonwoo Kim, Hanbyul Joo

机构 * Seoul National University(首尔国立大学) RLWRLD

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://tlb-miss.github.io/oor/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08333 2025-08-12 cs.CV 79%

DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models

Hyeonwoo Kim, Sangwon Baik, Hanbyul Joo

机构 * Seoul National University(首尔国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://snuvclab.github.io/david/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01440 2025-08-12 cs.CV 79%

BadPatch: Diffusion-Based Generation of Physical Adversarial Patches

Zhixiang Wang, Xingjun Ma, Yu-Gang Jiang

机构 * Shanghai Key Lab of Intell. Info. Processing, School of CS, Fudan University(上海智能信息处理关键实验室,计算机科学学院,复旦大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Code available at: https://github.com/Wwangb/BadPatch

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12777 2025-08-12 cs.CV cs.CL cs.CR cs.LG 79%

Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts

Hongcheng Gao, Tianyu Pang, Chao Du, Taihang Hu, Zhijie Deng, Min Lin

机构 * Sea AI Lab, Singapore(新加坡Sea AI实验室) University of Chinese Academy of Sciences(中国科学院大学) Shanghai Jiao Tong University(上海交通大学) Nankai University(南开大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07128 2025-08-12 cs.CV cs.AI 79%

Perceptual Evaluation of GANs and Diffusion Models for Generating X-rays

Gregory Schuit, Denis Parra, Cecilia Besa

机构 * Pontificia Universidad Católica de Chile(智利天主教大学) iHealth - Instituto Milenio en Ingeniería e Inteligencia Artificial para la Salud(iHealth - 毫米级工程与人工智能健康研究所) CENIA – Centro Nacional de Inteligencia Artificial(国家人工智能中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to the Workshop on Human-AI Collaboration at MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02056 2025-08-12 cs.CV 79%

StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion

Haoxin Yang, Weihong Chen, Xuemiao Xu, Cheng Xu, Peng Xiao, Cuifeng Sun, Shaoyu Huang, Shengfeng He

机构 * School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院) Centre for Smart Health, The Hong Kong Polytechnic University(香港理工大学智能健康中心) Cloud Computing Center, Chinese Academy of Sciences(中国科学院云计算中心) Guangzhou Yichuang Information Technology Co., Ltd.(广州亿创信息技术有限公司) School of Computing and Information Systems, Singapore Management University(新加坡管理大学计算机与信息系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14769 2025-08-12 cs.CV cs.RO 79%

CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion

Jiahua Ma, Yiran Qin, Yixiong Li, Xuanqi Liao, Yulan Guo, Ruimao Zhang

机构 * Sun Yat-sen University(中山大学) CUHK(SZ)(香港中文大学(深圳)) Oxford(牛津大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05855 2025-08-12 cs.RO cs.CV 79%

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Junjie Wen, Yichen Zhu, Jinming Li, Zhibin Tang, Chaomin Shen, Feifei Feng

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments The webpage is at https://dex-vla.github.io/. DexVLA is accepted by CoRL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07926 2025-08-12 cs.LG 78%

Score Augmentation for Diffusion Models

Liang Hou, Yuan Gao, Boyuan Jiang, Xin Tao, Qi Yan, Renjie Liao, Pengfei Wan, Di Zhang, Kun Gai

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07815 2025-08-12 eess.IV 78%

Deep Learning-Based Desikan-Killiany Parcellation of the Brain Using Diffusion MRI

Yousef Sadegheih, Dorit Merhof

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07420 2025-08-12 math.NA cs.NA math.AP 78%

Robust, fast, and adaptive splitting schemes for nonlinear doubly-degenerate diffusion equations

Ayesha Javed, Koondanibha Mitra, Iuliu Sorin Pop

专题命中 扩散模型 :diffusion(title,abstract)

Comments 39 pages, 20 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07243 2025-08-12 cs.LG cs.AI 78%

Causal Negative Sampling via Diffusion Model for Out-of-Distribution Recommendation

Chu Zhao, Eneng Yang, Yizhou Dang, Jianzhe Zhao, Guibing Guo, Xingwei Wang

机构 * Northeastern University(东北大学)

专题命中 扩散模型 :diffusion(title,abstract)

Comments 14 pages, 6 figures, Under-review

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16288 2025-08-12 math.PR math.OC 78%

Pontryagin Maximum Principle for McKean-Vlasov Stochastic Reaction-Diffusion Equations

Johan Benedikt Spille, Wilhelm Stannat

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏