arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-11-18 至 2025-11-18 共收录 87 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 87 篇

2408.00998 2025-11-18 cs.CV cs.AI 87%

FBSDiff: Plug-and-Play Frequency Band Substitution of Diffusion Features for Highly Controllable Text-Driven Image Translation

Xiang Gao, Jiaying Liu

机构 * Wangxuan Institute of Computer Technology, Peking University(王轩计算机技术研究所,北京大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);text-to-image(abstract);image synthesis(abstract)

Comments Accepted conference paper of ACM MM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05772 2025-11-18 cs.CV 83%

MAISI-v2: Accelerated 3D High-Resolution Medical Image Synthesis with Rectified Flow and Region-specific Contrastive Loss

Can Zhao, Pengfei Guo, Dong Yang, Yucheng Tang, Yufan He, Benjamin Simon, Mason Belue, Stephanie Harmon, Baris Turkbey, Daguang Xu

专题命中 扩散模型 :image synthesis(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06157 2025-11-18 cs.CV 83%

3D-free meets 3D priors: Novel View Synthesis from a Single Image with Pretrained Diffusion Guidance

Taewon Kang, Divya Kothandaraman, Dinesh Manocha, Ming C. Lin

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to The 40th Annual AAAI Conference on Artificial Intelligence (AAAI-26), AAAI 2026 Workshop on AI for Environmental Science (AI4ES). Due to arXiv's 1,920-character limit, the abstract here is shortened. Please refer to the paper (View PDF) to read the full abstract. 14 pages, 13 figures, v5: AAAI-26 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03457 2025-11-18 cs.GR cs.CV cs.SD eess.AS 81%

READ: Real-time and Efficient Asynchronous Diffusion for Audio-driven Talking Head Generation

Haotian Wang, Yuzhe Weng, Jun Du, Haoran Xu, Xiaoyan Wu, Shan He, Bing Yin, Cong Liu, Jianqing Gao, Qingfeng Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project page: https://readportrait.github.io/READ/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12280 2025-11-18 cs.CV cs.CL cs.LG 80%

D$^{3}$ToM: Decider-Guided Dynamic Token Merging for Accelerating Diffusion MLLMs

Shuochen Chang, Xiaofeng Zhang, Qingyang Liu, Li Niu

机构 * Project leader(项目负责人)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by AAAI Conference on Artificial Intelligence (AAAI) 2026. Code available at https://github.com/bcmi/D3ToM-Diffusion-MLLM

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13121 2025-11-18 cs.CV 79%

CloseUpShot: Close-up Novel View Synthesis from Sparse-views via Point-conditioned Diffusion Model

Yuqi Zhang, Guanying Chen, Jiaxing Chen, Chuanyu Fu, Chuan Huang, Shuguang Cui

机构 * Shenzhen Future Network of Intelligence Institute (FNii-Shenzhen)(深圳未来网络智能研究院(FNii-深圳)) Chinese University of Hong Kong at Shenzhen (CUHKSZ)(香港中文大学(深圳)) Sun Yat-sen University(中山大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Link: https://zyqz97.github.io/CloseUpShot/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12757 2025-11-18 cs.CV cs.AI 79%

Which Way from B to A: The role of embedding geometry in image interpolation for Stable Diffusion

Nicholas Karris, Luke Durell, Javier Flores, Tegan Emerson

机构 * Department of Mathematics University of California, San Diego(数学系 加州大学圣地亚哥分校) National Security Directorate Pacific Northwest National Laboratory(国家安全局 西部国家实验室) Earth and Biological Sciences Directorate Pacific Northwest National Laboratory(地球与生命科学局 西部国家实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12428 2025-11-18 cs.CV 79%

RedVTP: Training-Free Acceleration of Diffusion Vision-Language Models Inference via Masked Token-Guided Visual Token Pruning

Jingqi Xu, Jingxi Lu, Chenghao Li, Sreetama Sarkar, Souvik Kundu, Peter A. Beerel

机构 * University of Southern California(南加州大学) Intel Labs(英特尔实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12304 2025-11-18 cs.CV 79%

LiDAR-GS++:Improving LiDAR Gaussian Reconstruction via Diffusion Priors

Qifeng Chen, Jiarun Liu, Rengan Xie, Tao Tang, Sicong Du, Yiru Zhao, Yuchi Huo, Sheng Yang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14741 2025-11-18 cs.CV cs.AI 79%

DEXTER: Diffusion-Guided EXplanations with TExtual Reasoning for Vision Models

Simone Carnemolla, Matteo Pennisi, Sarinda Samarasinghe, Giovanni Bellitto, Simone Palazzo, Daniela Giordano, Mubarak Shah, Concetto Spampinato

机构 * University of Catania(卡塔尼亚大学) University of Central Florida(中央佛罗里达大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to NeurIPS 2025 (spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11334 2025-11-18 cs.CV 79%

GANDiff FR: Hybrid GAN Diffusion Synthesis for Causal Bias Attribution in Face Recognition

Md Asgor Hossain Reaj, Rajan Das Gupta, Md Yeasin Rahat, Nafiz Fahad, Md Jawadul Hasan, Tze Hui Liew

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This is the preprint version of the manuscript. It is currently being prepared for submission to an academic conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00652 2025-11-18 cs.CV cs.CR 79%

Video Signature: Implicit Watermarking for Video Diffusion Models

Yu Huang, Junhao Chen, Shuliang Liu, Hanqian Li, Jungang Li, Qi Zheng, Aiwei Liu, Yi R. Fung, Xuming Hu

机构 * AI Thrust, Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)人工智能 thrust) Hong Kong University of Science and Technology(香港科学与技术大学) School of Software, BNRist, Tsinghua University(清华大学软件学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 17 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11777 2025-11-18 cs.CV 79%

Self-NPO: Data-Free Diffusion Model Enhancement via Truncated Diffusion Fine-Tuning

Fu-Yun Wang, Keqiang Sun, Yao Teng, Xihui Liu, Jiale Yuan, Jiaming Song, Hongsheng Li

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21385 2025-11-18 cs.CV 79%

Physics-Guided Image Dehazing Diffusion

Shijun Zhou, Xing Xie, Baojie Fan, Jiandong Tian

机构 * State Key Laboratory of Robotics and Intelligent Systems(机器人与智能系统国家重点实验室) Shenyang Institute of Automation, Chinese Academy of Sciences(中国科学院沈阳自动化研究所) University of Chinese Academy of Sciences(中国科学院大学) Department of Automation, Nanjing University of Posts and Telecommunications(南京邮电大学自动化系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17347 2025-11-18 cs.CV 79%

Dereflection Any Image with Diffusion Priors and Diversified Data

Jichen Hu, Chen Yang, Zanwei Zhou, Jiemin Fang, Xiaokang Yang, Qi Tian, Wei Shen

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12015 2025-11-18 cs.CV 79%

QDM: Quadtree-Based Region-Adaptive Sparse Diffusion Models for Efficient Image Super-Resolution

Donglin Yang, Paul Vicol, Xiaojuan Qi, Renjie Liao, Xiaofan Zhang

机构 * The University of Hong Kong(香港大学) Google DeepMind(谷歌DeepMind) The University of British Columbia(不列颠哥伦比亚大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.00620 2025-11-18 cs.CV 79%

Lane Graph Extraction from Aerial Imagery via Lane Segmentation Refinement with Diffusion Models

Antonio Ruiz, Andrew Melnik, Nicolo Savioli, Dong Wang, Yanfeng Zhang, Helge Ritter

机构 * Riemann Lab , Huawei(里曼实验室,华为) Center for Cognitive Interaction Technology (CITEC), Faculty of Technology, Bielefeld University(认知交互技术中心(CITEC),技术学院,比勒菲尔德大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Remote Sensing, 17(16), 2845

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12099 2025-11-18 cs.CV 79%

Adaptive Begin-of-Video Tokens for Autoregressive Video Diffusion Models

Tianle Cheng, Zeyan Zhang, Kaifeng Gao, Jun Xiao

机构 * Zhejiang University(浙江大学) Manycore Tech Inc.(Manycore科技公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12072 2025-11-18 cs.MM cs.AI cs.SD 79%

ProAV-DiT: A Projected Latent Diffusion Transformer for Efficient Synchronized Audio-Video Generation

Jiahui Sun, Weining Wang, Mingzhen Sun, Yirong Yang, Xinxin Zhu, Jing Liu

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Beihang University(北航大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12056 2025-11-18 cs.CV cs.AI cs.DC 79%

PipeDiT: Accelerating Diffusion Transformers in Video Generation with Task Pipelining and Model Decoupling

Sijie Wang, Qiang Wang, Shaohuai Shi

机构 * Sijie Wang, Qiang Wang, Shaohuai Shi(作者)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11944 2025-11-18 cs.CV 79%

From Events to Clarity: The Event-Guided Diffusion Framework for Dehazing

Ling Wang, Yunfan Lu, Wenzong Ma, Huizai Yao, Pengteng Li, Hui Xiong

机构 * The Hong Kong University of Science and Technology (GuangZhou)(香港科技大学(广州))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 11 pages, 8 figures. Completed in April 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08291 2025-11-18 cs.CV 79%

SynWeather: Weather Observation Data Synthesis across Multiple Regions and Variables via a General Diffusion Transformer

Kaiyi Xu, Junchao Gong, Zhiwang Zhou, Zhangrui Li, Yuandong Pu, Yihao Liu, Ben Fei, Fenghua Ling, Wenlong Zhang, Lei Bai

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by AAAI-26 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15217 2025-11-18 cs.SD cs.AI cs.LG cs.MM 79%

DRAGON: Distributional Rewards Optimize Diffusion Generative Models

Yatong Bai, Jonah Casebeer, Somayeh Sojoudi, Nicholas J. Bryan

机构 * University of California, Berkeley(加州大学伯克利分校) Adobe Research(Adobe研究)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.MM

Comments Accepted to TMLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17459 2025-11-18 cs.CE 78%

Sparse Diffusion Autoencoder for Test-time Adapting Prediction of Complex Systems

Jingwen Cheng, Ruikun Li, Huandong Wang, Yong Li

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13186 2025-11-18 cs.LG cs.SY eess.SY 78%

DiffFP: Learning Behaviors from Scratch via Diffusion-based Fictitious Play

Akash Karthikeyan, Yash Vardhan Pant

机构 * Department of Electrical and Computer Engineering, University of Waterloo(滑铁卢大学电气与计算机工程系)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Initial results presented at the IJCAI 2025 Workshop on User-Aligned Assessment of Adaptive AI Systems. Project page: https://aku02.github.io/projects/difffp/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13137 2025-11-18 cs.AI 78%

Conditional Diffusion Model for Multi-Agent Dynamic Task Decomposition

Yanda Zhu, Yuanyang Zhu, Daoyi Dong, Caihua Chen, Chunlin Chen

专题命中 扩散模型 :diffusion(title,abstract)

Comments AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12912 2025-11-18 cs.RO 78%

DiffuDepGrasp: Diffusion-based Depth Noise Modeling Empowers Sim2Real Robotic Grasping

Yingting Zhou, Wenbo Cui, Weiheng Liu, Guixing Chen, Haoran Li, Dongbin Zhao

机构 * The State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Beijing Zhiwangweilai Technology Co., Ltd.(北京智王未来科技有限公司)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12843 2025-11-18 q-bio.BM 78%

Treatment of phenol wastewater by electro-Fenton oxidative degradation based on efficient iron-based-gas diffusion-photocatalysis

Zhang Junye, Zheng Hongyu, Cheng Jingran, Zhang Shengli

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12742 2025-11-18 cs.LG 78%

Stabilizing Self-Consuming Diffusion Models with Latent Space Filtering

Zhongteng Cai, Yaxuan Wang, Yang Liu, Xueru Zhang

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted by AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12221 2025-11-18 quant-ph 78%

Channel-Constrained Markovian Quantum Diffusion Model from Open System Perspective

Qin-Sheng Zhu, Geng Chen, Lian-Hui Yu, Xiaodong Xing, Xiao-Yu Li

专题命中 扩散模型 :diffusion(title,abstract)

Comments 38 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏