arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-01 至 2025-10-01 共收录 82 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 6 篇

2501.15562 2025-10-01 cs.CV cs.AI 88%

CE-SDWV: Effective and Efficient Concept Erasure for Text-to-Image Diffusion Models via a Semantic-Driven Word Vocabulary

Jiahang Tu, Qian Feng, Jiahua Dong, Hanbin Zhao, Chao Zhang, Nicu Sebe, Hui Qian

机构 * College of Computer Science and Technology, Zhejiang University, China(浙江大学计算机科学与技术学院) Mohamed bin Zayed University of Artificial Intelligence, Abu Dhabi(阿布扎赫德莫罕默德·本·扎耶德人工智能大学) Department of Information Engineering and Computer Science, University of Trento, Italy(意大利特伦托大学信息工程与计算机科学系)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

Comments 25 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25771 2025-10-01 cs.CV cs.AI 88%

Free Lunch Alignment of Text-to-Image Diffusion Models without Preference Image Pairs

Jia Jun Cheng Xian, Muchen Li, Haotian Yang, Xin Tao, Pengfei Wan, Leonid Sigal, Renjie Liao

机构 * University of British Columbia(不列颠哥伦比亚大学) Vector Institute for AI(人工智能向量研究所) Canada CIFAR AI Chair(加拿大CIFAR人工智能主席) Kling Team, Kuaishou Technology(快手科技 Kling 团队) NSERC CRC Chair(加拿大NSERC CRC主席)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04946 2025-10-01 cs.CV cs.CL 88%

Taming the Tri-Space Tension: ARC-Guided Hallucination Modeling and Control for Text-to-Image Generation

Jianjiang Yang, Ziyan Huang, Yanshu li, Da Peng, Huaiyuan Yao

机构 * University of Bristol(布里斯托大学) South China University of Technology(华南理工大学) Brown University(布朗大学) Xi’an Jiaotong University(西安交通大学) Arizona State University(亚利桑那州立大学)

专题命中 文生图 :text-to-image(title,abstract);image generation(title);diffusion(abstract);分类 cs.CV

Comments 9 pages,6 figures,7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26641 2025-10-01 cs.CV 85%

Query-Kontext: An Unified Multimodal Model for Image Generation and Editing

Yuxin Song, Wenkai Dong, Shizun Wang, Qi Zhang, Song Xue, Tao Yuan, Hu Yang, Haocheng Feng, Hang Zhou, Xinyan Xiao, Jingdong Wang

机构 * Baidu VIS(百度视觉) National University of Singapore(新加坡国立大学)

专题命中 文生图 :image generation(title,abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments 23 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23398 2025-10-01 cs.CV cs.CY 83%

A Large Scale Analysis of Gender Biases in Text-to-Image Generative Models

Leander Girrbach, Stephan Alaniz, Genevieve Smith, Zeynep Akata

机构 * Technical University of Munich, Munich Center for Machine Learning, MDSI(慕尼黑技术大学、慕尼黑机器学习中心、MDSI) LTCI, Télécom Paris, Institut Polytechnique de Paris(LTCI、巴黎电信学院、巴黎理工学院) Helmholtz Munich(海德堡-慕尼黑研究中心) University of California, Berkeley(加州大学伯克利分校)

专题命中 文生图 :text-to-image(title,abstract);image generation(abstract);分类 cs.CV

Comments 40 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26055 2025-10-01 cs.GR cs.CV cs.LG 62%

GaussEdit: Adaptive 3D Scene Editing with Text and Image Prompts

Zhenyu Shu, Junlong Yu, Kai Chao, Shiqing Xin, Ligang Liu

机构 * School of Computer and Data Engineering, NingboTech University(计算机与数据工程学院,宁波科技学院) Ningbo Institute, Zhejiang University(浙江大学宁波学院) School of Software Technology, Zhejiang University(软件技术学院,浙江大学) School of Big Data and Artificial Intelligence Management, Xi’an Jiaotong University(大数据与人工智能管理学院,西安交通大学) School of Computer Science and Technology, ShanDong University(计算机科学与技术学院,山东大学) Graphics & Geometric Computing Laboratory, School of Mathematical Sciences, University of Science and Technology of China(图形与几何计算实验室,数学科学学院,中国科学技术大学)

专题命中 文生图 :image synthesis(abstract);分类 cs.CV、cs.GR

Journal ref IEEE Transactions on Visualization and Computer Graphics. 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 扩散模型 63 篇

2505.06995 2025-10-01 cs.CV 88%

KDC-Diff: A Latent-Aware Diffusion Model with Knowledge Retention for Memory-Efficient Image Generation

Md. Naimur Asif Borno, Md Sakib Hossain Shovon, Asmaa Soliman Al-Moisheer, Mohammad Ali Moni

机构 * The University of Queensland(昆士兰大学) Rajshahi University of Engineering & Technology(拉贾沙希工程与技术大学) Imam Mohammad Ibn Saud Islamic University (IMSIU)(伊玛姆穆罕默德·本·沙特伊斯兰大学) Charles Sturt University(查尔斯·斯特劳特大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(title);text-to-image(abstract);分类 cs.CV

Comments Currently Under Review at IEEE open journal for Computer Society

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26538 2025-10-01 cs.AI 88%

HilbertA: Hilbert Attention for Image Generation with Diffusion Models

Shaoyi Zheng, Wenbo Lu, Yuxuan Xia, Haomin Liu, Shengjie Wang

机构 * Department of Computer Science, New York University(纽约大学计算机科学系)

专题命中 扩散模型 :image generation(title,abstract);diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26644 2025-10-01 cs.CV cs.AI cs.LG 83%

Stitch: Training-Free Position Control in Multimodal Diffusion Transformers

Jessica Bader, Mateusz Pach, Maria A. Bravo, Serge Belongie, Zeynep Akata

机构 * Technical University of Munich(慕尼黑技术大学) Helmholtz Munich(海德堡-慕尼黑研究所) Munich Center for Machine Learning(慕尼黑机器学习中心) University of Copenhagen(哥本哈根大学)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26436 2025-10-01 cs.CV 83%

Post-Training Quantization via Residual Truncation and Zero Suppression for Diffusion Models

Donghoon Kim, Dongyoung Lee, Ik Joon Chang, Sung-Ho Bae

机构 * Department of Artificial Intelligence(人工智能系) Kyung Hee University(庆尚大学) Department of Electrical Engineering(电气工程系) Department of Computer Science(计算机科学系)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25705 2025-10-01 cs.CV 83%

How Diffusion Models Memorize

Juyeop Kim, Songkuk Kim, Jong-Seok Lee

机构 * Yonsei University(延世大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13915 2025-10-01 cs.CV 83%

Binary Diffusion Probabilistic Model

Vitaliy Kinakh, Slava Voloshynovskiy

机构 * Department of Computer Science University of Geneva(计算机科学系日内瓦大学)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26231 2025-10-01 cs.CV 79%

IMG: Calibrating Diffusion Models via Implicit Multimodal Guidance

Jiayi Guo, Chuanhao Yan, Xingqian Xu, Yulin Wang, Kai Wang, Gao Huang, Humphrey Shi

机构 * Tsinghua University(清华大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26025 2025-10-01 cs.CV 79%

PatchVSR: Breaking Video Diffusion Resolution Limits with Patch-wise Video Super-Resolution

Shian Du, Menghan Xia, Chang Liu, Xintao Wang, Jing Wang, Pengfei Wan, Di Zhang, Xiangyang Ji

机构 * Tsinghua University(清华大学) Kling Team, Kuaishou Technology(快手科技 Kling 团队) Beijing Institute of Technology(北京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25739 2025-10-01 cs.CV 79%

LieHMR: Autoregressive Human Mesh Recovery with $SO(3)$ Diffusion

Donghwan Kim, Tae-Kyun Kim

机构 * School of Computing, KAIST(计算机学院,韩国科学技术院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 17 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25681 2025-10-01 cs.RO cs.CV 79%

dVLA: Diffusion Vision-Language-Action Model with Multimodal Chain-of-Thought

Junjie Wen, Minjie Zhu, Jiaming Liu, Zhiyuan Liu, Yicun Yang, Linfeng Zhang, Shanghang Zhang, Yichen Zhu, Yi Xu

机构 * Midea Group(美的集团) Peking University(北京大学) Shanghai Jiaotong University(上海交通大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments technique report

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25304 2025-10-01 cs.CV 79%

LUMA: Low-Dimension Unified Motion Alignment with Dual-Path Anchoring for Text-to-Motion Diffusion Model

Haozhe Jia, Wenshuo Chen, Yuqi Lin, Yang Yang, Lei Wang, Mang Ning, Bowen Tian, Songning Lai, Nanqian Jia, Yifan Chen, Yutao Yue

机构 * HKUST-GZ(香港科技大学-广州校区) Shandong University(山东大学) UESTC(电子科技大学) Griffith University(格里菲斯大学) Data61/CSIRO Utrecht University(乌得勒支大学) Beijing University of Posts and Telecommunications(北京邮电大学) Peking University(北京大学) Institute of Deep Perception Technology, JITRI(JITRI深度感知技术研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25280 2025-10-01 eess.IV cs.CV 79%

Anatomy-DT: A Cross-Diffusion Digital Twin for Anatomical Evolution

Moinak Bhattacharya, Gagandeep Singh, Prateek Prasanna

机构 * Stony Brook University(斯通布罗克大学) Columbia University(哥伦比亚大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13343 2025-10-01 cs.CV eess.IV 79%

Taming Diffusion Transformer for Efficient Mobile Video Generation in Seconds

Yushu Wu, Yanyu Li, Anil Kag, Ivan Skorokhodov, Willi Menapace, Ke Ma, Arpit Sahni, Ju Hu, Aliaksandr Siarohin, Dhritiman Sagar, Yanzhi Wang, Sergey Tulyakov

机构 * Snap Inc.(Snap公司) Northeastern University(东北大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 21 pages, 9 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03054 2025-10-01 cs.CV cs.AI 79%

LATTE: Latent Trajectory Embedding for Diffusion-Generated Image Detection

Ana Vasilcoiu, Ivona Najdenkoska, Zeno Geradts, Marcel Worring

机构 * University of Amsterdam(阿姆斯特丹大学) Netherlands Forensic Institute (NFI)(荷兰法医学研究所)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08643 2025-10-01 stat.ML cs.AI cs.CV cs.LG 79%

Rethinking Diffusion Model in High Dimension

Zhenxin Zheng, Zhenjie Zheng

机构 * MTLab, Meitu Inc.(美图公司MT实验室) Civil Engineering, University of Hong Kong(香港大学土木工程系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.11797 2025-10-01 math.AP 78%

Equilibrium and Non-Equilibrium diffusion approximation for the radiative transfer equation

Elena Demattè, Juan J. L. Velázquez

专题命中 扩散模型 :diffusion(title,abstract)

Comments 1 Figure, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26328 2025-10-01 cs.CL 78%

Fast-dLLM v2: Efficient Block-Diffusion LLM

Chengyue Wu, Hao Zhang, Shuchen Xue, Shizhe Diao, Yonggan Fu, Zhijian Liu, Pavlo Molchanov, Ping Luo, Song Han, Enze Xie

机构 * The University of Hong Kong(香港大学) NVIDIA(英伟达) MIT(麻省理工学院)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26095 2025-10-01 cond-mat.mtrl-sci cond-mat.supr-con 78%

The diffusion-driven orthorhombic to tetragonal transition in YBa$_2$Cu$_3$O$_7$ derived with a machine learning interatomic potential

Davide Gambino, Niccolò Di Eugenio, Jesper Byggmästar, Johan Klarbring, Daniele Torsello, Flyura Djurabekova, Francesco Laviano

专题命中 扩散模型 :diffusion(title,abstract)

Comments 30 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26054 2025-10-01 math.AP 78%

Initial traces and solvability of the fast diffusion equation with power-type nonlinearity

Kazuhiro Ishige, Nobuhito Miyake

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25978 2025-10-01 math.AP 78%

Weak-strong uniqueness for general cross-diffusion systems with volume filling

Maria Heitzinger, Ansgar Jüngel

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25604 2025-10-01 cs.CL cs.LG 78%

RFG: Test-Time Scaling for Diffusion Large Language Model Reasoning with Reward-Free Guidance

Tianlang Chen, Minkai Xu, Jure Leskovec, Stefano Ermon

机构 * Stanford University(斯坦福大学)

专题命中 扩散模型 :diffusion(title,abstract)

Comments 27 pages, 3 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25401 2025-10-01 cs.LG cs.AI cs.PF 78%

FlashOmni: A Unified Sparse Attention Engine for Diffusion Transformers

Liang Qiao, Yue Dai, Yeqi Huang, Hongyu Kan, Jun Shi, Hong An

机构 * University of Science and Technology of China(中国科学技术大学) University of Edinburgh(爱丁堡大学) University of Virginia(弗吉尼亚大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25375 2025-10-01 eess.SY cs.SY 78%

Safe and Stable Control via Lyapunov-Guided Diffusion Models

Xiaoyuan Cheng, Xiaohang Tang, Yiming Yang

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25365 2025-10-01 cond-mat.stat-mech 78%

Diffusion with doubly stochastic resetting

Maxence Arutkin, Shlomi Reuveni

专题命中 扩散模型 :diffusion(title,abstract)

Comments 13 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏