arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-11-11 至 2025-11-11 共收录 125 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 9 篇

2511.05616 2025-11-11 cs.CV cs.AI 92%

Personalized Image Editing in Text-to-Image Diffusion Models via Collaborative Direct Preference Optimization

Connor Dunlop, Matthew Zheng, Kavana Venkatesh, Pinar Yanardag

机构 * Virginia Tech(弗吉尼亚理工大学)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image editing(title,abstract);分类 cs.CV

Comments Published at NeurIPS'25 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19291 2025-11-11 cs.CV cs.AI 90%

TextDiffuser-RL: Efficient and Robust Text Layout Optimization for High-Fidelity Text-to-Image Synthesis

Kazi Mahathir Rahman, Showrin Rahman, Sharmin Sultana Srishty

机构 * BRAC University(布拉克大学)

专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract);image generation(abstract);diffusion(abstract)

Comments 19 pages, 36 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21227 2025-11-11 cs.CV cs.CL 86%

Evaluating the Evaluators: Metrics for Compositional Text-to-Image Generation

Seyed Amir Kasaei, Ali Aghayari, Arash Marioriyad, Niki Sepasian, MohammadAmin Fazli, Mahdieh Soleymani Baghshah, Mohammad Hossein Rohban

机构 * Department of Computer Engineering, Sharif University of Technology(谢里夫理工大学计算机工程系)

专题命中 文生图 :image generation(title,abstract);text-to-image(title);分类 cs.CV

Comments Accepted at GenProCC NeurIPS 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06724 2025-11-11 cs.CV cs.DC 83%

Argus: Quality-Aware High-Throughput Text-to-Image Inference Serving System

Shubham Agarwal, Subrata Mitra, Saud Iqbal

专题命中 文生图 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted at Middleware 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06876 2025-11-11 cs.CV 79%

Generating an Image From 1,000 Words: Enhancing Text-to-Image With Structured Captions

Eyal Gutflaish, Eliran Kachlon, Hezi Zisman, Tal Hacham, Nimrod Sarid, Alexander Visheratin, Saar Huberman, Gal Davidi, Guy Bukchin, Kfir Goldberg, Ron Mokady

机构 * BRIA AI(BRIA人工智能)

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21257 2025-11-11 cs.CV cs.CL 79%

Hallucination as an Upper Bound: A New Perspective on Text-to-Image Evaluation

Seyed Amir Kasaei, Mohammad Hossein Rohban

机构 * Department of Computer Engineering, Sharif University of Technology(谢里夫理工大学计算机工程系)

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

Comments Accepted at GenProCC NeurIPS 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06284 2025-11-11 cs.CV cs.CL cs.MM 73%

Enhancing Multimodal Misinformation Detection by Replaying the Whole Story from Image Modality Perspective

Bing Wang, Ximing Li, Yanjun Wang, Changchun Li, Lin Yuanbo Wu, Buyu Wang, Shengsheng Wang

专题命中 文生图 :image generation(abstract);text-to-image(abstract);分类 cs.CV、cs.MM

Comments Accepted by AAAI 2026. 13 pages, 6 figures. Code: https://github.com/wangbing1416/RETSIMD

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05573 2025-11-11 cs.CV cs.AI 70%

Video Text Preservation with Synthetic Text-Rich Videos

Ziyang Liu, Kevin Valencia, Justin Cui

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01894 2025-11-11 cs.CV cs.AI cs.HC 70%

LIVS: A Pluralistic Alignment Dataset for Inclusive Public Spaces

Rashid Mushkani, Shravan Nayak, Hugo Berard, Allison Cohen, Shin Koseki, Hadrien Bertrand

机构 * Mila–Quebec AI Institute(魁北克AI研究所)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 图像编辑 2 篇

2508.07607 2025-11-11 cs.CV 85%

X2Edit: Revisiting Arbitrary-Instruction Image Editing through Self-Constructed Data and Task-Aware Representation Learning

Jian Ma, Xujie Zhu, Zihao Pan, Qirong Peng, Xu Guo, Chen Chen, Haonan Lu

机构 * OPPO AI Center(OPPO人工智能中心)

专题命中 图像编辑 :image editing(title,abstract);image generation(abstract);diffusion(abstract);分类 cs.CV

Comments Accepted to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06365 2025-11-11 cs.CV 57%

V-Shuffle: Zero-Shot Style Transfer via Value Shuffle

Haojun Tang, Qiwei Lin, Tongda Xu, Lida Huang, Yan Wang

机构 * Tsinghua University(清华大学) Dalian University of Technology(大连理工大学) Beijing Institute of Radio Measurement(北京无线电测量研究所)

专题命中 图像编辑 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 扩散模型 89 篇

2511.05598 2025-11-11 cs.CR eess.IV 89%

Diffusion-Based Image Editing: An Unforeseen Adversary to Robust Invisible Watermarks

Wenkai Fu, Finn Carter, Yue Wang, Emily Davis, Bo Zhang

专题命中 扩散模型 :diffusion(title,abstract);image editing(title,abstract);image generation(abstract)

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07362 2025-11-11 cs.CV cs.AI cs.LG 85%

Inference-Time Scaling of Diffusion Models for Infrared Data Generation

Kai A. Horstmann, Maxim Clouser, Kia Khezeli

机构 * Cornell University(康奈尔大学) YRIKKA, Inc.(YRIKKA公司)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);text-to-image(abstract);分类 cs.CV

Comments Peer-reviewed workshop paper

Journal ref 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: Learning to Sense

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10037 2025-11-11 cs.CV 85%

Role Bias in Diffusion Models: Diagnosing and Mitigating through Intermediate Decomposition

Sina Malakouti, Adriana Kovashka

机构 * University of Pittsburgh(匹兹堡大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24434 2025-11-11 cs.LG cs.CV 83%

Graph Flow Matching: Enhancing Image Generation with Neighbor-Aware Flow Fields

Md Shahriar Rahim Siddiqui, Moshe Eliasof, Eldad Haber

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

Comments The 40th Annual AAAI Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18877 2025-11-11 cs.CV cs.CR cs.LG 83%

Mitigating Sexual Content Generation via Embedding Distortion in Text-conditioned Diffusion Models

Jaesin Ahn, Heechul Jung

机构 * Department of Artificial Intelligence(人工智能系) Kyungpook National University(全北国立大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments NeurIPS 2025 accepted. Official code: https://github.com/amoeba04/des

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09935 2025-11-11 eess.IV cs.CV physics.med-ph 83%

Physics-informed DeepCT: Sinogram Wavelet Decomposition Meets Masked Diffusion

Zekun Zhou, Tan Liu, Bing Yu, Yanru Gong, Liu Shi, Qiegen Liu

机构 * School of Mathematics and Computer Sciences, Nanchang University, Nanchang, China(南昌大学数学与计算机科学学院) School of Information Engineering, Nanchang University, Nanchang, China(南昌大学信息工程学院) Key Laboratory of Advanced Medical Imaging and Intelligent Computing of Guizhou Province, China(贵州省先进医学影像与智能计算重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.12727 2025-11-11 cs.LG cs.AI math.ST stat.ML stat.TH 82%

Diffusion Posterior Sampling is Computationally Intractable

Shivam Gupta, Ajil Jalal, Aditya Parulekar, Eric Price, Zhiyang Xun

机构 * UT Austin(德克萨斯大学) UC Berkeley(加州大学伯克利分校)

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10637 2025-11-11 cs.GR cs.CV 81%

Distilling Diversity and Control in Diffusion Models

Rohit Gandikota, David Bau

机构 * Northeastern University(东北大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Project Page: https://distillation.baulab.info/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07103 2025-11-11 cs.CV cs.AI 79%

GEWDiff: Geometric Enhanced Wavelet-based Diffusion Model for Hyperspectral Image Super-resolution

Sirui Wang, Jiang He, Natàlia Blasco Andreo, Xiao Xiang Zhu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This manuscript has been accepted for publication in AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07067 2025-11-11 cs.CV 79%

RaLD: Generating High-Resolution 3D Radar Point Clouds with Latent Diffusion

Ruijie Zhang, Bixin Zeng, Shengpeng Wang, Fuhui Zhou, Wei Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06948 2025-11-11 cs.CV 79%

PADM: A Physics-aware Diffusion Model for Attenuation Correction

Trung Kien Pham, Hoang Minh Vu, Anh Duc Chu, Dac Thai Nguyen, Trung Thanh Nguyen, Thao Nguyen Truong, Mai Hong Son, Thanh Trung Nguyen, Phi Le Nguyen

机构 * AI4LIFE, Hanoi University of Science and Technology, Vietnam(AI4LIFE,越南河内科学技术大学) Nagoya Univeristy, Japan(名古屋大学) National Institute of Advanced Industrial Science and Technology, Japan(日本国家先进工业科学和技术研究院) Military Central Hospital, Vietnam(越南108中央军医院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06823 2025-11-11 cs.CV 79%

Integrating Reweighted Least Squares with Plug-and-Play Diffusion Priors for Noisy Image Restoration

Ji Li, Chao Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05308 2025-11-11 cs.CV cs.AI cs.LG 79%

Rethinking Metrics and Diffusion Architecture for 3D Point Cloud Generation

Matteo Bastico, David Ryckelynck, Laurent Corté, Yannick Tillier, Etienne Decencière

机构 * Mines Paris, Université PSL(Mines Paris,Université PSL) Centre des Matériaux (MAT), UMR7633 CNRS(Centre des Matériaux(MAT),UMR7633 CNRS) Centre de Morphologie Mathématique (CMM)(Centre de Morphologie Mathématique(CMM)) Centre de Mise en Forme des Matériaux (CEMEF), UMR7635 CNRS(Centre de Mise en Forme des Matériaux(CEMEF),UMR7635 CNRS)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments This paper has been accepted at International Conference on 3D Vision (3DV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17383 2025-11-11 cs.LG cs.CV cs.CY 79%

The Evolving Nature of Latent Spaces: From GANs to Diffusion

Ludovica Schaerf

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Presented and published at Ethics and Aesthetics of Artificial Intelligence Conference (EA-AI'25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08009 2025-11-11 cs.CV cs.AI cs.LG 79%

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Xun Huang, Zhengqi Li, Guande He, Mingyuan Zhou, Eli Shechtman

机构 * Adobe Research(Adobe研究院) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments NeurIPS 2025 spotlight. Project website: http://self-forcing.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18167 2025-11-11 physics.med-ph cs.CV physics.bio-ph 79%

Scattering approach to diffusion quantifies axonal damage in brain injury

Ali Abdollahzadeh, Ricardo Coronado-Leija, Hong-Hsi Lee, Alejandra Sierra, Els Fieremans, Dmitry S. Novikov

机构 * Center for Biomedical Imaging, Department of Radiology, New York University School of Medicine, New York, NY, USA(纽约大学医学学院放射科生物医学成像中心) A.I. Virtanen Institute for Molecular Sciences, University of Eastern Finland, Kuopio, Finland(东部芬兰大学A.I. Virtanen分子科学研究所) Athinoula A. Martinos Center for Biomedical Imaging, Department of Radiology, Massachusetts General Hospital, Harvard Medical School, Boston, MA, USA(马萨诸塞州总医院哈佛医学院生物医学成像中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Nat Commun 16, 9808 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02261 2025-11-11 cs.CV 79%

Diffusion Implicit Policy for Unpaired Scene-aware Motion Synthesis

Jingyu Gong, Chong Zhang, Fengqi Liu, Ke Fan, Qianyu Zhou, Xin Tan, Zhizhong Zhang, Yuan Xie

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06422 2025-11-11 cs.CV 79%

DiffusionUavLoc: Visually Prompted Diffusion for Cross-View UAV Localization

Tao Liu, Kan Ren, Qian Chen

机构 * School of Electronic and Optical Engineering, Nanjing University of Science and Technology(电子与光学工程学院,南京理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06310 2025-11-11 cs.CV 79%

Adaptive 3D Reconstruction via Diffusion Priors and Forward Curvature-Matching Likelihood Updates

Seunghyeok Shin, Dabin Kim, Hongki Lim

机构 * Department of Electrical and Computer Engineering, Inha University(电子与计算机工程系,延世大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏