arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-08-06 至 2025-08-06 共收录 64 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 7 篇

2508.03481 2025-08-06 cs.CV cs.AI cs.CL 91%

Draw Your Mind: Personalized Generation via Condition-Level Modeling in Text-to-Image Diffusion Models

Hyungjin Kim, Seokho Ahn, Young-Duk Seo

机构 * Department of Electrical and Computer Engineering, Inha University(电子与计算机工程系,inha大学)

专题命中 文生图 :diffusion(title,abstract);personalized generation(title,abstract);text-to-image(title);分类 cs.CV

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03006 2025-08-06 cs.CV 89%

Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models

Fan Yang, Yihao Huang, Jiayi Zhu, Ling Shi, Geguang Pu, Jin Song Dong, Kailong Wang

机构 * Huazhong University of Science and Technology(华中科技大学) National University of Singapore(新加坡国立大学) East China Normal University(华东师范大学) Nanyang Technological University(南洋理工大学)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03320 2025-08-06 cs.CV 77%

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation

Peiyu Wang, Yi Peng, Yimeng Gan, Liang Hu, Tianyidan Xie, Xiaokun Wang, Yichen Wei, Chuanxin Tang, Bo Zhu, Changshi Li, Hongyang Wei, Eric Li, Xuchen Song, Yang Liu, Yahui Zhou

机构 * Multimodality Team, Skywork AI(Skywork AI 多模态团队)

专题命中 文生图 :image generation(abstract);text-to-image(abstract);image editing(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03415 2025-08-06 cs.CV cs.AI cs.GR 73%

Learning Latent Representations for Image Translation using Frequency Distributed CycleGAN

Shivangi Nigam, Adarsh Prasad Behera, Shekhar Verma, P. Nagabhushan

机构 * Department of Information Technology, Indian Institute of Information Technology, Allahabad, U.P.(信息科技系,印度信息技术研究所,阿勒普尔,乌塔兰邦) KTH Royal Institute of Technology(皇家理工学院) Department of Computer Science and Engineering, Vignan University, Guntur, Andhra Pradesh(计算机科学与工程系,维加恩大学,古特尔,安得拉邦)

专题命中 文生图 :diffusion(abstract);image synthesis(abstract);分类 cs.CV、cs.GR

Comments This paper is currently under review for publication in an IEEE Transactions. If accepted, the copyright will be transferred to IEEE

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02937 2025-08-06 cs.CY 71%

Documenting Patterns of Exoticism of Marginalized Populations within Text-to-Image Generators

Sourojit Ghosh, Sanjana Gautam, Pranav Venkit, Avijit Ghosh

专题命中 文生图 :text-to-image(title)

Comments Upcoming Publication, AIES 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03535 2025-08-06 cs.CV 70%

CoEmoGen: Towards Semantically-Coherent and Scalable Emotional Image Content Generation

Kaishen Yuan, Yuting Zhang, Shang Gao, Yijie Zhu, Wenshuo Chen, Yutao Yue

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments 10 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03091 2025-08-06 cs.AI cs.CR cs.CV 57%

T2UE: Generating Unlearnable Examples from Text Descriptions

Xingjun Ma, Hanxun Huang, Tianwei Song, Ye Sun, Yifeng Gao, Yu-Gang Jiang

机构 * Fudan University(复旦大学) The University of Melbourne(墨尔本大学)

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments To appear in ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 图像编辑 1 篇

2508.03480 2025-08-06 cs.CV cs.AI 57%

VideoGuard: Protecting Video Content from Unauthorized Editing

Junjie Cao, Kaizhou Li, Xinchun Yu, Hongxiang Li, Xiaoping Zhang

专题命中 图像编辑 :diffusion(abstract);分类 cs.CV

Comments ai security, 10pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 扩散模型 42 篇

2506.14404 2025-08-06 cs.CV cs.AI 83%

Causally Steered Diffusion for Automated Video Counterfactual Generation

Nikos Spyrou, Athanasios Vlontzos, Paraskevas Pegios, Thomas Melistas, Nefeli Gkouti, Yannis Panagakis, Giorgos Papanastasiou, Sotirios A. Tsaftaris

机构 * Spotify, UK(英国Spotify)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00482 2025-08-06 cs.CV cs.AI 83%

JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion Transformers

Kwon Byung-Ki, Qi Dai, Lee Hyoseok, Chong Luo, Tae-Hyun Oh

机构 * POSTECH Microsoft Research Asia(微软亚洲研究院) KAIST(韩国科学技术院)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

Comments Accepted to IEEE/CVF International Conference on Computer Vision (ICCV) 2025. Project page: https://byungki-k.github.io/JointDiT/ Code: https://github.com/kaist-ami/JointDiT

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14432 2025-08-06 cs.CV eess.IV 83%

IntroStyle: Training-Free Introspective Style Attribution using Diffusion Features

Anand Kumar, Jiteng Mu, Nuno Vasconcelos

机构 * University of California, San Diego(加州大学圣地亚哥分校)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV

Comments 17 pages, 16 figures

Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03645 2025-08-06 cs.RO cs.CV cs.LG 79%

DiWA: Diffusion Policy Adaptation with World Models

Akshay L Chandra, Iman Nematollahi, Chenguang Huang, Tim Welschehold, Wolfram Burgard, Abhinav Valada

机构 * University of Freiburg(弗赖堡大学) University of Technology Nuremberg(纽伦堡技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at the 2025 Conference on Robot Learning (CoRL)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03256 2025-08-06 cs.CV 79%

Beyond Isolated Words: Diffusion Brush for Handwritten Text-Line Generation

Gang Dai, Yifan Zhang, Yutao Qin, Qiangya Guo, Shuangping Huang, Shuicheng Yan

机构 * South China University of Technology(华南理工大学) MiroMind AI National University of Singapore(新加坡国立大学) Pazhou Laboratory(Pazhou 实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments To appear in ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02973 2025-08-06 cs.CV 79%

Diffusion Models with Adaptive Negative Sampling Without External Resources

Alakh Desai, Nuno Vasconcelos

机构 * University of California, San Diego(加州大学圣地亚哥分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02807 2025-08-06 cs.CV 79%

DreamVVT: Mastering Realistic Video Virtual Try-On in the Wild via a Stage-Wise Diffusion Transformer Framework

Tongchun Zuo, Zaiyu Huang, Shuliang Ning, Ente Lin, Chao Liang, Zerong Zheng, Jianwen Jiang, Yuan Zhang, Mingyuan Gao, Xin Dong

机构 * ByteDance Intelligent Creation(字节跳动智能创作)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 18 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17350 2025-08-06 cs.CV 79%

Decouple and Track: Benchmarking and Improving Video Diffusion Transformers for Motion Transfer

Qingyu Shi, Jianzong Wu, Jinbin Bai, Jiangning Zhang, Lu Qi, Yunhai Tong, Xiangtai Li

机构 * PKU(北京大学) NTU(国立新加坡大学) NUS(新加坡国立大学) ZJU(浙江大学) UC Merced(加州大学默塞德分校)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11801 2025-08-06 cs.GR cs.LG cs.RO 79%

Diffuse-CLoC: Guided Diffusion for Physics-based Character Look-ahead Control

Xiaoyu Huang, Takara Truong, Yunbo Zhang, Fangzhou Yu, Jean Pierre Sleiman, Jessica Hodgins, Koushil Sreenath, Farbod Farshidian

机构 * University of California, Berkeley(加州大学伯克利分校) RAI Institute(RAI研究所) Stanford University(斯坦福大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18375 2025-08-06 cs.CV eess.IV 79%

Individual Content and Motion Dynamics Preserved Pruning for Video Diffusion Models

Yiming Wu, Zhenghao Chen, Huan Wang, Dong Xu

机构 * The University of Hong Kong(香港大学) University of Newcastle(新castle大学) Westlake University(西湖大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17456 2025-08-06 cs.CV cs.LG eess.IV 79%

Generalized Compressed Sensing for Image Reconstruction with Diffusion Probabilistic Models

Ling-Qi Zhang, Zahra Kadkhodaie, Eero P. Simoncelli, David H. Brainard

机构 * Janelia Research Campus, Howard Hughes Medical Institute(贾勒尼亚研究campus,霍华德·休斯医学研究所) Flatiron Institute, Simons Foundation(Flatiron研究所,Simons基金会) New York University(纽约大学) University of Pennsylvania(宾夕法尼亚大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Transactions on Machine Learning Research (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08808 2025-08-06 cs.LG math.OC stat.ML 79%

A First-order Generative Bilevel Optimization Framework for Diffusion Models

Quan Xiao, Hui Yuan, A F M Saif, Gaowen Liu, Ramana Kompella, Mengdi Wang, Tianyi Chen

机构 * Rensselaer Polytechnic Institute(伦斯勒理工学院) Princeton University(普林斯顿大学) Cisco Research(思科研究) Cornell University(康奈尔大学)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Cameral-ready version: added experiments using the HPSv2 reward, improved notation consistency for the diffusion model, and added related works

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03617 2025-08-06 math.ST stat.TH 78%

Expanding the Standard Diffusion Process to Specified Non-Gaussian Marginal Distributions

Robert Richardson, H. Dennis Tolley, Kenneth Kuttler

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11117 2025-08-06 math.AP 78%

Global existence and decay rates of strong solutions to the diffusion approximation model in radiation hydrodynamics

Peng Jiang, Fucai Li, Jinkai Ni

专题命中 扩散模型 :diffusion(title,abstract)

Comments 29 pages

Journal ref SIAM Journal on Mathematical Analysis 57 (2025), no. 4, 3755--3783

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03123 2025-08-06 cs.SD cs.AI eess.AS 78%

Fine-Tuning Text-to-Speech Diffusion Models Using Reinforcement Learning with Human Feedback

Jingyi Chen, Ju Seung Byun, Micha Elsner, Pichao Wang, Andrew Perrault

机构 * Department of Linguistics(语言学系) Department of Computer Science and Engineering(计算机科学与工程系)

专题命中 扩散模型 :diffusion(title,abstract)

Comments 4 pages, 1 figure, INTERSPEECH 2025. arXiv admin note: text overlap with arXiv:2405.14632

Journal ref INTERSPEECH 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03083 2025-08-06 cs.AI 78%

MissDDIM: Deterministic and Efficient Conditional Diffusion for Tabular Data Imputation

Youran Zhou, Mohamed Reda Bouadjenek, Sunil Aryal

机构 * Deakin University(德克萨斯大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03042 2025-08-06 cs.LG 78%

Urban In-Context Learning: Bridging Pretraining and Inference through Masked Diffusion for Urban Profiling

Ruixing Zhang, Bo Wang, Tongyu Zhu, Leilei Sun, Weifeng Lv

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12313 2025-08-06 cond-mat.stat-mech cond-mat.soft math-ph math.MP physics.chem-ph 78%

Velocity Distribution and Diffusion of an Athermal Inertial Run-and-Tumble Particle in a Shear-Thickening Medium

Subhanker Howlader, Sayantan Mondal, Prasenjit Das

专题命中 扩散模型 :diffusion(title,abstract)

Comments 17 Pages, 6 Figures, Accepted in Phys. Rev. E

Journal ref Phys. Rev. E 112, 025403 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.19502 2025-08-06 cs.RO 78%

Simultaneous Pick and Place Detection by Combining SE(3) Diffusion Models with Differential Kinematics

Tianyi Ko, Takuya Ikeda, Balazs Opra, Koichi Nishiwaki

机构 * Woven by Toyota, Inc.(丰田公司)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Accepted for IROS2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13754 2025-08-06 cond-mat.stat-mech math-ph math.MP math.PR 78%

Large Deviations in Switching Diffusion: from Free Cumulants to Dynamical Transitions

Mathis Guéneau, Satya N. Majumdar, Gregory Schehr

专题命中 扩散模型 :diffusion(title,abstract)

Comments Letter: 7+2 pages and 3 figures; Supp. Mat.: 32 pages and 9 figures

Journal ref Phys. Rev. Lett. 135, 067102 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14499 2025-08-06 cs.CR cs.LG 78%

PrivDiffuser: Privacy-Guided Diffusion Model for Data Obfuscation in Sensor Networks

Xin Yang, Omid Ardakanian

机构 * University of Alberta(阿尔伯塔大学)

专题命中 扩散模型 :diffusion(title,abstract)

Journal ref Proceedings on Privacy Enhancing Technologies, 2025(4), 40-55

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05222 2025-08-06 cond-mat.stat-mech 71%

Diffusion cascade in a model of interacting random walkers

Abhishek Raj, Paolo Glorioso, Sarang Gopalakrishnan, Vadim Oganesyan

专题命中 扩散模型 :diffusion(title)

Comments 12 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏