arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-07-31 至 2025-07-31 共收录 41 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 5 篇

2507.21391 2025-07-31 cs.CV cs.AI cs.CL 88%

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation

Shijie Zhou, Ruiyi Zhang, Huaisheng Zhu, Branislav Kveton, Yufan Zhou, Jiuxiang Gu, Jian Chen, Changyou Chen

机构 * University at Buffalo(布法罗大学) Adobe Research(Adobe研究) Pennsylvania State University(宾夕法尼亚州立大学)

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

Comments Accepted at ICCV 2025. Code available at https://github.com/sjz5202/LLaVA-Reward

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22100 2025-07-31 cs.CV 83%

Trade-offs in Image Generation: How Do Different Dimensions Interact?

Sicheng Zhang, Binzhu Xie, Zhonghao Yan, Yuli Zhang, Donghao Zhou, Xiaofei Chen, Shi Qiu, Jiaqi Liu, Guoyang Xie, Zhichao Lu

机构 * Khalifa University(卡利法大学) The Chinese University of Hong Kong(香港中文大学) Queen Mary University of London(伦敦大学玛丽女王学院) Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学) City University of Hong Kong(香港城市大学)

专题命中 文生图 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted in ICCV 2025, Codebase: https://github.com/fesvhtr/TRIG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22076 2025-07-31 cs.LG 82%

Test-time Prompt Refinement for Text-to-Image Models

Mohammad Abdul Hafeez Khan, Yash Jain, Siddhartha Bhattacharyya, Vibhav Vineet

机构 * Florida Institute of Technology(佛罗里达理工学院) Microsoft Research(微软研究院)

专题命中 文生图 :text-to-image(title,abstract);image generation(abstract)

Comments Accepted to ICCV 2025, MARS2 Workshop. Total 14 pages, 12 figures and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22617 2025-07-31 cs.CR cs.CV 70%

Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions

Yiting Qu, Ziqing Yang, Yihan Ma, Michael Backes, Savvas Zannettou, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA 欧洲信息安全研究中心) TU Delft(代尔夫特理工大学)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22499 2025-07-31 cs.LG cs.AI 67%

LoReUn: Data Itself Implicitly Provides Cues to Improve Machine Unlearning

Xiang Li, Qianli Shen, Haonan Wang, Kenji Kawaguchi

机构 * School of Computing(计算学院) National University of Singapore(新加坡国立大学)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract)

Comments 23 pages

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 扩散模型 29 篇

2507.00983 2025-07-31 eess.IV cs.CV 83%

DMCIE: Diffusion Model with Concatenation of Inputs and Errors to Improve the Accuracy of the Segmentation of Brain Tumors in MRI Images

Sara Yavari, Rahul Nitin Pandya, Jacob Furst

机构 * School of Computing, DePaul University(计算学院,德保罗大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22813 2025-07-31 cs.CV 79%

DISTIL: Data-Free Inversion of Suspicious Trojan Inputs via Latent Diffusion

Hossein Mirzaei, Zeinab Taghavi, Sepehr Rezaee, Masoud Hadi, Moein Madadi, Mackenzie W. Mathis

机构 * École Polytechnique Fédérale de Lausanne(洛桑联邦理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22615 2025-07-31 cs.CV 79%

Generative Active Learning for Long-tail Trajectory Prediction via Controllable Diffusion Model

Daehee Park, Monu Surana, Pranav Desai, Ashish Mehta, Reuben MV John, Kuk-Jin Yoon

机构 * Intelligent Systems and Learning Lab., DGIST, Korea(智能系统与学习实验室,DGIST,韩国) Qualcomm Research, USA(高通研究,美国) Visual Intelligence Lab., KAIST, Korea(视觉智能实验室,KAIST,韩国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22604 2025-07-31 cs.CV 79%

ShortFT: Diffusion Model Alignment via Shortcut-based Fine-Tuning

Xiefan Guo, Miaomiao Cui, Liefeng Bo, Di Huang

机构 * State Key Laboratory of Complex and Critical Software Environment(复杂与关键软件环境国家重点实验室) Beihang University(北京航空航天大学) School of Computer Science and Engineering(计算机科学与工程学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22501 2025-07-31 cs.CV eess.IV 79%

DACA-Net: A Degradation-Aware Conditional Diffusion Network for Underwater Image Enhancement

Chang Huang, Jiahang Cao, Jun Ma, Kieren Yu, Cong Li, Huayong Yang, Kaishun Wu

机构 * The Hong Kong University of Science and Technology (Guangzhou), Southern Marine Science and Engineering Guangdong Laboratory (Guangzhou)(香港科学与技术大学(广州)、广东海洋工程科学实验室(广州)) The Hong Kong University of Science(香港科学与技术大学) Engineering Guangdong Laboratory (Guangzhou)(广东工程实验室(广州))

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments accepted by ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22454 2025-07-31 cs.CV eess.IV 79%

TopoLiDM: Topology-Aware LiDAR Diffusion Models for Interpretable and Realistic LiDAR Point Cloud Generation

Jiuming Liu, Zheng Huang, Mengmeng Liu, Tianchen Deng, Francesco Nex, Hao Cheng, Hesheng Wang

机构 * School of Automation and Intelligent Sensing, Shanghai Jiao Tong University, Shanghai 200240(上海交通大学自动化与智能感知学院) Key Laboratory of System Control and Information Processing, Ministry of Education of China, Shanghai 200240(中国教育部系统控制与信息处理重点实验室) University of Twente, Netherlands(荷兰埃因霍温理工大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by IROS 2025. Code:https://github.com/IRMVLab/TopoLiDM

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22360 2025-07-31 cs.CV cs.AI 79%

GVD: Guiding Video Diffusion Model for Scalable Video Distillation

Kunyang Li, Jeffrey A Chan Santiago, Sarinda Dhanesh Samarasinghe, Gaowen Liu, Mubarak Shah

机构 * Center for Research in Computer Vision, University of Central Florida(计算机视觉研究中心,中央佛罗里达大学) Cisco Research(思科研究)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20590 2025-07-31 cs.CV 79%

Harnessing Diffusion-Yielded Score Priors for Image Restoration

Xinqi Lin, Fanghua Yu, Jinfan Hu, Zhiyuan You, Wu Shi, Jimmy S. Ren, Jinjin Gu, Chao Dong

机构 * Shenzhen Institute of Advanced Technology, CAS(深圳先进技术研究院,中国科学院) University of Chinese Academy of Sciences(中国科学院大学) The Chinese University of Hong Kong(香港中文大学) SenseTime Research(商汤科技研究院) Hong Kong Metropolitan University(香港 Metropolitan 大学) INSAIT, Sofia University(INSAIT,索菲亚大学) Shenzhen Institutes of Advanced Technology, CAS(深圳先进技术研究院,中国科学院) Shenzhen University of Advanced Technology(深圳大学先进技术学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05363 2025-07-31 cs.CV 79%

Seed Selection for Human-Oriented Image Reconstruction via Guided Diffusion

Yui Tatsumi, Ziyue Zeng, Hiroshi Watanabe

机构 * Graduate School of FSE, Waseda University Tokyo, Japan(FSE研究生院,早稻田大学东京,日本)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by 2025 IEEE 14th Global Conference on Consumer Electronics (GCCE 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04283 2025-07-31 cs.AI cs.CV 79%

Clustering via Self-Supervised Diffusion

Roy Uziel, Irit Chelly, Oren Freifeld, Ari Pakman

机构 * Department of Computer Science, Ben-Gurion University of the Negev, Beer Sheva, Israel(计算机科学系,贝尔谢巴以色列内盖夫比尔亚尔大学) Department of Industrial Engineering and Management, Ben-Gurion University of the Negev, Beer Sheva, Israel(工业工程与管理系,贝尔谢巴以色列内盖夫比尔亚尔大学) The School of Brain Sciences and Cognition, Ben-Gurion University of the Negev, Beer Sheva, Israel(脑科学与认知学院,贝尔谢巴以色列内盖夫比尔亚尔大学) Data Science Research Center, Ben-Gurion University of the Negev, Beer Sheva, Israel(数据科学研究中心,贝尔谢巴以色列内盖夫比尔亚尔大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20104 2025-07-31 cs.CV cs.AI cs.LG cs.RO 79%

SyncDiff: Synchronized Motion Diffusion for Multi-Body Human-Object Interaction Synthesis

Wenkun He, Yun Liu, Ruitao Liu, Li Yi

机构 * Tsinghua University(清华大学) Shanghai Qi Zhi Institute(上海启智研究院) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 27 pages, 10 figures, 20 tables. Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22589 2025-07-31 cs.SI 78%

Diffusion Models for Influence Maximization on Temporal Networks: A Guide to Make the Best Choice

Aaqib Zahoor, Iqra Altaf Gillani, Janibul Bashir

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22452 2025-07-31 math.DS 78%

Saddle-point structure of fixed points in a reaction-diffusion equation with discontinuous nonlinearity

José Valero

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22416 2025-07-31 math.DS 78%

Arnold diffusion in the elliptic Hill four-body problem: geometric method and numerical verification

Jaime Burgos, Marian Gidea, Claudio Sierpe

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22385 2025-07-31 math.OC cs.LG cs.SY eess.SY math.PR stat.ME 78%

Set Invariance with Probability One for Controlled Diffusion: Score-based Approach

Wenqing Wang, Alexis M. H. Teter, Murat Arcak, Abhishek Halder

机构 * Department of Aerospace Engineering, Iowa State University(航空航天工程系,爱荷华州立大学) Department of Applied Mathematics, University of California, Santa Cruz(应用数学系,加州大学圣克鲁兹分校) Department of Electrical Engineering and Computer Sciences, University of California, Berkeley(电气工程与计算机科学系,加州大学伯克利分校)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06427 2025-07-31 physics.med-ph eess.IV 78%

Estimation of Multi-Component Flow in the Kidney with Multi-b-value Spectral Diffusion

Mira M. Liu, Thomas Gladytz, Jonathan Dyke, Ian Bolger, Jonas Jasse, Sergio Calle, Tanner Crews, Surya Seshan, Steven Salvatore, Isaac Stillman, Thangamani Muthukumar, Bachir Taouli, Samira Farouk, Octavia Bane, Sara Lewis

专题命中 扩散模型 :diffusion(title,abstract)

Comments Version accepted for publication in Magnetic Resonance in Imaging. Published version available at https://doi.org/10.1002/mrm.30644

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10289 2025-07-31 cs.CV 74%

MaterialMVP: Illumination-Invariant Material Generation via Multi-view PBR Diffusion

Zebin He, Mingxin Yang, Shuhui Yang, Yixuan Tang, Tao Wang, Kaihao Zhang, Guanying Chen, Yuhong Liu, Jie Jiang, Chunchao Guo, Wenhan Luo

机构 * Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区) Tencent Hunyuan(腾讯文元) The Hong Kong University of Science and Technology(香港科技大学) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))

专题命中 扩散模型 :diffusion(title);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22692 2025-07-31 cs.CV 57%

Zero-Shot Image Anomaly Detection Using Generative Foundation Models

Lemar Abdi, Amaan Valiuddin, Francisco Caetano, Christiaan Viviers, Fons van der Sommen

机构 * Eindhoven University of Technology(埃因霍温理工大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Accepted at the workshop of Anomaly Detection with Foundation Models, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22469 2025-07-31 cs.CV cs.AI cs.LG 57%

Visual Language Models as Zero-Shot Deepfake Detectors

Viacheslav Pirogov

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Accepted to the ICML 2025 Workshop on Reliable and Responsible Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21132 2025-07-31 cs.CV 57%

Learning to See in the Extremely Dark

Hai Jiang, Binhao Guan, Zhen Liu, Xiaohong Liu, Jian Yu, Zheng Liu, Songchen Han, Shuaicheng Liu

机构 * School of Aeronautics and Astronautics, Sichuan University(四川大学航空宇航学院) University of Electronic Science and Technology of China(电子科技大学) Shanghai Jiao Tong University(上海交通大学) National Innovation Center for UHD Video Technology(超高清视频技术国家创新中心)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22880 2025-07-31 cs.IR 50%

AUV-Fusion: Cross-Modal Adversarial Fusion of User Interactions and Visual Perturbations Against VARS

Hai Ling, Tianchi Wang, Xiaohao Liu, Zhulin Tao, Lifang Yang, Xianglin Huang

专题命中 扩散模型 :diffusion(abstract)

Comments 14 pages,6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22803 2025-07-31 physics.optics 50%

Design and Analysis of Plasmonic-Nanorod-Enhanced Lead-Free Inorganic Perovskite/Silicon Heterojunction Tandem Solar Cell Exceeding the Shockley-Queisser Limit

Md. Sad Abdullah Sami, Arpan Sur, Ehsanur Rahman

专题命中 扩散模型 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22657 2025-07-31 physics.med-ph 50%

Viscoelastic Profiling of Rare Pediatric Extracranial Tumors using Multifrequency MR Elastography: A Pilot Study

C. Metz, S. Veldhoen, H. E. Deubzer, F. Mollica, T. Meyer, K. Hauptmann, A. H. Hagemann, A. Eggert, I. Sack, M. S. Anders

专题命中 扩散模型 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22582 2025-07-31 math.AP 50%

Regularity Properties of Solutions of a Model for Morphoelastic Growth in the Presence of Nutrients in One Spatial Dimension

Julian Blawid, Georg Dolzmann

专题命中 扩散模型 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06387 2025-07-31 math.AP 50%

Three solutions for a double phase variable exponent Kirchhoff problem

Mustafa Avci

专题命中 扩散模型 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏