arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-09-26 至 2025-09-26 共收录 49 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 33 篇

2506.06721 2025-09-26 cond-mat.mes-hall cond-mat.mtrl-sci cond-mat.str-el physics.comp-ph 50%

Electronic structure and transport in materials with flat bands: 2D materials and quasicrystals

Guy Trambly de Laissardière, Somepalli Venkateswarlu, Ahmed Missaoui, Ghassen Jemaï, Khouloud Chika, Javad Vahedi, Omid Faizy Namarvar, Jean-Pierre Julien, Andreas Honecker, Laurence Magaud, Jouda Jemaa Khabthani, Didier Mayou

专题命中 扩散模型 :diffusion(abstract)

Comments review

Journal ref Physica E 175 (2026) 116362

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13253 2025-09-26 cs.LG 50%

Conditional Denoising Meets Polynomial Modeling: A Flexible Decoupled Framework for Time Series Forecasting

Jintao Zhang, Mingyue Cheng, Xiaoyu Tao, Zhiding Liu, Daoyu Wang

机构 * State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(认知智能国家重点实验室,中国科学技术大学)

专题命中 扩散模型 :diffusion(abstract)

Journal ref Proceedings of the Thirty-Fourth International Joint Conference on Artificial Intelligence 2025, Main Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08350 2025-09-26 math.DS physics.flu-dyn 50%

Dissolution of variable-in-shape drug particles via the level-set method

Emiliano Cristiani, Mario Grassi, Francesca L. Ignoto, Giuseppe Pontrelli

专题命中 扩散模型 :diffusion(abstract)

Journal ref Appl. Math. Model., 142 (2025), 115966

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.11901 2025-09-26 cs.SI cs.LG 50%

A Deep Learning Framework for Evaluating Dynamic Network Generative Models and Anomaly Detection

Alireza Rashnu, Sadegh Aliakbary

专题命中 扩散模型 :diffusion(abstract)

Journal ref Journal of Innovations in Computer Science and Engineering (JICSE), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.06719 2025-09-26 quant-ph physics.comp-ph 50%

Quantum Algorithm for Smoothed Particle Hydrodynamics

Rhonda Au-Yeung, Anthony J. Williams, Viv M. Kendon, Steven J. Lind

专题命中 扩散模型 :diffusion(abstract)

Comments Published in Computer Physics Communications, 14 pages

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 可控生成 5 篇

2509.20524 2025-09-26 cs.CV cs.AI 83%

InstructVTON: Optimal Auto-Masking and Natural-Language-Guided Interactive Style Control for Inpainting-Based Virtual Try-On

Julien Han, Shuwen Qiu, Qi Li, Xingzi Xu, Mehmet Saygin Seyfioglu, Kavosh Asadi, Karim Bouyarmane

机构 * Amazon(亚马逊公司) University of California, Los Angeles (UCLA)(加州大学洛杉矶分校) Duke University(杜克大学)

专题命中 可控生成 :inpainting(title,abstract);image generation(abstract);分类 cs.CV

Comments Submitted to CVPR 2025 and Published at CVPR 2025 AI for Content Creation workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13065 2025-09-26 cs.CV 79%

Odo: Depth-Guided Diffusion for Identity-Preserving Body Reshaping

Siddharth Khandelwal, Sridhar Kamath, Arjun Jain

机构 * Fast Code AI Consult Pvt. Ltd.(Fast Code AI咨询私有有限公司)

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21119 2025-09-26 cs.CV 57%

MotionFlow:Learning Implicit Motion Flow for Complex Camera Trajectory Control in Video Generation

Guojun Lei, Chi Wang, Yikai Wang, Hong Li, Ying Song, Weiwei Xu

机构 * Zhejiang University(浙江大学) Tsinghua University(清华大学) Beihang University(北航) Zhejiang Gongshang University(浙江工商大学) ShengShu(盛舒)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments ICME2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15269 2025-09-26 cs.CV cs.AI 57%

Conditional Video Generation for High-Efficiency Video Compression

Fangqiu Yi, Jingyu Xu, Jiawei Shao, Chi Zhang, Xuelong Li

机构 * Institute of Artificial Intelligence (TeleAI)(人工智能研究院)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08705 2025-09-26 cs.CV cs.AI 57%

Instance-aware Image Colorization with Controllable Textual Descriptions and Segmentation Masks

Yanru An, Ling Gui, Chunlei Cai, Tianxiao Ye, JIangchao Yao, Guangtao Zhai, Qiang Hu, Xiaoyun Zhang

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 图像修复 1 篇

2509.20852 2025-09-26 cs.LG cs.AI cs.CE cs.CV 79%

FHRFormer: A Self-supervised Transformer Approach for Fetal Heart Rate Inpainting and Forecasting

Kjersti Engan, Neel Kanwal, Anita Yeconia, Ladislaus Blacy, Yuda Munyaw, Estomih Mduma, Hege Ersdal

机构 * University of Stavanger(斯塔万格大学) Haydom Lutheran Hospital(海多姆路德医院) Stavanger University Hospital(斯塔万格大学医院)

专题命中 图像修复 :inpainting(title,abstract);分类 cs.CV

Comments Submitted to IEEE JBHI

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 个性化与一致性 2 篇

2509.20756 2025-09-26 cs.CV 81%

FreeInsert: Personalized Object Insertion with Geometric and Style Control

Yuhong Zhang, Han Wang, Yiwen Wang, Rong Xie, Li Song

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 个性化与一致性 :image generation(abstract);text-to-image(abstract);diffusion(abstract);image editing(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20775 2025-09-26 cs.CV cs.AI 70%

CusEnhancer: A Zero-Shot Scene and Controllability Enhancement Method for Photo Customization via ResInversion

Maoye Ren, Praneetha Vaddamanu, Jianjin Xu, Fernando De la Torre Frade

机构 * College of Computer Science and Technology, East China University of Science and Technology(计算机科学与技术学院,东华大学) Carnegie Mellon University(卡内基梅隆大学) Microsoft(微软公司)

专题命中 个性化与一致性 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 图像生成评测 1 篇

2506.23275 2025-09-26 cs.CV cs.AI 70%

Why Settle for One? Text-to-ImageSet Generation and Evaluation

Chengyou Jia, Xin Shen, Zhuohang Dang, Zhuohang Dang, Changliang Xia, Weijia Wu, Xinyu Zhang, Hangwei Qian, Ivor W. Tsang, Minnan Luo

机构 * Xi’an Jiaotong University(西安交通大学) National University of Singapore(新加坡国立大学) A*STAR(新加坡科技研究局)

专题命中 图像生成评测 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 效率与蒸馏 5 篇

2412.02687 2025-09-26 cs.CV 88%

Supercharged One-step Text-to-Image Diffusion Models with Negative Prompts

Viet Nguyen, Anh Nguyen, Trung Dao, Khoi Nguyen, Cuong Pham, Toan Tran, Anh Tran

机构 * Qualcomm AI Research(高通人工智能研究)

专题命中 效率与蒸馏 :diffusion(title,abstract);text-to-image(title);image synthesis(abstract);分类 cs.CV

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.02080 2025-09-26 cs.LG stat.CO stat.ML 78%

Energy based diffusion generator for efficient sampling of Boltzmann distributions

Yan Wang, Ling Guo, Hao Wu, Tao Zhou

机构 * School of Mathematical Sciences, Tongji University(同济大学数学科学学院) Department of Mathematics, Shanghai Normal University(上海师范大学数学系) Institute of Computational Mathematics and Scientific/Engineering Computing, AMSS, Chinese Academy of Sciences(中国科学院数学与系统科学研究院)

专题命中 效率与蒸馏 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21318 2025-09-26 cs.CV cs.AI 57%

SD3.5-Flash: Distribution-Guided Distillation of Generative Flows

Hmrishav Bandyopadhyay, Rahim Entezari, Jim Scott, Reshinth Adithyan, Yi-Zhe Song, Varun Jampani

机构 * Stability AI SketchX, University of Surrey(SketchX,大学)

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

Comments Project Page: https://hmrishavbandy.github.io/sd35flash/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21027 2025-09-26 cs.RO cs.CV 57%

KeyWorld: Key Frame Reasoning Enables Effective and Efficient World Models

Sibo Li, Qianyue Hao, Yu Shang, Yong Li

机构 * Department of Electronic Engineering, BNRist, Tsinghua University(电子工程系、北京理工大学、清华大学)

专题命中 效率与蒸馏 :inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20744 2025-09-26 cs.AI 50%

Parallel Thinking, Sequential Answering: Bridging NAR and AR for Efficient Reasoning

Qihang Ai, Haiyun Jiang

机构 * Nanyang Technological University(南洋理工大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 效率与蒸馏 :diffusion(abstract)

Comments 4 pages

详情

展开后加载摘要…

URL PDF HTML 收藏