arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-09-16 至 2025-09-16 共收录 76 信号源:cs.CV, cs.GR, cs.MM

1. 可控生成 8 篇

2507.20083 2025-09-16 cs.CV 83%

KB-DMGen: Knowledge-Based Global Guidance and Dynamic Pose Masking for Human Image Generation

Shibang Liu, Xuemei Xie, Guangming Shi

机构 * Shibang Liu Xuemei Xie Guangming Shi

专题命中 可控生成 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04126 2025-09-16 cs.CV cs.AI 83%

MEPG:Multi-Expert Planning and Generation for Compositionally-Rich Image Generation

Yuan Zhao, Lin Liu

专题命中 可控生成 :image generation(title);text-to-image(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10463 2025-09-16 cs.LG cs.CV 61%

The 1st International Workshop on Disentangled Representation Learning for Controllable Generation (DRL4Real): Methods and Results

Qiuyu Chen, Xin Jin, Yue Song, Xihui Liu, Shuai Yang, Tao Yang, Ziqiang Li, Jianguo Huang, Yuntao Wei, Ba'ao Xie, Nicu Sebe, Wenjun, Zeng, Jooyeol Yun, Davide Abati, Mohamed Omran, Jaegul Choo, Amir Habibian, Auke Wiggers, Masato Kobayashi, Ning Ding, Toru Tamaki, Marzieh Gheisari, Auguste Genovesio, Yuheng Chen, Dingkun Liu, Xinyao Yang, Xinping Xu, Baicheng Chen, Dongrui Wu, Junhao Geng, Lexiang Lv, Jianxin Lin, Hanzhe Liang, Jie Zhou, Xuanxin Chen, Jinbao Wang, Can Gao, Zhangyi Wang, Zongze Li, Bihan Wen, Yixin Gao, Xiaohan Pan, Xin Li, Zhibo Chen, Baorui Peng, Zhongming Chen, Haoran Jin

专题命中 可控生成 :diffusion(abstract,comments);分类 cs.CV

Comments Workshop summary paper for ICCV 2025, 9 accepted papers, 9 figures, IEEE conference format, covers topics including diffusion models, controllable generation, 3D-aware disentanglement, autonomous driving applications, and EEG analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11840 2025-09-16 cs.CV 57%

Synthetic Captions for Open-Vocabulary Zero-Shot Segmentation

Tim Lebailly, Vijay Veerabadran, Satwik Kottur, Karl Ridgeway, Michael Louis Iuzzolino

机构 * Meta KU Leuven(鲁汶大学)

专题命中 可控生成 :generative vision(abstract);分类 cs.CV

Comments ICCV 2025 CDEL Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.19189 2025-09-16 cs.CV cs.LG 57%

An End-to-End Depth-Based Pipeline for Selfie Image Rectification

Ahmed Alhawwary, Janne Mustaniemi, Phong Nguyen-Ha, Janne Heikkilä

机构 * Center for Machine Vision and Signal Analysis (CMVS), University of Oulu(机器视觉与信号分析中心(CMVS)、奥卢大学) Qualcomm AI Research(高通人工智能研究)

专题命中 可控生成 :inpainting(abstract);分类 cs.CV

Comments Accepted at IEEE TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02353 2025-09-16 cs.CV cs.AI cs.LG 57%

Semantic Augmentation in Images using Language

Sahiti Yerramilli, Jayant Sravan Tamarapalli, Tanmay Girish Kulkarni, Jonathan Francis, Eric Nyberg

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10466 2025-09-16 cs.CV cs.HC 57%

A Real-Time Diminished Reality Approach to Privacy in MR Collaboration

Christian Fane

专题命中 可控生成 :inpainting(abstract);分类 cs.CV

Comments 50 pages, 12 figures | Demo video: https://youtu.be/udBxj35GEKI?t=499 | Code: https://github.com/c1h1r1i1s1 (multiple repositories)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 图像修复 1 篇

2509.10894 2025-09-16 physics.app-ph 50%

A novel IR-SRGAN assisted super-resolution evaluation of photothermal coherence tomography for impact damage in toughened thermoplastic CFRP laminates under room temperature and low temperature

Pengfei Zhu, Hai Zhang, Stefano Sfarra, Fabrizio Sarasini, Zijing Ding, Clemente Ibarra-Castanedo, Xavier Maldague

专题命中 图像修复 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 个性化与一致性 2 篇

2509.12001 2025-09-16 eess.IV cs.CV 57%

Data-driven Smile Design: Personalized Dental Aesthetics Outcomes Using Deep Learning

Marcus Lin, Jennifer Lai

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

Comments 6 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11092 2025-09-16 cs.CV cs.AI 57%

PanoLora: Bridging Perspective and Panoramic Video Generation with LoRA Adaptation

Zeyu Dong, Yuyang Yin, Yuqi Li, Eric Li, Hao-Xiang Guo, Yikai Wang

机构 * School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院) Beijing Jiaotong University(北京交通大学) The City College of New York(纽约城市学院) Skywork AI

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 图像生成评测 2 篇

2505.18985 2025-09-16 cs.LG cs.CL cs.CV 77%

STRICT: Stress Test of Rendering Images Containing Text

Tianyu Zhang, Xinyu Wang, Lu Li, Zhenghan Tai, Jijun Chi, Jingrui Tian, Hailin He, Suyuchen Wang

机构 * Mila, University of Montreal(蒙特利尔大学Mila) McGill University(麦吉尔大学) University of Pennsylvania(宾夕法尼亚大学) University of Toronto(多伦多大学) University of California, Los Angeles(加州大学洛杉矶分校) Southwestern University of Finance and Economics(西南财经大学)

专题命中 图像生成评测 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments Accepted as a main conference paper at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11959 2025-09-16 cs.CV cs.RO 57%

Learning to Generate 4D LiDAR Sequences

Ao Liang, Youquan Liu, Yu Yang, Dongyue Lu, Linfeng Li, Lingdong Kong, Huaici Zhao, Wei Tsang Ooi

机构 * NUS(国立新加坡大学) UCAS(中国科学院大学) SIA, CAS(中国科学院上海自动化研究所) FDU(福建大学) ZJU(浙江大学)

专题命中 图像生成评测 :diffusion(abstract);分类 cs.CV

Comments Abstract Paper (Non-Archival) @ ICCV 2025 Wild3D Workshop; GitHub Repo at https://lidarcrafter.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 效率与蒸馏 4 篇

2310.06389 2025-09-16 cs.CV cs.LG stat.ML 83%

Learning Stackable and Skippable LEGO Bricks for Efficient, Reconfigurable, and Variable-Resolution Diffusion Modeling

Huangjie Zheng, Zhendong Wang, Jianbo Yuan, Guanghan Ning, Pengcheng He, Quanzeng You, Hongxia Yang, Mingyuan Zhou

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校) ByteDance Inc.(字节跳动公司) Microsoft Azure AI(微软Azure人工智能)

专题命中 效率与蒸馏 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11047 2025-09-16 cs.LG cs.CV 79%

Data-Efficient Ensemble Weather Forecasting with Diffusion Models

Kevin Valencia, Ziyang Liu, Justin Cui

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 效率与蒸馏 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11284 2025-09-16 cs.LG physics.comp-ph 50%

PINGS: Physics-Informed Neural Network for Fast Generative Sampling

Achmad Ardani Prasha, Clavino Ourizqi Rachmadi, Muhamad Fauzan Ibnu Syahlan, Naufal Rahfi Anugerah, Nanda Garin Raditya, Putri Amelia, Sabrina Laila Mutiara, Hilman Syachr Ramadhan

机构 * Faculty of Computer Science, Universitas Mercu Buana(计算机科学学院,默克鲁巴纳大学) Alumni, Universitas Mercu Buana(默克鲁巴纳大学校友)

专题命中 效率与蒸馏 :diffusion(abstract)

Comments 19 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20514 2025-09-16 hep-th math-ph math.MP 50%

Spontaneous quantization of the Yang--Mills gradient flow

Alexander Migdal

专题命中 效率与蒸馏 :diffusion(abstract)

Comments 26 pages, 5 figures, This is the final revision submitted to NPB, with added discussion of previous analysis of minimal surfaces as solution of the loop equation

详情

展开后加载摘要…

URL PDF HTML 收藏