arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-12-04 至 2025-12-04 共收录 5 信号源:cs.CV, cs.GR, cs.MM

1. 效率与蒸馏 5 篇

2512.04039 2025-12-04 cs.CV cs.AI cs.LG 70%

Fast & Efficient Normalizing Flows and Applications of Image Generative Models

高效且快速的归一化流及图像生成模型的应用

Sandeep Nagar

机构 * International Institute of Information Technology (Deemed to be University)(国际信息技术学院(认定为大学)) IIIT Hyderabad(IIIT海得拉尔)

专题命中 效率与蒸馏 :diffusion(abstract);inpainting(abstract);分类 cs.CV

AI总结 本研究提出高效归一化流及图像生成模型的应用,包括超分辨率、农业质量评估、地质制图、隐私保护及艺术修复等

Comments PhD Thesis

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04040 2025-12-04 cs.CV 57%

RELIC: Interactive Video World Model with Long-Horizon Memory

RELIC:具有长时间记忆的交互式视频世界模型

Yicong Hong, Yiqun Mei, Chongjian Ge, Yiran Xu, Yang Zhou, Sai Bi, Yannick Hold-Geoffroy, Mike Roberts, Matthew Fisher, Eli Shechtman, Kalyan Sunkavalli, Feng Liu, Zhengqi Li, Hao Tan

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

AI总结 RELIC通过统一框架实现长时记忆与实时交互,提升视频世界建模的准确性与稳定性。

Comments 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03796 2025-12-04 cs.CV 57%

LSRS: Latent Scale Rejection Sampling for Visual Autoregressive Modeling

LSRS: 隐式尺度拒绝采样用于视觉自回归建模

Hong-Kai Zheng, Piji Li

机构 * College of Artificial Intelligence, Nanjing University of Aeronautics and Astronautics, China(南京航空航天大学人工智能学院) MIIT Key Laboratory of Pattern Analysis and Machine Intelligence, Nanjing, China(信息产业部模式分析与机器智能重点实验室) The Key Laboratory of Brain-Machine Intelligence Technology, Ministry of Education, Nanjing, China(教育部脑机智能技术重点实验室)

专题命中 效率与蒸馏 :image generation(abstract);分类 cs.CV

AI总结 LSRS通过隐式尺度拒绝采样方法,在保持计算效率的同时提升视觉自回归模型的生成质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08136 2025-12-04 cs.CV 57%

FantasyStyle: Controllable Stylized Distillation for 3D Gaussian Splatting

FantasyStyle: 用于3D高斯点云的可控风格化蒸馏

Yitong Yang, Yinglin Wang, Changshuo Wang, Huajie Wang, Shuting He

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

AI总结 FantasyStyle通过可控风格化蒸馏和多视图频率一致性,解决3DGS风格迁移中的多视图不一致和内容泄漏问题,提升风格化质量与视觉真实感。

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03125 2025-12-04 cs.LG cs.AI 50%

Mitigating Intra- and Inter-modal Forgetting in Continual Learning of Unified Multimodal Models

缓解统一多模态模型持续学习中的模态内和模态间遗忘

Xiwen Wei, Mustafa Munir, Radu Marculescu

专题命中 效率与蒸馏 :image generation(abstract)

AI总结 本文提出MoDE,通过解耦模态以缓解统一多模态模型中的模态内和模态间遗忘问题。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏