arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-02-02 至 2026-02-02 共收录 3 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3 篇

2509.25562 2026-02-02 cs.AI cs.CL cs.CV cs.LG 85%

IRIS: Intrinsic Reward Image Synthesis

IRIS:内在奖励图像合成

Yihang Chen, Yuanhao Ban, Yunqi Hong, Cho-Jui Hsieh

机构 * Department of Computer Science, University of California, Los Angeles, USA(计算机科学系,加州大学洛杉矶分校)

专题命中 文生图 :image synthesis(title,abstract);image generation(abstract);text-to-image(abstract);分类 cs.CV

AI总结 IRIS通过内在奖励提升自回归T2I模型性能,改进图像生成质量并促进细致推理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22630 2026-02-02 cs.CV 70%

LINA: Linear Autoregressive Image Generative Models with Continuous Tokens

LINA: 基于连续令牌的线性自回归图像生成模型

Jiahao Wang, Ting Pan, Haoge Deng, Dongchen Han, Taiqiang Wu, Xinlong Wang, Ping Luo

机构 * The University of Hong Kong(香港大学) University of Chinese Academy of Sciences(中国科学院大学) Tsinghua University(清华大学) Beijing Academy of Artificial Intelligence(北京人工智能研究院)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

AI总结 LINA是一种基于线性注意力的高效文本到图像生成模型,通过改进归一化和门控机制,在计算效率和生成质量上取得平衡。

Comments 20 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02725 2026-02-02 cs.AI 67%

Advances in Artificial Intelligence: A Review for the Creative Industries

人工智能进展:面向创意产业的综述

Nantheera Anantrasirichai, Fan Zhang, David Bull

机构 * Visual Information Laboratory, University of Bristol, Bristol, UK(布里斯托大学视觉信息实验室)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract)

AI总结 本文综述了自2022年以来人工智能在创意产业中的进展,探讨了生成式AI、大语言模型和扩散模型等技术对创意生产流程的影响,并分析了人类与AI协作的新趋势及面临的挑战。

Comments This is an updated review of our previous paper (see https://doi.org/10.1007/s10462-021-10039-7), and has been accepted by Artificial Intelligence Review journal

详情

展开后加载摘要…

URL PDF HTML 收藏