arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-11-27 至 2025-11-27 共收录 5 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 5 篇

2406.13933 2025-11-27 cs.CR 88%

EnTruth: Enhancing the Traceability of Unauthorized Dataset Usage in Text-to-image Diffusion Models with Minimal and Robust Alterations

EnTruth:通过最小和稳健的修改增强文本到图像扩散模型中未经授权数据集使用的可追溯性

Jie Ren, Yingqian Cui, Chen Chen, Yue Xing, Hui Liu, Lingjuan Lyu

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract)

AI总结 EnTruth通过模板记忆技术,提升文本到图像扩散模型中未经授权数据集使用的可追溯性,实现版权保护与生成模型检测的创新应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21185 2025-11-27 cs.CV cs.AI 85%

Progress by Pieces: Test-Time Scaling for Autoregressive Image Generation

分块进展:用于自回归图像生成的测试时扩展

Joonhyung Park, Hyeongwon Jang, Joowon Kim, Eunho Yang

专题命中 文生图 :image generation(title,abstract);text-to-image(abstract);image editing(abstract);分类 cs.CV

AI总结 GridAR通过网格分区逐步生成和布局指定提示重述策略,在有限测试时扩展下提升视觉自回归模型的生成质量与编辑效果。

Comments Project page: https://grid-ar.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21530 2025-11-27 cs.CV 70%

The Age-specific Alzheimer 's Disease Prediction with Characteristic Constraints in Nonuniform Time Span

具有非均匀时间跨度的特征约束的年龄特异性阿尔茨海默病预测

Xin Hong, Kaifeng Huang

机构 * Department of Software, School of Computer Science and Technology, Huaqiao University, China(软件系,计算机科学与技术学院,华侨大学)

专题命中 文生图 :image generation(abstract);image synthesis(abstract);分类 cs.CV

AI总结 本研究提出了一种基于定量指标和年龄缩放因子的序列图像生成方法,用于提高非均匀时间跨度下阿尔茨海默病预测的准确性。

Comments 16 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04678 2025-11-27 cs.CV 70%

Unsupervised Segmentation by Diffusing, Walking and Cutting

无监督分割的扩散、行走与切割

Daniela Ivanova, Marco Aversa, Paul Henderson, John Williamson

机构 * University of Glasgow, UK(格拉斯哥大学)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

AI总结 基于预训练扩散模型的自注意力特征,提出无监督图像分割方法,通过递归归一化切割和随机行走实现语义层次分割,无需额外训练。

Comments Accepted to The IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21547 2025-11-27 cs.HC 50%

Seeing Twice: How Side-by-Side T2I Comparison Changes Auditing Strategies

双倍观察:如何通过并列的T2I比较改变审计策略

Matheus Kunzler Maldaner, Wesley Hanwen Deng, Jason I. Hong, Kenneth Holstein, Motahhare Eslami

专题命中 文生图 :text-to-image(abstract)

AI总结 MIRAGE工具通过并列展示多个文本到图像模型,帮助用户发现生成模型的偏见,改变审计策略。

Comments 8 pages, 6 figures. Presented at ACM Collective Intelligence (CI), 2025. Available at https://ci.acm.org/2025/wp-content/uploads/101-Maldaner.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏