arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-08-26 至 2025-08-26 共收录 101 信号源:cs.CV, cs.GR, cs.MM

1. 图像修复 1 篇

2508.17708 2025-08-26 cs.CV 57%

CATformer: Contrastive Adversarial Transformer for Image Super-Resolution

Qinyi Tian, Spence Cox, Laura E. Dalton

专题命中 图像修复 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 个性化与一致性 3 篇

2508.16863 2025-08-26 cs.CV 83%

Delta-SVD: Efficient Compression for Personalized Text-to-Image Models

Tangyuan Zhang, Shangyu Chen, Qixiang Chen, Jianfei Cai

机构 * Monash University(墨尔本大学)

专题命中 个性化与一致性 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16696 2025-08-26 cs.GR cs.AI 57%

DecoMind: A Generative AI System for Personalized Interior Design Layouts

Reema Alshehri, Rawan Alotaibi, Leen Almasri, Rawan Altaweel

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.GR

Comments ~7 pages; ~32 figures; compiled with pdfLaTeX. Primary category: cs.CV. (Secondary: cs.AI)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18118 2025-08-26 cs.IR cs.CL 50%

HLLM-Creator: Hierarchical LLM-based Personalized Creative Generation

Junyi Chen, Lu Chi, Siliang Xu, Shiwei Ran, Bingyue Peng, Zehuan Yuan

机构 * ByteDance(字节跳动)

专题命中 个性化与一致性 :personalized generation(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 效率与蒸馏 5 篇

2508.17868 2025-08-26 cs.SD cs.AI cs.LG eess.AS stat.ML 78%

FasterVoiceGrad: Faster One-step Diffusion-Based Voice Conversion with Adversarial Diffusion Conversion Distillation

Takuhiro Kaneko, Hirokazu Kameoka, Kou Tanaka, Yuto Kondo

机构 * NTT, Inc.(日本NTT公司)

专题命中 效率与蒸馏 :diffusion(title,abstract)

Comments Accepted to Interspeech 2025. Project page: https://www.kecl.ntt.co.jp/people/kaneko.takuhiro/projects/fastervoicegrad/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16939 2025-08-26 cs.LG math.PR stat.ML 78%

Sig-DEG for Distillation: Making Diffusion Models Faster and Lighter

Lei Jiang, Wen Ge, Niels Cariou-Kotlarek, Mingxuan Yi, Po-Yu Chen, Lingyi Yang, Francois Buet-Golfouse, Gaurav Mittal, Hao Ni

机构 * University College London(伦敦大学学院) JPMorganChase(摩根大通) University of Oxford(牛津大学) AIML Global Markets, Barclays(巴克莱证券AIML全球市场)

专题命中 效率与蒸馏 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17426 2025-08-26 cs.LG 67%

Modular MeanFlow: Towards Stable and Scalable One-Step Generative Modeling

Haochen You, Baojing Liu, Hongyang He

机构 * Graduate School of Arts and Sciences, Columbia University, New York, USA(哥伦比亚大学艺术与科学研究生院) Department of Computer Science, University of Warwick, Coventry, UK(沃里克大学计算机科学系)

专题命中 效率与蒸馏 :diffusion(abstract);image synthesis(abstract)

Comments Accepted as a conference paper at PRCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15420 2025-08-26 cs.CV cs.AI cs.LG 57%

Visual Generation Without Guidance

Huayu Chen, Kai Jiang, Kaiwen Zheng, Jianfei Chen, Hang Su, Jun Zhu

机构 * Department of Computer Science \& Technology, Tsinghua University

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

Comments Accepted to ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16917 2025-08-26 cs.CV 57%

Structural Energy-Guided Sampling for View-Consistent Text-to-3D

Qing Zhang, Jinguang Tong, Jie Hong, Jing Zhang, Xuesong Li

机构 * The Australian National University(澳大利亚国立大学) CSIRO(澳大利亚联邦科学与工业研究组织) The University of Hong Kong(香港大学)

专题命中 效率与蒸馏 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 其他图像生成 2 篇

2508.17199 2025-08-26 cs.CV 79%

MMCIG: Multimodal Cover Image Generation for Text-only Documents and Its Dataset Construction via Pseudo-labeling

Hyeyeon Kim, Sungwoo Han, Jingun Kwon, Hidetaka Kamigaito, Manabu Okumura

机构 * Chungnam National University(Chungnam 国立大学) Nara Institute of Science and Technology (NAIST)(Nara 科学技术研究所) Institute of Science Tokyo(东京科学研究所)

专题命中 其他图像生成 :image generation(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.11804 2025-08-26 cs.CR cs.AI cs.LG 50%

Adversarial Illusions in Multi-Modal Embeddings

Tingwei Zhang, Rishi Jha, Eugene Bagdasaryan, Vitaly Shmatikov

机构 * Cornell University(康奈尔大学) University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校) Cornell Tech(康奈尔科技)

专题命中 其他图像生成 :image generation(abstract)

Comments In USENIX Security'24

详情

展开后加载摘要…

URL PDF HTML 收藏