arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-28 至 2025-10-28 共收录 6 信号源:cs.CV, cs.GR, cs.MM

1. 个性化与一致性 6 篇

2510.22994 2025-10-28 cs.CV 70%

SceneDecorator: Towards Scene-Oriented Story Generation with Scene Planning and Scene Consistency

Quanjian Song, Donghao Zhou, Jingyu Lin, Fei Shen, Jiaze Wang, Xiaowei Hu, Cunjian Chen, Pheng-Ann Heng

机构 * Monash University(莫纳什大学) The Chinese University of Hong Kong(香港中文大学) National University of Singapore(新加坡国立大学) South China University of Technology(华南理工大学)

专题命中 个性化与一致性 :image generation(abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted by NeurIPS 2025; Project Page: https://lulupig12138.github.io/SceneDecorator

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23605 2025-10-28 cs.CV cs.AI cs.GR cs.LG cs.RO 62%

Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling

Shuhong Zheng, Ashkan Mirzaei, Igor Gilitschenski

机构 * University of Toronto(多伦多大学) Vector Institute(向量研究所) Snap Inc.(Snap公司)

专题命中 个性化与一致性 :inpainting(abstract);分类 cs.CV、cs.GR

Comments NeurIPS 2025, 38 pages, 22 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22810 2025-10-28 cs.CV 57%

MAGIC-Talk: Motion-aware Audio-Driven Talking Face Generation with Customizable Identity Control

Fatemeh Nazarieh, Zhenhua Feng, Diptesh Kanojia, Muhammad Awais, Josef Kittler

机构 * School of Computer Science and Electronic Engineering, University of Surrey(Surrey大学计算机科学与电子工程学院) School of Artificial Intelligence and Computer Science, Jiangnan University(江南大学人工智能与计算机科学学院) Centre for Vision, Speech and Signal Processing, University of Surrey(Surrey大学视觉、语音和信号处理中心)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04705 2025-10-28 cs.CV 57%

Identity-Preserving Text-to-Video Generation Guided by Simple yet Effective Spatial-Temporal Decoupled Representations

Yuji Wang, Moran Li, Xiaobin Hu, Ran Yi, Jiangning Zhang, Han Feng, Weijian Cao, Yabiao Wang, Chengjie Wang, Lizhuang Ma

机构 * Shanghai Jiao Tong University, Tencent Youtu Lab(上海交通大学,腾讯云图实验室) Tencent Youtu Lab(腾讯云图实验室) Shanghai Jiao Tong University(上海交通大学) Tencent(腾讯)

专题命中 个性化与一致性 :text-to-image(abstract);分类 cs.CV

Comments ACM Multimedia 2025; code URL: https://github.com/rain152/IPVG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21857 2025-10-28 cs.CV cs.AI 57%

Poisson Flow Consistency Training

Anthony Zhang, Mahmut Gokmen, Dennis Hein, Rongjun Ge, Wenjun Xia, Ge Wang, Jin Chen

机构 * Pratt School of Engineering Duke University(普拉特工程学院 哥伦比亚大学) Biomedical Imaging Center Rensselaer Polytechnic Institute(生物医学成像中心 莱文森理工学院) Department of Computer Science University of Kentucky(计算机科学系 肯塔基大学) Department of Medicine Department of Biomedical Informations University of Alabama at Birmingham(医学系 生物医学信息学系 亚拉巴马大学伯明翰分校) School of Engineering Sciences Department of Physics KTH Royal Institute of Technology(工程科学学院 物理系 瑞典皇家理工学院)

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

Comments 5 pages, 3 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23106 2025-10-28 cs.LG 50%

Sampling from Energy distributions with Target Concrete Score Identity

Sergei Kholkin, Francisco Vargas, Alexander Korotin

机构 * Applied AI Institute(应用人工智能研究所)

专题命中 个性化与一致性 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏