arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-17 至 2025-10-17 共收录 8 信号源:cs.CV, cs.GR, cs.MM

1. 可控生成 8 篇

2509.18092 2025-10-17 cs.CV 85%

ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation

Guocheng Gordon Qian, Daniil Ostashev, Egor Nemchinov, Avihay Assouline, Sergey Tulyakov, Kuan-Chieh Jackson Wang, Kfir Aberman

机构 * Snap Inc.

专题命中 可控生成 :image generation(title);text-to-image(abstract);diffusion(abstract);image synthesis(abstract)

Comments Accepted to SIGGRAPH Asia 2025, webpage: https://snap-research.github.io/composeme/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14975 2025-10-17 cs.CV cs.AI 83%

WithAnyone: Towards Controllable and ID Consistent Image Generation

Hengyuan Xu, Wei Cheng, Peng Xing, Yixiao Fang, Shuhan Wu, Rui Wang, Xianfang Zeng, Daxin Jiang, Gang Yu, Xingjun Ma, Yu-Gang Jiang

机构 * Fudan University(复旦大学) StepFun Project(StepFun项目) MultiID-2M MultiID-Bench

专题命中 可控生成 :image generation(title);text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments 23 Pages; Project Page: https://doby-xu.github.io/WithAnyone/; Code: https://github.com/Doby-Xu/WithAnyone

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14179 2025-10-17 cs.CV cs.AI 79%

Virtually Being: Customizing Camera-Controllable Video Diffusion Models with Multi-View Performance Captures

Yuancheng Xu, Wenqi Xian, Li Ma, Julien Philip, Ahmet Levent Taşel, Yiwei Zhao, Ryan Burgert, Mingming He, Oliver Hermann, Oliver Pilarski, Rahul Garg, Paul Debevec, Ning Yu

机构 * Eyeline Labs(Eyeline实验室)

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to SIGGRAPH Asia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14882 2025-10-17 cs.CV 77%

ScaleWeaver: Weaving Efficient Controllable T2I Generation with Multi-Scale Reference Attention

Keli Liu, Zhendong Wang, Wengang Zhou, Shaodong Xu, Ruixiao Dong, Houqiang Li

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 可控生成 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14976 2025-10-17 cs.CV cs.GR cs.RO 62%

Ponimator: Unfolding Interactive Pose for Versatile Human-human Interaction Animation

Shaowei Liu, Chuan Guo, Bing Zhou, Jian Wang

专题命中 可控生成 :diffusion(abstract);分类 cs.CV、cs.GR

Comments Accepted to ICCV 2025. Project page: https://stevenlsw.github.io/ponimator/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14874 2025-10-17 cs.CV 57%

TOUCH: Text-guided Controllable Generation of Free-Form Hand-Object Interactions

Guangyi Han, Wei Zhai, Yuhang Yang, Yang Cao, Zheng-Jun Zha

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14536 2025-10-17 cs.CV 57%

Exploring Image Representation with Decoupled Classical Visual Descriptors

Chenyuan Qu, Hao Chen, Jianbo Jiao

机构 * The MIx Group University of Birmingham, UK(米克斯小组英国伯明翰大学) University of Cambridge Cambridge, UK(剑桥大学) Allsee Technologies Ltd Birmingham, UK(Allsee技术有限公司)

专题命中 可控生成 :image generation(abstract);分类 cs.CV

Comments Accepted by The 36th British Machine Vision Conference (BMVC 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15195 2025-10-17 math.PR 50%

Control of Conditional Processes and Fleming--Viot Dynamics

Philipp Jettkant

专题命中 可控生成 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏