arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

视频大模型

视频理解、视频生成、视频语言模型和时序视觉推理。

2025-10-21 至 2025-10-21 共收录 3 信号源:cs.CV, eess.IV, cs.MM

1. 视频生成 3 篇

2510.16833 2025-10-21 cs.CV cs.GR 79%

From Mannequin to Human: A Pose-Aware and Identity-Preserving Video Generation Framework for Lifelike Clothing Display

Xiangyu Mu, Dongliang Zhou, Jie Hou, Haijun Zhang, Weili Guan

机构 * Harbin Institute of Technology(哈尔滨工业大学)

专题命中 视频生成 :video generation(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14588 2025-10-21 cs.CV cs.AI 79%

STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding

Zhifei Chen, Tianshuo Xu, Leyi Wu, Luozhou Wang, Dongyu Yan, Zihan You, Wenting Luo, Guo Zhang, Yingcong Chen

机构 * HKUST(GZ)(香港科技大学(珠海)) HKUST(香港科技大学) XMU(厦门大学) MIT(麻省理工学院)

专题命中 视频生成 :video generation(title,abstract);分类 cs.CV

Comments Code, model, and demos can be found at https://envision-research.github.io/STANCE/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17041 2025-10-21 cs.CV cs.AI cs.LG 79%

Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM

Jaemin Kim, Bryan Sangwoo Kim, Jong Chul Ye

机构 * Graduate School of AI, KAIST(人工智能研究生院,韩国科学技术院)

专题命中 视频生成 :text-to-video(title,abstract);分类 cs.CV

Comments ICCV 2025 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏