arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 70277 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70277 篇

2503.03708 2025-03-28 cs.CV cs.AI 79%

Rethinking Video Tokenization: A Conditioned Diffusion-based Approach

Nianzu Yang, Pandeng Li, Liming Zhao, Yang Li, Chen-Wei Xie, Yehui Tang, Xudong Lu, Zhihang Liu, Yun Zheng, Yu Liu, Junchi Yan

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07776 2025-03-28 cs.CV cs.AI cs.LG 79%

Video Motion Transfer with Diffusion Transformers

Alexander Pondaven, Aliaksandr Siarohin, Sergey Tulyakov, Philip Torr, Fabio Pizzati

机构 * University of Oxford(牛津大学) Snap Inc.(Snap公司) MBZUAI(穆罕默德·本·扎耶德人工智能大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025 - Project page: https://ditflow.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03044 2025-03-28 cs.CV 79%

Frequency-Guided Diffusion Model with Perturbation Training for Skeleton-Based Video Anomaly Detection

Xiaofeng Tan, Hongsong Wang, Xin Geng, Liang Wang

机构 * School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院) Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(教育部新一代人工智能技术及其交叉应用重点实验室(东南大学)) State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所多模态人工智能系统国家重点实验室(MAIS)) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21144 2025-03-28 cs.CV 79%

ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model

Jinwei Qi, Chaonan Ji, Sheng Xu, Peng Zhang, Bang Zhang, Liefeng Bo

机构 * Tongyi Lab, Alibaba Group(阿里巴巴集团通义实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://humanaigc.github.io/chat-anyone/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21082 2025-03-28 cs.CV 79%

Can Video Diffusion Model Reconstruct 4D Geometry?

Jinjie Mai, Wenxuan Zhu, Haozhe Liu, Bing Li, Cheng Zheng, Jürgen Schmidhuber, Bernard Ghanem

机构 * King Abdullah University of Science and Technology (KAUST)(阿卜杜拉国王科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.07421 2025-03-28 cs.LG cs.CV 79%

U-Turn Diffusion

Hamidreza Behjoo, Michael Chertkov

机构 * University of Arizona(亚利桑那大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Entropy 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20571 2025-03-27 eess.IV cs.CV 79%

Exploring Robustness of Cortical Morphometry in the presence of white matter lesions, using Diffusion Models for Lesion Filling

Vinzenz Uhr, Ivan Diaz, Christian Rummel, Richard McKinley

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20483 2025-03-27 cs.CV cs.LG 79%

Dissecting and Mitigating Diffusion Bias via Mechanistic Interpretability

Yingdong Shi, Changming Li, Yifan Wang, Yongxiang Zhao, Anqi Pang, Sibei Yang, Jingyi Yu, Kan Ren

机构 * ShanghaiTech University(上海科技大学) Stony Brook University(石溪大学) Tencent PCG(腾讯互动娱乐事业群)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025; Project Page: https://foundation-model-research.github.io/difflens

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20268 2025-03-27 cs.CV 79%

EGVD: Event-Guided Video Diffusion Model for Physically Realistic Large-Motion Frame Interpolation

Ziran Zhang, Xiaohui Li, Yihao Liu, Yujin Wang, Yueting Chen, Tianfan Xue, Shi Guo

机构 * Zhejiang University(浙江大学) Shanghai AI Laboratory(上海人工智能实验室) Shanghai Jiao Tong University(上海交通大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19357 2025-03-27 cs.CV 79%

Correcting Deviations from Normality: A Reformulated Diffusion Model for Multi-Class Unsupervised Anomaly Detection

Farzad Beizaee, Gregory A. Lodygensky, Christian Desrosiers, Jose Dolz

机构 * ÉTS Montreal(蒙特利尔高等技术学院) CHU-Sainte-Justine(圣茹斯特ine医院中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16302 2025-03-27 cs.CV cs.AI eess.IV 79%

Unleashing Vecset Diffusion Model for Fast Shape Generation

Zeqiang Lai, Yunfei Zhao, Zibo Zhao, Haolin Liu, Fuyun Wang, Huiwen Shi, Xianghui Yang, Qingxiang Lin, Jingwei Huang, Yuhong Liu, Jie Jiang, Chunchao Guo, Xiangyu Yue

机构 * MMLab, CUHK(香港中文大学MMLab) Tencent Hunyuan(腾讯混元) VISG, NJU(南京大学VISG) ShanghaiTech(上海科技大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Technical report

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13726 2025-03-27 cs.CV cs.AI 79%

DAWN: Dynamic Frame Avatar with Non-autoregressive Diffusion Framework for Talking Head Video Generation

Hanbo Cheng, Limin Lin, Chenyu Liu, Pengcheng Xia, Pengfei Hu, Jiefeng Ma, Jun Du, Jia Pan

机构 * University of Science and Technology of China(中国科学技术大学) iFLYTEK Research(讯飞研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06072 2025-03-27 cs.CV 79%

CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Zhuoyi Yang, Jiayan Teng, Wendi Zheng, Ming Ding, Shiyu Huang, Jiazheng Xu, Yuanming Yang, Wenyi Hong, Xiaohan Zhang, Guanyu Feng, Da Yin, Yuxuan Zhang, Weihan Wang, Yean Cheng, Bin Xu, Xiaotao Gu, Yuxiao Dong, Jie Tang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by ICLR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06218 2025-03-27 cs.CV cs.AI cs.SE 79%

Data Augmentation in Earth Observation: A Diffusion Model Approach

Tiago Sousa, Benoît Ries, Nicolas Guelfi

机构 * University of Luxembourg(卢森堡大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 25 pages, 12 figures

Journal ref Information 2025, 16, 81

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19881 2025-03-26 cs.CV 79%

Mask$^2$DiT: Dual Mask-based Diffusion Transformer for Multi-Scene Long Video Generation

Tianhao Qi, Jianlong Yuan, Wanquan Feng, Shancheng Fang, Jiawei Liu, SiYu Zhou, Qian He, Hongtao Xie, Yongdong Zhang

机构 * University of Science and Technology of China(中国科学技术大学) Bytedance Intelligent Creation(字节跳动智能创作) Yuanshi Inc.(元视科技公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07761 2025-03-26 cs.CV 79%

Repurposing Pre-trained Video Diffusion Models for Event-based Video Interpolation

Jingxi Chen, Brandon Y. Feng, Haoming Cai, Tianfu Wang, Levi Burner, Dehao Yuan, Cornelia Fermuller, Christopher A. Metzler, Yiannis Aloimonos

机构 * University of Maryland, College Park(马里兰大学帕克分校) Massachusetts Institute of Technology(麻省理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19798 2025-03-26 cs.CV eess.IV 79%

Unpaired Object-Level SAR-to-Optical Image Translation for Aircraft with Keypoints-Guided Diffusion Models

Ruixi You, Hecheng Jia, Feng Xu

机构 * School of Information Science and Technology, Fudan University(复旦大学信息科学与技术学院) Key Laboratory for Information Science of Electromagnetic Waves (Ministry of Education)(电磁波信息科学教育部重点实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19462 2025-03-26 cs.CV 79%

AccVideo: Accelerating Video Diffusion Model with Synthetic Dataset

Haiyu Zhang, Xinyuan Chen, Yaohui Wang, Xihui Liu, Yunhong Wang, Yu Qiao

机构 * Beihang University(北京航空航天大学) Shanghai AI Laboratory(上海人工智能实验室) The University of Hong Kong(香港大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project Page: https://aejion.github.io/accvideo/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19448 2025-03-26 cs.CV eess.IV 79%

Towards Robust Time-of-Flight Depth Denoising with Confidence-Aware Diffusion Model

Changyong He, Jin Zeng, Jiawei Zhang, Jiajie Guo

机构 * SenseTime Research(商汤科技研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19429 2025-03-26 cs.LG cs.CV 79%

Quantifying the Ease of Reproducing Training Data in Unconditional Diffusion Models

Masaya Hasegawa, Koji Yasuda

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19340 2025-03-26 cs.CV 79%

BADGR: Bundle Adjustment Diffusion Conditioned by GRadients for Wide-Baseline Floor Plan Reconstruction

Yuguang Li, Ivaylo Boyadzhiev, Zixuan Liu, Linda Shapiro, Alex Colburn

机构 * University of Washington(华盛顿大学) Zillow Group(齐洛集团)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19283 2025-03-26 cs.CV 79%

ISPDiffuser: Learning RAW-to-sRGB Mappings with Texture-Aware Diffusion Models and Histogram-Guided Color Consistency

Yang Ren, Hai Jiang, Menglong Yang, Wei Li, Shuaicheng Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19001 2025-03-26 cs.CV cs.AI 79%

DisentTalk: Cross-lingual Talking Face Generation via Semantic Disentangled Diffusion Model

Kangwei Liu, Junwu Liu, Yun Cao, Jinlin Guo, Xiaowei Yi

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Journal ref Accpeted by ICME 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16942 2025-03-26 cs.CV 79%

Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model

Yingying Fan, Quanwei Yang, Kaisiyuan Wang, Hang Zhou, Yingying Li, Haocheng Feng, Errui Ding, Yu Wu, Jingdong Wang

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) University of Science and Technology of China(中国科学技术大学) Baidu Inc.(百度公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16396 2025-03-26 cs.CV 79%

SV4D 2.0: Enhancing Spatio-Temporal Consistency in Multi-View Video Diffusion for High-Quality 4D Generation

Chun-Han Yao, Yiming Xie, Vikram Voleti, Huaizu Jiang, Varun Jampani

机构 * Stability AI Northeastern University(东北大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Project page: https://sv4d20.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15851 2025-03-26 cs.CV 79%

Zero-1-to-A: Zero-Shot One Image to Animatable Head Avatars Using Video Diffusion

Zhenglin Zhou, Fan Ma, Hehe Fan, Tat-Seng Chua

机构 * Zhejiang University(浙江大学) National University of Singapore(新加坡国立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by CVPR 2025, project page: https://zhenglinzhou.github.io/Zero-1-to-A/

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19694 2025-03-26 cs.CV cs.AI cs.LG 79%

BEVDiffuser: Plug-and-Play Diffusion Model for BEV Denoising with Ground-Truth Guidance

Xin Ye, Burhaneddin Yaman, Sheng Cheng, Feng Tao, Abhirup Mallik, Liu Ren

机构 * Bosch Research North America(博世北美研究院) Bosch Center for Artificial Intelligence (BCAI)(博世人工智能中心)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08244 2025-03-26 cs.CV 79%

FloVD: Optical Flow Meets Video Diffusion Model for Enhanced Camera-Controlled Video Synthesis

Wonjoon Jin, Qi Dai, Chong Luo, Seung-Hwan Baek, Sunghyun Cho

机构 * POSTECH(浦项科技大学) Microsoft Research Asia(微软亚洲研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Our paper has been accepted to CVPR 2025. Website: https://jinwonjoon.github.io/flovd_site/ Code: https://github.com/JinWonjoon/FloVD

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06011 2025-03-26 eess.IV cs.CV 79%

TopoCellGen: Generating Histopathology Cell Topology with a Diffusion Model

Meilong Xu, Saumya Gupta, Xiaoling Hu, Chen Li, Shahira Abousamra, Dimitris Samaras, Prateek Prasanna, Chao Chen

机构 * Stony Brook University(石溪大学) Massachusetts General Hospital(麻省总医院) Harvard Medical School(哈佛医学院) Stanford University(斯坦福大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by CVPR 2025. 15 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14103 2025-03-26 cs.CV cs.AI 79%

Extreme Precipitation Nowcasting using Multi-Task Latent Diffusion Models

Li Chaorong, Ling Xudong, Yang Qiang, Qin Fengqing, Huang Yuanyuan

机构 * School of Computer Science and Technology, Yibin University(宜宾学院计算机科学与技术学院) School of Artificial Intelligence, Chengdu University of Information Technology(成都信息工程大学人工智能学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 15 pages, 14figures

详情

展开后加载摘要…

URL PDF HTML 收藏