arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-02-13 至 2026-02-13 共收录 60 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3 篇

2602.12004 2026-02-13 cs.AI 88%

CSEval: A Framework for Evaluating Clinical Semantics in Text-to-Image Generation

CSEval: 一个用于评估文本到图像生成中临床语义的框架

Robert Cronshaw, Konstantinos Vilouras, Junyu Yan, Yuning Du, Feng Chen, Steven McDonagh, Sotirios A. Tsaftaris

机构 * School of Engineering, University of Edinburgh, United Kingdom(爱丁堡大学工程学院)

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)

AI总结 CSEval通过语言模型评估文本到图像生成中临床语义的一致性,弥补现有方法在临床相关性评估上的不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19578 2026-02-13 cs.CV 83%

Multiscale Vector-Quantized Variational Autoencoder for Endoscopic Image Synthesis

多尺度向量量化变分自编码器用于内窥镜图像合成

Dimitrios E. Diamantis, Dimitris K. Iakovidis

专题命中 文生图 :image synthesis(title,abstract);image generation(abstract);分类 cs.CV

AI总结 本文提出多尺度向量量化变分自编码器,用于生成内窥镜图像中的异常,提升临床决策支持系统的性能。

Journal ref Proc. IEEE International Conference on Imaging Systems and Techniques (IST 2025), Strasburg, France

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11980 2026-02-13 cs.CV 81%

Spatial Chain-of-Thought: Bridging Understanding and Generation Models for Spatial Reasoning Generation

空间链式思考:连接理解与生成模型以实现空间推理生成

Wei Chen, Yancheng Long, Mingqiao Liu, Haojie Ding, Yankai Yang, Hongyang Wei, Yi-Fan Zhang, Bin Wen, Fan Yang, Tingting Gao, Han Li, Long Chen

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Tsinghua University(清华大学) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)

专题命中 文生图 :image generation(abstract);diffusion(abstract);image editing(abstract);image synthesis(abstract)

AI总结 本文提出SCoT框架,通过结合MLLMs的推理能力与扩散模型的生成能力,提升空间推理生成任务的性能。

Comments 19 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 扩散模型 45 篇

2602.11401 2026-02-13 cs.CV cs.LG 86%

Latent Forcing: Reordering the Diffusion Trajectory for Pixel-Space Image Generation

潜在强制:重新排列扩散轨迹以生成像素空间图像

Alan Baade, Eric Ryan Chan, Kyle Sargent, Changan Chen, Justin Johnson, Ehsan Adeli, Li Fei-Fei

机构 * stanford(斯坦福大学) umich(密歇根大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(title);分类 cs.CV

AI总结 本文提出潜在强制方法,通过重新排列扩散轨迹提升像素空间图像生成效率,实现端到端建模优势,取得ImageNet上的新突破。

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11875 2026-02-13 cs.CV cs.RO 85%

DiffPlace: Street View Generation via Place-Controllable Diffusion Model Enhancing Place Recognition

DiffPlace: 通过可控制位置的扩散模型增强位置识别生成街道视图

Ji Li, Zhiwei Li, Shihao Li, Zhenjiang Yu, Boyang Wang, Haiou Liu

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);image synthesis(abstract);分类 cs.CV

AI总结 DiffPlace通过引入位置ID控制器,实现了位置可控的多视角图像生成,提升了视觉位置识别任务中的生成质量和训练支持。

Comments accepted by ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22650 2026-02-13 cs.CV 83%

Self-Attention Decomposition For Training Free Diffusion Editing

自注意力分解用于训练自由扩散编辑

Tharun Anand, Mohammad Hassan Vali, Arno Solin, Green Rosh, BH Pawan Prasad

机构 * ELLIS Institute Finland(芬兰ELLIS研究所) Department of Computer Science, Aalto University, Finland(芬兰奥卢大学计算机科学系)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

AI总结 本文提出了一种基于自注意力权重矩阵特征向量提取的扩散模型编辑方法,无需额外数据或微调,实现高效高质量的图像编辑。

Comments ICASSP 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.21088 2026-02-13 cs.LG cs.CR cs.CV 83%

Shallow Diffuse: Robust and Invisible Watermarking through Low-Dimensional Subspaces in Diffusion Models

浅层扩散:通过扩散模型中低维子空间实现的鲁棒且不可见的水印技术

Wenda Li, Huijie Zhang, Qing Qu

机构 * University of Michigan(密歇根大学)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

AI总结 本文提出Shallow Diffuse,一种通过扩散模型中低维子空间实现鲁棒且不可见水印的技术,提升了水印的检测性和数据生成的一致性。

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17770 2026-02-13 eess.IV cs.CV 83%

CT Synthesis with Conditional Diffusion Models for Abdominal Lymph Node Segmentation

基于条件扩散模型的CT合成用于腹部淋巴结分割

Yongrui Yu, Hanyu Chen, Zitian Zhang, Qiong Xiao, Wenhui Lei, Linrui Dai, Yu Fu, Hui Tan, Guan Wang, Peng Gao, Xiaofan Zhang

机构 * Shanghai Jiao Tong University, Shanghai, China(上海交通大学) Department of Surgical Oncology and General Surgery, Key Laboratory of Precision Diagnosis and Treatment of Gastrointestinal Tumors, Ministry of Education, The First Hospital of China Medical University, Shenyang, China(外科肿瘤科和普通外科,精准诊断与治疗胃肠肿瘤国家重点实验室,教育部,中国医科大学第一附属医院,沈阳,中国) Department of Radiology, The First Hospital of China Medical University, Shenyang, China(放射科,中国医科大学第一附属医院,沈阳,中国) Shanghai AI Laboratory, Shanghai, China(上海人工智能实验室,上海,中国)

专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV

AI总结 本文提出LN-DDPM模型,通过条件扩散模型生成腹部淋巴结数据,提升分割性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11704 2026-02-13 eess.IV cs.CV 79%

U-DAVI: Uncertainty-Aware Diffusion-Prior-Based Amortized Variational Inference for Image Reconstruction

U-DAVI:基于不确定性意识的扩散先验的图像重建近似变分推断

Ayush Varshney, Katherine L. Bouman, Berthy T. Feng

机构 * Computing and Mathematical Sciences, California Institute of Technology(计算与数学科学系,加州理工学院) Mathematical Sciences, California Institute of Technology(数学科学系,加州理工学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 U-DAVI通过引入空间自适应扰动和不确定性估计,改进了基于扩散的图像重建方法,实现更高质量的重建效果。

Comments Accepted at ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11703 2026-02-13 cs.CV cs.AI 79%

Semantically Conditioned Diffusion Models for Cerebral DSA Synthesis

具有语义条件的扩散模型用于脑部DSA合成

Qiwen Xu, David Rügamer, Holger Wenz, Johann Fontana, Nora Meggyeshazi, Andreas Bender, Máté E. Maros

机构 * Department of Statistics, LMU Munich(统计学系) Department of Biomedical Informatics (DBMI), Medical Faculty Mannheim, Heidelberg University(生物医学信息学系) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) Clinic for Diagnostic and Interventional Neuroradiology, Medical Faculty Mannheim, Heidelberg University(诊断和介入神经放射科) Department of Anesthesiology and Intensive Care Medicine, BG Trauma Center Tuebingen(麻醉和重症医学科)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本研究提出一种具有语义条件的扩散模型,用于生成具有真实感的脑部DSA图像,以支持医学研究和算法开发。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11653 2026-02-13 cs.CV 79%

GR-Diffusion: 3D Gaussian Representation Meets Diffusion in Whole-Body PET Reconstruction

GR-Diffusion:3D高斯表示与扩散在全身体PET重建中的结合

Mengxiao Geng, Zijie Chen, Ran Hong, Bingxuan Li, Qiegen Liu

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 GR-Diffusion结合3D高斯表示与扩散模型,提升全身体PET重建的图像质量和细节保留能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11545 2026-02-13 cs.CV 79%

Supervise-assisted Multi-modality Fusion Diffusion Model for PET Restoration

监督辅助多模态融合扩散模型用于PET修复

Yingkai Zhang, Shuang Chen, Ye Tian, Yunyi Gao, Jianyong Jiang, Ying Fu

机构 * Beijing Institute of Technology(北京理工大学) Research and Development Center of Agricultural Bank of China(中国农业银行研发中心) Peking University(北京大学) Beijing Normal University(北京师范大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出监督辅助多模态融合扩散模型,用于从低剂量PET和MR图像中恢复高质量SPET图像,通过优化特征融合和两阶段监督学习策略提升修复效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11446 2026-02-13 cs.CV cs.AI 79%

Enhanced Portable Ultra Low-Field Diffusion Tensor Imaging with Bayesian Artifact Correction and Deep Learning-Based Super-Resolution

增强型便携式超低场扩散张量成像与贝叶斯伪影校正及基于深度学习的超分辨率

Mark D. Olchanyi, Annabel Sorby-Adams, John Kirsch, Brian L. Edlow, Ava Farnan, Renfei Liu, Matthew S. Rosen, Emery N. Brown, W. Taylor Kimberly, Juan Eugenio Iglesias

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出了一种便携式超低场扩散张量成像序列及贝叶斯伪影校正算法,结合深度学习超分辨率技术,提升白质成像质量并应用于阿尔茨海默病分类。

Comments 38 pages, 8 figures, 2 supplementary figures, and 3 supplementary tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11214 2026-02-13 cs.CV cs.RO 79%

DD-MDN: Human Trajectory Forecasting with Diffusion-Based Dual Mixture Density Networks and Uncertainty Self-Calibration

DD-MDN: 基于扩散的双混合密度网络与不确定性自校准的人体轨迹预测

Manuel Hetzel, Kerim Turacan, Hannes Reichert, Konrad Doll, Bernhard Sick

机构 * Faculty of Engineering, University of Applied Sciences Aschaffenburg(亚琛应用科学大学工程学院) Intelligent Embedded Systems Lab, University of Kassel(卡塞尔大学智能嵌入式系统实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 DD-MDN通过结合扩散模型和双混合密度网络,实现高精度的人体轨迹预测,具备自校准的不确定性建模和对短观察期的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01497 2026-02-13 cs.CV 79%

Learning A Physical-aware Diffusion Model Based on Transformer for Underwater Image Enhancement

基于变压器的物理感知扩散模型学习用于水下图像增强

Chen Zhao, Chenyu Dong, Weiling Cai, Yueyue Wang

机构 * School of Intelligence Science and Technology, Nanjing University(智能科学与技术学院,南京大学) School of Artificial Intelligence, Nanjing Normal University(人工智能学院,南京师范大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出基于物理知识的扩散模型PA-Diff,通过物理先验生成和隐式神经重建提升水下图像增强效果,实验表明其在水下图像增强任务中表现最优。

Comments IEEE Transactions on Geoscience and Remote Sensing (TGRS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09181 2026-02-13 cs.CV 79%

Improving Efficiency of Diffusion Models via Multi-Stage Framework and Tailored Multi-Decoder Architectures

通过多阶段框架和定制多解码器架构提高扩散模型的效率

Huijie Zhang, Yifu Lu, Ismail Alkhouri, Saiprasad Ravishankar, Dogyoon Song, Qing Qu

机构 * University of Michigan(密歇根大学) Michigan State University(密歇根州立大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出多阶段框架和定制多解码器架构,以提高扩散模型的训练和采样效率。

Comments The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12229 2026-02-13 cs.LG 78%

Diffusion Alignment Beyond KL: Variance Minimisation as Effective Policy Optimiser

扩散对齐超越KL:方差最小化作为有效的策略优化器

Zijing Ou, Jacob Si, Junyi Zhu, Ondrej Bohdal, Mete Ozay, Taha Ceritli, Yingzhen Li

机构 * Imperial College London(帝国理工学院伦敦分校) Samsung R&D Institute UK(三星英国研发中心)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出方差最小化策略优化方法,通过最小化对数重要权重的方差来实现扩散对齐,超越传统KL优化,提供新的设计方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12045 2026-02-13 cs.LG cs.AI 78%

Fourier Transformers for Latent Crystallographic Diffusion and Generative Modeling

用于潜在晶体学扩散和生成建模的傅里叶变换器

Jed A. Duersch, Elohan Veillon, Astrid Klipfel, Adlane Sayede, Zied Bouraoui

机构 * CRIL UMR 8188, Universit\'e d'Artois, CNRS, France

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出一种基于傅里叶变换的生成模型,用于高效生成具有周期性和对称性的晶体结构,通过潜在扩散模型在压缩空间中实现高精度重建和生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22163 2026-02-13 quant-ph 78%

A quantum advection-diffusion solver using the quantum singular value transform

基于量子奇异值变换的量子对流-扩散求解器

Gard Olav Helle, Tommaso Benacchio, Anna Bomme Ousager, Jørgen Ellegaard Andersen

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出利用量子奇异值变换和高阶有限差分算子,实现对流-扩散方程的高效量子求解,降低门和量子比特需求。

Comments 45 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19078 2026-02-13 cs.LG 78%

Diffusion Bridge Variational Inference for Deep Gaussian Processes

扩散桥变分推断用于深度高斯过程

Jian Xu, Qibin Zhao, John Paisley, Delu Zeng

机构 * RIKEN AIP Columbia University(哥伦比亚大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出DBVI,通过学习的数据依赖初始分布改进DDVI,提升大规模深度高斯过程的推断效率和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04100 2026-02-13 math.DG math.AT 78%

Manifold Diffusion Geometry: Curvature, Tangent Spaces, and Dimension

流形扩散几何:曲率、切空间与维度

Iolo Jones

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出基于扩散几何的流形数据曲率、切空间和维度估计方法,具有较高的鲁棒性和准确性,尤其在噪声或稀疏数据情况下表现更优。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11540 2026-02-13 physics.bio-ph q-bio.PE 78%

Transition from traveling fronts to diffusion-limited growth in expanding populations

旅行波向扩散限制生长的转变在扩展种群中

Louis Brezin, Kyle J. Shaffer, Kirill S. Korolev

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 该研究通过分析反应-扩散方程,揭示了扩展种群中旅行波向扩散限制生长转变的机制,解释了菌落面积随时间线性增长的原因。

Comments To be published in Phys. Rev. E

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11477 2026-02-13 eess.AS cs.CE 78%

SLD-L2S: Hierarchical Subspace Latent Diffusion for High-Fidelity Lip to Speech Synthesis

SLD-L2S: 基于分层子空间潜在扩散的高质量唇部到语音合成

Yifan Liang, Andong Li, Kang Yang, Guochen Yu, Fangkun Liu, Lingling Dai, Xiaodong Li, Chengshi Zheng

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 SLD-L2S通过分层子空间潜在扩散模型直接将唇部运动映射到语音编解码器的潜在空间,提升语音合成质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11364 2026-02-13 cs.CL cs.AI 78%

The Energy of Falsehood: Detecting Hallucinations via Diffusion Model Likelihoods

虚假信息的能量:通过扩散模型似然检测幻觉

Arpit Singh Gautam, Kailash Talreja, Saurabh Jha

机构 * Dell Technologies, CSG CTO Team(戴尔技术,CSG CTO团队)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 DiffuTruth通过扩散模型似然检测幻觉,利用非平衡热力学原理,提出语义能量度量和混合校准方法,在FEVER和HOVER数据集上取得显著性能提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14312 2026-02-13 cs.LG cs.AI 78%

H-LDM: Hierarchical Latent Diffusion Models for Controllable and Interpretable PCG Synthesis from Clinical Metadata

H-LDM:用于从临床元数据生成可控且可解释的PCG合成的分层潜在扩散模型

Chenyang Xu, Siming Li, Hao Wang

机构 * School of Cyber Engineering Xidian University Xi'an, China(电子科技大学信息工程学院西安中国) School of Telecommunications Engineering Xidian University Xi'an, China(电子科技大学电信工程学院西安中国)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 H-LDM通过分层潜在扩散模型从临床元数据生成可控且可解释的PCG信号,提升心血管疾病诊断的准确性和有效性。

Comments This paper was accepted by IEEE BIBM 2025 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03692 2026-02-13 math.PR 78%

Non-negative diffusion bridge of the McKean-Vlasov type: analysis of singular diffusion and application to fish migration

非负的 McKean-Vlasov 类扩散桥:对奇异扩散的分析及其在鱼类迁徙中的应用

Hidekazu Yoshioka

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出了一种非负的 McKean-Vlasov 类扩散桥,分析了奇异扩散系数对鱼类迁徙模型的影响,并应用于实际数据研究。

Comments Updated on February 12, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12003 2026-02-13 cs.CV 70%

Projected Representation Conditioning for High-fidelity Novel View Synthesis

投影表示条件化用于高质量新视角合成

Min-Seop Kwak, Minkyung Kwon, Jinhyeok Choi, Jiho Park, Seungryong Kim

专题命中 扩散模型 :diffusion(abstract);inpainting(abstract);分类 cs.CV

AI总结 本文提出ReNoV框架,通过投影外部表示提升扩散模型在高质量新视角合成中的几何一致性和重建保真度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11942 2026-02-13 cs.CV cs.AI 57%

Synthesis of Late Gadolinium Enhancement Images via Implicit Neural Representations for Cardiac Scar Segmentation

通过隐式神经表示合成晚期钆增强图像用于心脏瘢痕分割

Soufiane Ben Haddou, Laura Alvarez-Florez, Erik J. Bekkers, Fleur V. Y. Tjong, Ahmad S. Amin, Connie R. Bezzina, Ivana Išgum

机构 * Amsterdam UMC, The Netherlands(阿姆斯特丹大学医学中心,荷兰) University of Amsterdam, The Netherlands(阿姆斯特丹大学,荷兰) Informatics Institute, University of Amsterdam, The Netherlands(阿姆斯特丹大学信息学院,荷兰) Amsterdam Cardiovascular Sciences, Amsterdam UMC, The Netherlands(阿姆斯特丹心血管科学,阿姆斯特丹大学医学中心,荷兰) Amsterdam Machine Learning Lab, University of Amsterdam, The Netherlands(阿姆斯特丹机器学习实验室,阿姆斯特丹大学,荷兰) Department of Biomedical Engineering and Physics, Amsterdam UMC, The Netherlands(生物医学工程与物理系,阿姆斯特丹大学医学中心,荷兰) Department of Experimental Cardiology, Amsterdam Cardiovascular Sciences, Heart Failure & Arrhythmias, Amsterdam UMC, The Netherlands(实验心脏病学系,阿姆斯特丹心血管科学,心力衰竭与心律失常,阿姆斯特丹大学医学中心,荷兰) Department of Cardiology, Amsterdam UMC, The Netherlands(心脏病学系,阿姆斯特丹大学医学中心,荷兰) Department of Radiology and Nuclear Medicine, Amsterdam UMC, The Netherlands(放射学与核医学系,阿姆斯特丹大学医学中心,荷兰) Department of Radiology, Mayo Clinic, Rochester, United States of America(放射学系,梅奥诊所,罗切斯特,美国)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出利用隐式神经表示和扩散模型合成LGE图像,以提高心脏瘢痕分割的性能,通过生成合成数据缓解标注数据不足的问题。

Comments Paper accepted at SPIE Medical Imaging 2026 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15253 2026-02-13 cs.CV 57%

Dual Frequency Branch Framework with Reconstructed Sliding Windows Attention for AI-Generated Image Detection

双频分支框架与重构滑动窗口注意力用于AI生成图像检测

Jiazhen Yan, Ziqiang Li, Fan Wang, Ziwen He, Zhangjie Fu

机构 * Engineering Research Center of Digital Forensics, Ministry of Education, Nanjing University of Information Science and Technology(数字取证工程研究中心,教育 ministry,南京信息科学技术大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出双频域分支框架与重构滑动窗口注意力机制,用于提升AI生成图像检测的泛化能力,实验表明在多种生成模型图像上检测准确率提升2.13%。

Comments Accepted by IEEE Transactions on Information Forensics and Security

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11769 2026-02-13 cs.CV 57%

Light4D: Training-Free Extreme Viewpoint 4D Video Relighting

Light4D: 无需训练的极端视角4D视频重映射

Zhenghuang Wu, Kang Chen, Zeyu Zhang, Hao Tang

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 Light4D提出一种无需训练的框架,通过解耦流引导和时间一致注意力,实现极端视角下的4D视频重映射,提升时间一致性和光照保真度。

详情

展开后加载摘要…

URL PDF HTML 收藏