arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-12-09 至 2025-12-09 共收录 70 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 70 篇

2410.06055 2025-12-09 cs.CV 88%

RepLDM: Reprogramming Pretrained Latent Diffusion Models for High-Quality, High-Efficiency, High-Resolution Image Generation

RepLDM: 重编程预训练的潜在扩散模型以实现高质量、高效、高分辨率图像生成

Boyuan Cao, Jiaxin Ye, Yujie Wei, Hongming Shan

专题命中 扩散模型 :image generation(title,abstract);diffusion(title,abstract);分类 cs.CV

AI总结 RepLDM通过重编程预训练的潜在扩散模型,实现高质量、高效、高分辨率的图像生成,优于现有方法。

Comments NeurIPS 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16302 2025-12-09 cs.LG cs.AI cs.CR cs.CV 85%

Towards Resilient Safety-driven Unlearning for Diffusion Models against Downstream Fine-tuning

面向对抗下游微调的鲁棒安全驱动遗忘方法用于扩散模型

Boheng Li, Renjie Gu, Junjie Wang, Leyi Qi, Yiming Li, Run Wang, Zhan Qin, Tianwei Zhang

机构 * Nanyang Technological University, Singapore(南洋理工大学,新加坡) Central South University, China(中南大学,中国) Key Laboratory of Aerospace Information Security and Trusted Computing, Ministry of Education, School of Cyber Science and Engineering, Wuhan University, China(航空航天信息安全部门,教育部,武汉大学,中国) State Key Laboratory of Blockchain and Data Security, Zhejiang University, China(区块链与数据安全国家重点实验室,浙江大学,中国)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);text-to-image(abstract);分类 cs.CV

AI总结 本文提出ResAlign,一种针对扩散模型对抗下游微调的鲁棒安全驱动遗忘框架,通过隐含优化问题建模和元学习策略提升安全性和生成能力。

Comments Accepted to the 39th Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02483 2025-12-09 cs.CV 83%

Event-Customized Image Generation

事件定制图像生成

Zhen Wang, Yilei Jiang, Dong Zheng, Jun Xiao, Long Chen

机构 * Zhejiang University, Hangzhou, China(浙江大学) The Hong Kong University of Science and Technology(香港科技大学)

专题命中 扩散模型 :image generation(title,abstract);diffusion(abstract);分类 cs.CV

AI总结 本文提出FreeEvent方法,通过引入实体切换和事件转移路径,实现事件定制化图像生成,提升复杂场景下的定制化能力。

Journal ref ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06353 2025-12-09 cs.CV 83%

TreeQ: Pushing the Quantization Boundary of Diffusion Transformer via Tree-Structured Mixed-Precision Search

TreeQ: 通过树结构混合精度搜索推动扩散变换器的量化边界

Kaicheng Yang, Kaisen Yang, Baiting Wu, Xun Zhang, Qianrui Yang, Haotong Qin, He Zhang, Yulun Zhang

机构 * Shanghai Jiao Tong University(上海交通大学) Tsinghua University(清华大学) ETH Zürich(苏黎世联邦理工学院) Adobe Research(Adobe研究院)

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV

AI总结 TreeQ通过树结构混合精度搜索方法,首次在DiT模型上实现接近无损的4位PTQ性能,解决了DiT量化中的关键挑战。

Comments Code and Supplementary Material could be found at https://github.com/racoonykc/TreeQ

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22505 2025-12-09 cs.RO cs.CV 80%

RealD$^2$iff: Bridging Real-World Gap in Robot Manipulation via Depth Diffusion

RealD$^2$iff: 通过深度扩散弥合机器人操作中的现实差距

Xiujian Liang, Jiacheng Liu, Mingyang Sun, Qichen He, Cewu Lu, Jianhua Sun

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 RealD$^2$iff通过深度扩散模型弥合现实与模拟之间的差距,实现无需额外微调的零样本机器人操作,并生成现实世界般的深度数据集。

Comments We are the author team of the paper "RealD$^2$iff: Bridging Real-World Gap in Robot Manipulation via Depth Diffusion". After self-examination, our team discovered inappropriate wording in the citation of related work, the introduction, and the contribution statement, which may affect the contribution of other related works. Therefore, we have decided to revise the paper and request its withdrawal

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07745 2025-12-09 cs.CV 79%

DiffusionDriveV2: Reinforcement Learning-Constrained Truncated Diffusion Modeling in End-to-End Autonomous Driving

DiffusionDriveV2: 结束到结束自动驾驶中的强化学习约束截断扩散建模

Jialv Zou, Shaoyu Chen, Bencheng Liao, Zhiyu Zheng, Yuehao Song, Lefei Zhang, Qian Zhang, Wenyu Liu, Xinggang Wang

机构 * School of Electronic Information and Communications, Huazhong University of Science & Technology(华中科技大学电子信息与通信学院) Institute of Artificial Intelligence, Huazhong University of Science & Technology(华中科技大学人工智能研究院) Horizon Robotics(地平线科技) School of Computer Science, Wuhan University(武汉大学计算机学院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 DiffusionDriveV2通过强化学习约束和探索,解决了自动驾驶中截断扩散模型在多样性和高质量之间的平衡问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07480 2025-12-09 cs.CV 79%

Single-step Diffusion-based Video Coding with Semantic-Temporal Guidance

单步扩散视频编码与语义-时间引导

Naifu Xue, Zhaoyang Jia, Jiahao Li, Bin Li, Zihan Zheng, Yuan Zhang, Yan Lu

机构 * Communication University of China(中国通信大学) University of Science and Technology of China(中国科学技术大学) Microsoft Research Asia(微软亚洲研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 S2VC通过单步扩散与语义-时间引导,实现低比特率下的高质量视频编码,比特率节省达52.73%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07251 2025-12-09 cs.CV 79%

See More, Change Less: Anatomy-Aware Diffusion for Contrast Enhancement

看得更多,改变更少:面向对比增强的解剖感知扩散模型

Junqi Liu, Zejun Wu, Pedro R. A. S. Bassi, Xinze Zhou, Wenxuan Li, Ibrahim E. Hamamci, Sezgin Er, Tianyu Lin, Yi Luo, Szymon Płotka, Bjoern Menze, Daguang Xu, Kai Ding, Kang Wang, Yang Yang, Yucheng Tang, Alan L. Yuille, Zongwei Zhou

机构 * Johns Hopkins University(约翰霍普金斯大学) University of Copenhagen(哥本哈根大学) University of Virginia(弗吉尼亚大学) University of Bologna(博洛尼亚大学) Italian Institute of Technology(意大利理工学院) University of Zurich(苏黎世大学) ETH AI Center(ETH人工智能中心) Istanbul Medipol University(伊斯坦布尔梅迪波尔大学) Jagiellonian University(雅盖隆大学) NVIDIA(NVIDIA公司) Johns Hopkins Medicine(约翰霍普金斯医学) University of California, San Francisco(旧金山大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 SMILE是一种解剖感知扩散模型,通过结构感知监督和统一推理,在医学图像对比增强中实现更准确且临床有价值的图像生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07201 2025-12-09 cs.CV cs.LG 79%

Understanding Diffusion Models via Code Execution

通过代码执行理解扩散模型

Cheng Yu

机构 * School of Artificial Intelligence, Chongqing University of Technology, Chongqing , China(人工智能学院,重庆理工大学,重庆,中国) DiAi Corporation, China(迪爱公司,中国)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文通过代码实现从代码执行角度解释扩散模型,保留核心组件并去除冗余细节,为研究人员提供实践层面的理解。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01267 2025-12-09 cs.CV 79%

Diffusion-based Adversarial Purification from the Perspective of the Frequency Domain

基于频域视角的扩散对抗净化

Gaozheng Pei, Ke Ma, Yingfei Sun, Qianqian Xu, Qingming Huang

机构 * School of Electronic, Electrical and Communication Engineering, UCAS, Beijing.(电子、电气与通信工程学院,中国科学院大学,北京) Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, CAS, Beijing.(智能信息处理重点实验室,计算技术研究所,中国科学院,北京) School of Computer Science and Technology, UCAS, Beijing.(计算机科学与技术学院,中国科学院大学,北京) Key Laboratory of Big Data Mining and Knowledge Management, UCAS, Beijing.(大数据挖掘与知识管理重点实验室,中国科学院大学,北京)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出基于频域视角的扩散对抗净化方法,通过频域成分分离和针对性处理,有效消除对抗扰动并保留图像内容与结构。

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18723 2025-12-09 eess.IV cs.CV cs.LG 79%

MRI Reconstruction with Regularized 3D Diffusion Model (R3DM)

利用正则化的3D扩散模型(R3DM)进行MRI重建

Arya Bangun, Zhuo Cao, Alessio Quercia, Hanno Scharr, Elisabeth Pfaehler

机构 * IAS-8 INM-4 Forschungszentrum Jülich(INM-4 研究中心) RWTH Aachen University(亚琛工业大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出了一种基于正则化3D扩散模型的MRI重建方法,通过结合优化方法提升图像质量与保真度,实验表明其在欠采样数据下的重建性能优于现有方法。

Comments Accepted to WACV 2025,17 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.18886 2025-12-09 cs.LG cs.AI 79%

A Survey on Diffusion Models for Time Series and Spatio-Temporal Data

时间序列和时空数据扩散模型综述

Yiyuan Yang, Ming Jin, Haomin Wen, Chaoli Zhang, Yuxuan Liang, Lintao Ma, Yi Wang, Chenghao Liu, Bin Yang, Zenglin Xu, Shirui Pan, Qingsong Wen

机构 * University of Oxford(牛津大学) Griffith University(格里菲斯大学) Carnegie Mellon University(卡内基梅隆大学) Zhejiang Normal University(浙江师范大学) Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) The University of Hong Kong(香港大学) Salesforce Research(Salesforce研究) East China Normal University(东华大学) Fudan University(复旦大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文综述了扩散模型在时间序列和时空数据中的应用,系统梳理了模型类别、任务类型及实际应用场景,为后续研究提供基础。

Comments Accepted by ACM Computing Surveys; 37 pages; Github Repo: https://github.com/yyysjz1997/Awesome-TimeSeries-SpatioTemporal-Diffusion-Model

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07744 2025-12-09 cond-mat.stat-mech cond-mat.str-el 78%

Anomalous coarsening and nonlinear diffusion of kinks in an one-dimensional quasi-classical Holstein model

一维准经典Holstein模型中kinks的异常粗化与非线性扩散

Ho Jang, Yang Yang, Gia-Wei Chern

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 研究一维准经典Holstein模型中kinks的异常粗化与非线性扩散,揭示电子统计与kink跳跃对慢粗化机制的影响。

Comments 13 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07555 2025-12-09 q-fin.MF math.PR 78%

On the structure of increasing profits in a 1D general diffusion market with interest rates

关于具有利率的1D一般扩散市场中增加利润结构的研究

Alexis Anagnostakis, David Criens, Mikhail Urusov

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文研究了具有利率的1D一般扩散市场中增加利润的结构,通过分析确定性测度和交易策略,揭示了无套利理论与随机过程理论之间的新联系。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07350 2025-12-09 cs.DC 78%

Communication-Efficient Serving for Video Diffusion Models with Latent Parallelism

面向视频扩散模型的通信高效服务与潜在并行性

Zhiyuan Wu, Shuai Wang, Li Chen, Kaihui Gao, Dan Li, Yanyu Ren, Qiming Zhang, Yong Wang

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出潜在并行性策略,通过动态旋转潜在空间维度,显著降低视频扩散模型服务的通信开销,同时保持生成质量。

Comments 19 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07289 2025-12-09 cond-mat.mtrl-sci cs.LG 78%

Equivariant Diffusion for Crystal Structure Prediction

等价扩散用于晶体结构预测

Peijia Lin, Pin Chen, Rui Jiao, Qing Mo, Jianhuan Cen, Wenbing Huang, Yang Liu, Dan Huang, Yutong Lu

机构 * School of Computer Science Engineering, Sun Yat-sen University, Guangzhou, China National Supercomputer Center in Guangzhou, China Dept. of Comp. Sci. \& Tech., Institute for AI, Tsinghua University, Beijing, China Institute for AIR, Tsinghua University, Beijing, China Gaoling School of Artificial Intelligence, Renmin University of China, Beijing, China Beijing Key Laboratory of Big Data Management Analysis Methods, Beijing, China

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 EquiCSP通过等价扩散模型解决晶体结构预测中的对称性问题,提升生成结构的准确性并加快训练收敛速度。

Comments ICML 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07184 2025-12-09 cs.LG 78%

UniDiff: A Unified Diffusion Framework for Multimodal Time Series Forecasting

UniDiff: 一种用于多模态时间序列预测的统一扩散框架

Da Zhang, Bingyu Li, Zhuyuan Zhao, Junyu Gao, Feiping Nie, Xuelong Li

机构 * School of Artificial Intelligence, OPtics and ElectroNics (iOPEN), Northwestern Polytechnical University, Xi’an 710072, China(人工智能学院、光学与电子学(iOPEN)、西北工业大学,西安710072,中国) Institute of Artificial Intelligence (TeleAI), China Telecom, China(人工智能研究所(TeleAI)、中国电信,中国)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 UniDiff提出了一种统一的扩散框架,通过融合文本和时间戳信息,提升多模态时间序列预测的准确性与鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06977 2025-12-09 eess.IV cs.LG 78%

Physics-Guided Diffusion Priors for Multi-Slice Reconstruction in Scientific Imaging

用于科学成像多切片重建的物理引导扩散先验

Laurentius Valdy, Richard D. Paul, Alessio Quercia, Zhuo Cao, Xuan Zhao, Hanno Scharr, Arya Bangun

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出了一种结合物理约束和扩散先验的框架,用于提高多切片重建的效率和质量,适用于医学和科学成像领域。

Comments 8 pages, 5 figures, AAAI AI2ASE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02768 2025-12-09 eess.SP 78%

Diffusion-Prior Split Gibbs Sampling for Synthetic Aperture Radar Imaging under Incomplete Measurements

扩散先验分割吉布斯采样用于不完全测量下的合成孔径雷达成像

Hefei Gao, Tianyao Huang, Letian Guo, Jie He, Yonina C. Eldar

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出一种基于扩散的分割吉布斯采样框架,通过整合测量保真度与学习的扩散先验,提升SAR成像在不完整测量下的重建性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21417 2025-12-09 cs.LG 78%

Self-diffusion for Solving Inverse Problems

自扩散用于求解逆问题

Guanxiong Luo, Shoujin Huang, Yanlong Yang

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 自扩散通过自包含的迭代过程,无需预训练模型,有效解决逆问题,适应性强且性能优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.03930 2025-12-09 eess.AS cs.AI cs.CL cs.LG cs.SD 78%

DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation

DiTAR:基于扩散变换器的语音生成自回归建模

Dongya Jia, Zhuo Chen, Jiawei Chen, Chenpeng Du, Jian Wu, Jian Cong, Xiaobin Zhuang, Chumin Li, Zhen Wei, Yuping Wang, Yuxuan Wang

机构 * ByteDance Seed(字节跳动种子)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 DiTAR通过结合语言模型和扩散变换器,提出一种基于补丁的自回归框架,有效提升连续语音生成的效率与质量。

Comments ByteDance Seed template, ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08846 2025-12-09 math.PR 78%

On the number of crossings and bouncings of a diffusion at a sticky threshold

关于扩散在粘性阈值处交叉和反弹数量的研究

Alexis Anagnostakis, Sara Mazzonetto

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文研究了粘性扩散在阈值处交叉和反弹数量的渐近行为,并提出了粘性参数的估计方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.04456 2025-12-09 physics.soc-ph stat.OT 78%

Modeling diffusion in networks with communities: a multitype branching process approach

在具有社区结构的网络中建模扩散:一种多类型分支过程方法

Alina Dubovskaya, Caroline B. Pena, David J. P. O'Sullivan

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出利用多类型分支过程方法,在具有社区结构的网络中建模和分析扩散过程,通过度分布计算传播动态特征,并验证了其在不同网络结构中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06183 2025-12-09 cs.LG 78%

PMA-Diffusion: A Physics-guided Mask-Aware Diffusion Framework for TSE from Sparse Observations

PMA-Diffusion:一种基于物理的掩码感知扩散框架,用于从稀疏观测中重建交通流

Lindong Liu, Zhixiong Jin, Seongjin Choi

机构 * Department of Civil, Environmental, and Geo- Engineering, University of Minnesota, 500 Pillsbury Dr. SE, Minneapolis, MN 55455, USA(土木、环境与地球工程系,明尼苏达大学) Univ. Gustave Eiffel, ENTPE, EMob-Lab, Lyon, F-69675, France(格勒诺布尔伊夫·欧贝尔大学,ENTPE,EMob-Lab,法国里昂)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 PMA-Diffusion通过结合掩码感知扩散先验和物理引导后验采样器,有效重建稀疏观测下的交通流状态。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17482 2025-12-09 cs.RO 78%

Variational Shape Inference for Grasp Diffusion on SE(3)

变分形状推断用于SE(3)上的抓取扩散

S. Talha Bukhari, Kaivalya Agrawal, Zachary Kingston, Aniket Bera

机构 * Department of Computer Science, Purdue University(计算机科学系,普渡大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出一种基于变分形状推断的SE(3)抓取扩散框架,通过隐式神经表示训练自编码器并引导扩散模型,实现鲁棒的多模态抓取合成,实验显示在ACRONYM数据集上性能提升6.3%并具备零样本迁移能力。

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07130 2025-12-09 cs.RO cs.CV 74%

Mimir: Hierarchical Goal-Driven Diffusion with Uncertainty Propagation for End-to-End Autonomous Driving

Mimir:基于不确定性传播的分层目标驱动扩散用于端到端自动驾驶

Zebin Xing, Yupeng Zheng, Qichao Zhang, Zhixing Ding, Pengxuan Yang, Songen Gu, Zhongpu Xia, Dongbin Zhao

机构 * Institue of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) China University of Geosciences(中国地质大学) Fudan University(复旦大学)

专题命中 扩散模型 :diffusion(title);分类 cs.CV

AI总结 Mimir通过不确定性传播和多速率指导机制,提升端到端自动驾驶中轨迹生成的鲁棒性和推理效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.15758 2025-12-09 math.PR 71%

Up-down ordered Chinese restaurant processes with two-sided immigration, emigration and diffusion limits

上下有序的中国餐厅过程及其双侧移民、退居和扩散极限

Quan Shi, Matthias Winkel

专题命中 扩散模型 :diffusion(title)

AI总结 本文研究了上下有序中国餐厅过程的缩放极限,扩展了三参数族模型,揭示了泊松-迪里赫利特分布的参数范围扩展及其在相关过程中的应用。

Comments 52 pages, 4 figures, to appear in the Annals of Applied Probability

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01302 2025-12-09 cs.CV 70%

DCText: Scheduled Attention Masking for Visual Text Generation via Divide-and-Conquer Strategy

DCText: 通过分而治之策略实现视觉文本生成的调度注意力遮罩

Jaewoo Song, Jooyoung Choi, Kanghyun Baek, Sangyub Lee, Daemin Park, Sungroh Yoon

专题命中 扩散模型 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

AI总结 DCText通过分而治之策略和注意力遮罩技术,实现高效视觉文本生成,提升长文本和多文本处理能力,同时保持图像质量和生成效率。

Comments Accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09365 2025-12-09 eess.IV cs.CV cs.LG 70%

A Biophysically-Conditioned Generative Framework for 3D Brain Tumor MRI Synthesis

一种基于生物物理条件的生成框架用于3D脑肿瘤MRI合成

Valentin Biller, Lucas Zimmer, Ayhan Can Erdur, Sandeep Nagar, Daniel Rückert, Niklas Bubeck, Jonas Weidner

机构 * Technical University of Munich(慕尼黑技术大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) Imperial College London(伦敦帝国学院)

专题命中 扩散模型 :diffusion(abstract);inpainting(abstract);分类 cs.CV

AI总结 本文提出了一种基于生物物理条件的生成框架,用于生成高保真的3D脑肿瘤MRI,通过条件生成模型实现肿瘤合成和健康组织修复,实验结果显示在PSNR指标上取得良好效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23880 2025-12-09 cs.CV cs.GR 62%

TRELLISWorld: Training-Free World Generation from Object Generators

TRELLISWorld: 从物体生成器中无需训练的世界生成

Hanke Chen, Yuan Liu, Minchen Li

机构 * Carnegie Mellon University(卡内基梅隆大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV、cs.GR

AI总结 TRELLISWorld通过无需训练的模块化瓷砖生成器实现通用语言驱动的3D场景生成,支持多样布局、高效生成和灵活编辑。

详情

展开后加载摘要…

URL PDF HTML 收藏