arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-01-27 至 2026-01-27 共收录 92 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 7 篇

2505.09166 2026-01-27 cs.HC cs.AI 86%

An Exploration of Default Images in Text-to-Image Generation

文本到图像生成中默认图像的探索

Hannu Simonen, Atte Kiviniemi, Hannah Johnston, Helena Barranha, Jonas Oppenlaender

机构 * University of Oulu Oulu Finland Carleton University Ottawa Canada IST, University of Lisbon \& IHA-NOVA FCSH / IN2PAST Lisbon Portugal University of Oulu Carleton University IST, University of Lisbon \& IHA-NOVA FCSH / IN2PAST

专题命中 文生图 :text-to-image(title,abstract);image generation(title)

AI总结 本文首次探讨了文本到图像生成中默认图像的特性,通过实验分析揭示了默认图像的一致性及其对用户满意度的影响。

Comments 21 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17666 2026-01-27 cs.CV 85%

Training-Free Text-to-Image Compositional Food Generation via Prompt Grafting

无需训练的文本到图像组合食品生成 via 提示嫁接

Xinyue Pan, Yuhao Chen, Fengqing Zhu

机构 * Elmore Family School of Electrical and Computer Engineering, Purdue University(电气与计算机工程学院,普渡大学) Department of Systems Design Engineering, University of Waterloo(系统设计工程系,滑铁卢大学)

专题命中 文生图 :text-to-image(title,abstract);image generation(abstract);diffusion(abstract);分类 cs.CV

AI总结 无需训练的文本到图像组合食品生成方法通过提示嫁接实现对多食物图像生成的可控分离。

Comments Accepted by CAI2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17027 2026-01-27 cs.CV cs.AI 83%

Scientific Image Synthesis: Benchmarking, Methodologies, and Downstream Utility

科学图像合成:基准测试、方法论与下游应用

Honglin Lin, Chonghan Qin, Zheng Liu, Qizhi Pei, Yu Li, Zhanping Zhong, Xin Gao, Yanfeng Wang, Conghui He, Lijun Wu

机构 * Shanghai Jiao Tong University(上海交通大学) OpenDataLab, Shanghai Artificial Intelligence Laboratory(OpenDataLab,上海人工智能实验室) The University of Hong Kong(香港大学) Peking University(北京大学)

专题命中 文生图 :image synthesis(title,abstract);text-to-image(abstract);分类 cs.CV

AI总结 本文提出ImgCoder框架和SciGenBench基准,通过逻辑驱动方法提升科学图像生成的结构精度,并展示微调LMMs在科学图像上的效果,验证了高保真合成在多模态推理中的潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02567 2026-01-27 cs.CV 77%

Unified Multimodal Understanding and Generation Models: Advances, Challenges, and Opportunities

统一的多模态理解和生成模型:进展、挑战与机遇

Shanshan Zhao, Xinjie Zhang, Jintao Guo, Jiakui Hu, Lunhao Duan, Minghao Fu, Yong Xien Chng, Guo-Hua Wang, Qing-Guo Chen, Zhao Xu, Weihua Luo, Kaifu Zhang

机构 * Alibaba Group(阿里巴巴集团) Hong Kong University of Science and Technology(香港科学与技术大学) Nanjing University(南京大学) Peking University(北京大学) Tsinghua University(清华大学)

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV

AI总结 本文探讨了多模态理解和生成模型的统一进展、挑战与机遇,分析了不同架构范式及未来研究方向。

Comments In this version, we incorporate new papers (after Aug. 2025), datasets, and benchmarks. This work is still in progress; Github project: https://github.com/AIDC-AI/Awesome-Unified-Multimodal-Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18624 2026-01-27 cs.CV 70%

GeneMAN: Generalizable Single-Image 3D Human Reconstruction from Multi-Source Human Data

GeneMAN: 从多源人类数据中通用的单图像3D人体重建

Wentao Wang, Hang Ye, Fangzhou Hong, Xue Yang, Jianfu Zhang, Yizhou Wang, Ziwei Liu, Liang Pan

机构 * Shanghai AI Laboratory(上海人工智能实验室) Peking University(北京大学) Nanyang Technological University(南洋理工大学) SAIS & SCS, Shanghai Jiao Tong University(上海交通大学SAIS与SCS)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

AI总结 GeneMAN通过多源高质量数据集和深度学习方法,实现从单张图像到高保真3D人体的通用重建。

Comments Accepted by NeurIPS 2025; Project page: https://roooooz.github.io/GeneMAN/

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17706 2026-01-27 cs.CL cs.CV 57%

A Computational Approach to Visual Metonymy

一种视觉隐喻的计算方法

Saptarshi Ghosh, Linfeng Liu, Tianyu Jiang

机构 * University of Cincinnati(辛辛那提大学)

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

AI总结 本文提出了一种基于符号学理论的计算方法,通过构建ViMET数据集评估多模态语言模型在理解视觉隐喻方面的认知推理能力,并揭示了机器在处理间接视觉参考上的局限性。

Comments EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26861 2026-01-27 cs.IR cs.CL 50%

Evaluating Perspectival Biases in Cross-Modal Retrieval

评估跨模态检索中的视角偏差

Teerapol Saengsukhiran, Peerawat Chomphooyod, Narabodee Rodjananant, Chompakorn Chaksangchaichot, Patawee Prakrankamanant, Witthawin Sripheanpol, Pak Lovichit, Sarana Nutanong, Ekapol Chuangsuwanich

机构 * Department of Computer Engineering, Chulalongkorn University(楚莱隆坤大学计算机工程系) School of Information Science and Technology, VISTEC(信息科学与技术学院,VISTEC)

专题命中 文生图 :text-to-image(abstract)

AI总结 本文提出3XCM基准,揭示跨模态检索中语言和文化偏见的影响,强调需通过解耦策略实现公平检索。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 图像编辑 4 篇

2506.07611 2026-01-27 cs.CV 79%

DragNeXt: Rethinking Drag-Based Image Editing

DragNeXt: 重新思考基于拖动的图像编辑

Yuan Zhou, Junbao Zhou, Qingshan Xu, Kesen Zhao, Yuxuan Wang, Hao Fei, Richang Hong, Hanwang Zhang

专题命中 图像编辑 :image editing(title,abstract);分类 cs.CV

AI总结 DragNeXt通过重新定义拖动为变形、旋转和翻译,提出简单有效的编辑框架,提升图像编辑质量。

Comments https://github.com/zhouyuan888888/DragNeXt

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04058 2026-01-27 cs.GR cs.CV 62%

SMooGPT: Stylized Motion Generation using Large Language Models

SMooGPT:利用大语言模型进行风格化动作生成

Lei Zhong, Yi Yang, Changjian Li

机构 * University of Edinburgh(爱丁堡大学)

专题命中 图像编辑 :diffusion(abstract);分类 cs.CV、cs.GR

AI总结 SMooGPT通过利用大语言模型的推理、组合和生成能力,实现风格化动作生成,提升动作控制和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03906 2026-01-27 cs.CV 61%

From Filters to VLMs: Benchmarking Defogging Methods through Object Detection and Segmentation Performance

从滤波器到视觉语言模型:通过目标检测和分割性能评估去雾方法

Ardalan Aryashad, Parsa Razmara, Amin Mahjoub, Seyedarmin Azizi, Mahdi Salmani, Arad Firouzkouhi

机构 * University of Southern California(南加州大学)

专题命中 图像编辑 :image editing(abstract);分类 cs.CV;diffusion(comments)

AI总结 本文通过目标检测和分割性能评估,探讨了去雾方法在真实与合成环境中的有效性,揭示了视觉语言模型在恶劣天气下的应用潜力。

Comments Accepted at WACV 2026 Proceedings (Oral), 5th Workshop on Image, Video, and Audio Quality Assessment in Computer Vision, with a focus on VLM and Diffusion Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12954 2026-01-27 cs.CV 57%

StyMam: A Mamba-Based Generator for Artistic Style Transfer

StyMam:基于Mamba的风格迁移生成器

Zhou Hong, Ning Dong, Yicheng Di, Xiaolong Xu, Rongsheng Hu, Yihua Shao, Run Ling, Yun Wang, Juqin Wang, Zhanjie Zhang, Ao Ma

机构 * School of Information Engineering, Suqian University(信息工程学院,苏庆大学) School of Artificial Intelligence and Computer Science, Jiangnan University(人工智能与计算机科学学院,江南大学) Jiangsu Province Engineering Research Center of Smart Poultry Farming and Intelligent Equipment(江苏省智能养鸡场与智能设备工程研究中心) School of Computer Science, Nanjing University of Posts and Telecommunications(计算机科学学院,南京邮电大学) Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院) School of Internet of Things, Wuxi University of Technology(物联网学院,无锡科技大学)

专题命中 图像编辑 :diffusion(abstract);分类 cs.CV

AI总结 StyMam通过引入基于Mamba的生成器,结合残差双路径扫描机制和通道加权空间注意力模块,实现了高质量且快速的图像风格迁移。

Comments Accepted by ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 扩散模型 62 篇

2601.17927 2026-01-27 cs.CV cs.MM 86%

RemEdit: Efficient Diffusion Editing with Riemannian Geometry

RemEdit: 基于黎曼几何的高效扩散编辑

Eashan Adhikarla, Brian D. Davison

专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);image editing(abstract);分类 cs.CV、cs.MM

AI总结 RemEdit通过基于黎曼几何的潜在空间导航和任务特定注意力剪枝机制,实现了高效且保真的图像编辑,同时保持实时性能。

Journal ref IEEE/CVF Winter Conference on Applications of Computer Vision, WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17259 2026-01-27 cs.CV cs.GR cs.LG 86%

Inference-Time Loss-Guided Colour Preservation in Diffusion Sampling

推理时的损失引导颜色保持在扩散采样中

Angad Singh Ahuja, Aarush Ram Anandh

机构 * Constrained Image-Synthesis Lab(受限图像合成实验室)

专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);inpainting(abstract);分类 cs.CV、cs.GR

AI总结 本文提出一种无需额外训练的推理时颜色保持方法,通过区域约束和复合损失引导扩散模型,实现精准颜色控制。

Comments 25 Pages, 12 Figures, 3 Tables, 5 Appendices, 8 Algorithms

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26442 2026-01-27 cs.IT math.IT 82%

Diffusion-Aided Bandwidth-Efficient Semantic Communication with Adaptive Requests

扩散辅助的带宽高效语义通信与自适应请求

Xuesong Wang, Xinyan Xie, Mo Li, Zhaoqian Liu

专题命中 扩散模型 :diffusion(title,abstract);inpainting(abstract)

AI总结 本文提出了一种基于反馈的闭环语义通信方案,通过自适应请求减少潜在块的传输,提高带宽效率和语义一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.02334 2026-01-27 cs.IT cs.CV cs.MM eess.SP math.IT 81%

Communicate Less, Synthesize the Rest: Latency-aware Intent-based Generative Semantic Multicasting with Diffusion Models

少传多合成:基于意图的生成语义多播与扩散模型

Xinkai Liu, Mahdi Boloursaz Mashhadi, Li Qiao, Yi Ma, Rahim Tafazolli, Mehdi Bennis

机构 * GIC & 6GIC, Institute for Communication Systems (ICS), University of Surrey(5GIC与6GIC,通信系统研究所(ICS),萨里大学) School of Information and Electronics, Beijing Institute of Technology(信息电子学院,北京理工大学) Centre for Wireless Communications, University of Oulu(无线通信中心,奥卢大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.MM

AI总结 本文提出一种基于意图的生成语义多播方法,利用预训练扩散模型减少传输延迟,同时保持高质量信号合成。

Comments Accepted at IEEE Transactions on Vehicular Technology

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02831 2026-01-27 cs.CV 80%

No Other Representation Component Is Needed: Diffusion Transformers Can Provide Representation Guidance by Themselves

无需其他表示组件:扩散变换器本身可以提供表示指导

Dengyang Jiang, Mengmeng Wang, Liuzhuozheng Li, Lei Zhang, Haoyu Wang, Wei Wei, Guang Dai, Yanning Zhang, Jingdong Wang

机构 * Northwestern Polytechnical University(西北工业大学) SGIT AI Lab, State Grid Corporation of China(国网SGIT人工智能实验室) Zhejiang University of Technology(浙江工业大学) Baidu Inc.(百度公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出SRA方法,利用扩散变换器自身内部表示进行对齐,无需外部组件即可提升生成模型性能。

Comments ICLR 2026. Self-Representation Alignment for Diffusion Transformers. Code: https://github.com/vvvvvjdy/SRA

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18556 2026-01-27 cs.CV cs.LG 79%

Generative Diffusion Augmentation with Quantum-Enhanced Discrimination for Medical Image Diagnosis

生成扩散增强与量子增强判别用于医学图像诊断

Jingsong Xia, Siqi Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 SDA-QEC通过生成扩散增强与量子增强判别技术,提升医学图像分类的诊断准确性与平衡性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18168 2026-01-27 cs.CV 79%

TempDiffReg: Temporal Diffusion Model for Non-Rigid 2D-3D Vascular Registration

TempDiffReg:用于非刚性2D-3D血管配准的时序扩散模型

Zehua Liu, Shihao Zou, Jincai Huang, Yanfang Zhang, Chao Tong, Weixin Si

机构 * School of Computer Science and Engineering, Beihang University, Beijing, China(北京航空航天大学计算机科学与工程学院) State Key Laboratory of Virtual Reality Technology and Systems, Beihang University, Beijing, China(北京航空航天大学虚拟现实技术与系统国家重点实验室) Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences, Shenzhen, China(中国科学院深圳先进技术研究所) School of Computer Science and Control Engineering, Shenzhen University of Advanced Technology, Shenzhen, China(深圳先进技术大学计算机科学与控制工程学院) Department of Interventional Radiology, Shenzhen People’s Hospital, Shenzhen, China(深圳人民医院介入放射科)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 TempDiffReg通过时序扩散模型和结构感知视角n点模块,实现高精度的2D-3D血管配准,提升TACE手术的准确性和安全性。

Comments Accepted by IEEE BIBM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18832 2026-01-27 cs.CV 79%

Localizing Knowledge in Diffusion Transformers

在扩散变换器中定位知识

Arman Zarei, Samyadeep Basu, Keivan Rezaei, Zihao Lin, Sayan Nag, Soheil Feizi

机构 * University of Maryland(马里兰大学) University of California, Davis(加州大学戴维斯分校) Adobe(Adobe公司)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出了一种方法,用于在扩散变换器中定位特定类型的知识,通过评估不同模型和知识类别,实现了高效且有针对性的模型更新。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17228 2026-01-27 cs.CV q-bio.QM 79%

Semi-Supervised Domain Adaptation with Latent Diffusion for Pathology Image Classification

半监督域适应与潜在扩散用于病理图像分类

Tengyue Zhang, Ruiwen Ding, Luoting Zhuang, Yuxiao Wu, Erika F. Rodriguez, William Hsu

机构 * Medical & Imaging Informatics, Department of Radiological Sciences, David Geffen School of Medicine, University of California, Los Angeles(医学与影像信息学系,放射科学系,大卫·盖弗医学院,加州大学洛杉矶分校) Bioengineering Department, Henry Samueli School of Engineering, UCLA(生物工程系,亨利·萨缪尔西工程学院,UCLA) Department of Pathology & Laboratory Sciences, David Geffen School of Medicine, UCLA(病理学与实验室科学系,大卫·盖弗医学院,UCLA)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本研究提出利用潜在扩散模型生成目标领域意识的合成数据,提升计算病理学中的域泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10917 2026-01-27 cs.CV cs.AI cs.LG 79%

Self-learned representation-guided latent diffusion model for breast cancer classification in deep ultraviolet whole surface images

自学习表征引导的潜在扩散模型用于深紫外全表面图像中的乳腺癌分类

Pouya Afshin, David Helminiak, Tianling Niu, Julie M. Jorns, Tina Yen, Bing Yu, Dong Hye Ye

机构 * Department of Computer Science, Georgia State University(计算机科学系,佐治亚州立大学) Department of Electrical and Computer Engineering, Marquette University(电气与计算机工程系,马基特大学) Joint Dept. of Biomedical Eng., Marquette Univ. and Med. Coll. of Wisconsin(马基特大学与威斯康星医学院联合生物医学工程系)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

AI总结 本文提出自监督学习引导的潜在扩散模型,通过生成高质量合成数据提升深紫外全表面图像中乳腺癌分类的准确率和FID分数。

Comments This paper has been accepted for the IEEE International Symposium on Biomedical Imaging (ISBI) 2026, London, UK, and will be presented in the corresponding session

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.03502 2026-01-27 eess.IV cs.AI cs.GR 79%

DC-VSR: Spatially and Temporally Consistent Video Super-Resolution with Video Diffusion Prior

DC-VSR: 基于视频扩散先验的时空一致视频超分辨率

Janghyeok Han, Gyujin Sim, Geonung Kim, Hyun-seung Lee, Kyuha Choi, Youngseok Han, Sunghyun Cho

机构 * POSTECH Visual Display Business, Samsung Electronics(视觉显示业务,三星电子)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.GR

AI总结 DC-VSR通过引入空间和时间注意力传播机制及细节抑制自注意力引导,实现了视频超分辨率的时空一致性与高质量纹理恢复。

Comments Equal contributions from first two authors

Journal ref In Proceedings of the Special Interest Group on Computer Graphics and Interactive Techniques Conference Conference Papers (SIGGRAPH Conference Papers 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18615 2026-01-27 cs.LG 78%

Geometry-Free Conditional Diffusion Modeling for Solving the Inverse Electrocardiography Problem

无几何条件的条件扩散建模用于解决逆电生理图问题

Ramiro Valdes Jara, Adam Meyers

机构 * Department of Industrial and Systems Engineering(工业与系统工程系) University of Miami(迈阿密大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出了一种无几何条件的条件扩散模型,用于解决逆电生理图问题,通过数据驱动的方法提高心脏电生理成像的重建精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18069 2026-01-27 cs.NI cs.AI 78%

Diffusion Model-based Reinforcement Learning for Version Age of Information Scheduling: Average and Tail-Risk-Sensitive Control

基于扩散模型的强化学习用于版本信息年龄调度:平均与尾风险敏感控制

Haoyuan Pan, Sizhao Chen, Zhaorui Wang, Tse-Tin Chan

机构 * College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) School of Computing and Information Technology, Great Bay University(广东大湾区大学计算机与信息科技学院) Department of Mathematics and Information Technology, The Education University of Hong Kong(香港教育大学数学与信息技术系)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出基于扩散模型的强化学习方法,用于多用户无线系统中平均和尾风险敏感的版本信息年龄调度,通过分布性critic和扩散-based actor实现稳健的尾风险优化。

Comments 16 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14784 2026-01-27 eess.AS 78%

MELA-TTS: Joint transformer-diffusion model with representation alignment for speech synthesis

MELA-TTS:联合变压器-扩散模型与表征对齐用于语音合成

Keyu An, Zhiyu Zhang, Changfeng Gao, Yabin Li, Zhendong Peng, Haoxu Wang, Zhihao Du, Han Zhao, Zhifu Gao, Xiangang Li

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 MELA-TTS通过联合变压器-扩散模型与表征对齐,实现端到端文本到语音合成,提升连续特征建模效率和跨模态一致性。

Comments accepted by ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07363 2026-01-27 cs.IT math.IT 78%

Knowledge Distillation Driven Semantic NOMA for Image Transmission with Diffusion Model

基于扩散模型的语义NOMA知识蒸馏图像传输

Qifei Wang, Zhen Gao, Shuo Sun, Zhijin Qin, Xiaodong Xu, Meixia Tao

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出KDD-SemNOMA方法,结合知识蒸馏和扩散模型,提升多用户无线图像传输的鲁棒性和图像恢复质量。

Comments 16 pages, accepted by IEEE Transactions on Wireless Communications

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19738 2026-01-27 math.NA cs.NA 78%

Numerical Identification of a Time-Dependent Coefficient in a Time-Fractional Diffusion Equation with Integral Constraints

时间分数扩散方程中时间依赖系数的数值识别

Arshyn Altybay

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出了一种高效算法用于时间分数扩散方程中时间依赖系数的数值识别,通过积分形式和有限差分方案实现了稳定性与收敛性的分析,并验证了算法在噪声数据下的鲁棒性。

Journal ref Zeitschrift fur Angewandte Mathematik und Physik, 77, 41 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17902 2026-01-27 cs.SD eess.AS 78%

dLLM-ASR: A Faster Diffusion LLM-based Framework for Speech Recognition

dLLM-ASR:一种更高效的基于扩散大语言模型的语音识别框架

Wenjie Tian, Bingshen Mu, Guobin Ma, Xuelong Geng, Zhixian Zhao, Lei Xie

机构 * Northwestern Polytechnical University(西北工业大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 dLLM-ASR通过将扩散大语言模型的解码过程转化为先验引导的去噪流程,实现高效的语音识别,达到与自回归模型相当的准确率并提升4.44倍推理速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17375 2026-01-27 math.NA cs.NA 78%

Operator splitting based diffusion samplers and improved convergence analysis

基于运算符分裂的扩散采样器及改进的收敛性分析

Peiyi Liu, Zhaoqiang Liu, Yiqi Gu

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出基于运算符分裂的扩散采样器,并改进收敛性分析,建立了更精确的误差界,验证了误差与步长的二次关系。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17224 2026-01-27 cs.LG 78%

Parameter Inference and Uncertainty Quantification with Diffusion Models: Extending CDI to 2D Spatial Conditioning

基于扩散模型的参数推断与不确定性量化:扩展CDI以支持二维空间条件

Dmitrii Torbunov, Yihui Ren, Lijun Wu, Yimei Zhu

机构 * Brookhaven National Laboratory(布鲁克海文国家实验室)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出基于扩散模型的CDI方法,扩展至二维空间条件,用于CBED参数推断,提供更准确的不确定性量化。

详情

展开后加载摘要…

URL PDF HTML 收藏