arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-08-14 至 2025-08-14 共收录 63 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 4 篇

2508.09575 2025-08-14 cs.CV 89%

Dual Recursive Feedback on Generation and Appearance Latents for Pose-Robust Text-to-Image Diffusion

Jiwon Kim, Pureum Kim, SeonHwa Kim, Soobin Park, Eunju Cha, Kyong Hwan Jin

机构 * Korea University(韩国大学) Sookmyung Women’s University(肃明女子大学)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09293 2025-08-14 cs.CY cs.AI 78%

Ethical Medical Image Synthesis

Weina Jin, Ashish Sinha, Kumar Abhishek, Ghassan Hamarneh

专题命中 文生图 :image synthesis(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09919 2025-08-14 eess.IV cs.AI cs.CV 57%

T-CACE: A Time-Conditioned Autoregressive Contrast Enhancement Multi-Task Framework for Contrast-Free Liver MRI Synthesis, Segmentation, and Diagnosis

Xiaojiao Xiao, Jianfeng Zhao, Qinmin Vivian Hu, Guanghui Wang

机构 * Department of Computer Science, Toronto Metropolitan University(计算机科学系,多伦多 Metropolitan 大学) School of Biomedical Engineering, Western University(生物医学工程学院,西部大学)

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments IEEE Journal of Biomedical and Health Informatics, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12341 2025-08-14 cs.MM 57%

Multimodal LLM-based Query Paraphrasing for Video Search

Jiaxin Wu, Chong-Wah Ngo, Wing-Kwong Chan, Sheng-Hua Zhong, Xiong-Yong Wei, Qing Li

专题命中 文生图 :text-to-image(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 扩散模型 42 篇

2508.09903 2025-08-14 quant-ph 88%

Hybrid Quantum-Classical Latent Diffusion Models for Medical Image Generation

Kübra Yeter-Aydeniz, Nora M. Bauer, Pranay Jain, Max Masnick

专题命中 扩散模型 :image generation(title,abstract);diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09385 2025-08-14 cs.LG cs.AI 86%

Understanding Dementia Speech Alignment with Diffusion-Based Image Generation

Mansi, Anastasios Lepipas, Dominika Woszczyk, Yiying Guan, Soteris Demetriou

机构 * Imperial College London(帝国理工学院伦敦校区)

专题命中 扩散模型 :image generation(title);diffusion(title);text-to-image(abstract)

Comments Paper accepted at Interspeech 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09968 2025-08-14 cs.LG cs.CV 83%

Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models

Luca Eyring, Shyamgopal Karthik, Alexey Dosovitskiy, Nataniel Ruiz, Zeynep Akata

机构 * Technical University of Munich(慕尼黑技术大学) Munich Center of Machine Learning(慕尼黑机器学习中心) Helmholtz Munich(海德堡-慕尼黑 Helmholtz 中心) University of Tübingen(图宾根大学) Inceptive(Inceptive 公司) Google(谷歌公司)

专题命中 扩散模型 :diffusion(title,abstract);generative vision(abstract);分类 cs.CV

Comments Project page: https://noisehypernetworks.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15290 2025-08-14 cs.GR cs.AI cs.CV cs.HC 81%

Human Motion Capture from Loose and Sparse Inertial Sensors with Garment-aware Diffusion Models

Andela Ilic, Jiaxi Jiang, Paul Streli, Xintong Liu, Christian Holz

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV、cs.GR

Comments Accepted by IJCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09949 2025-08-14 cs.CV cs.LG 79%

Stable Diffusion Models are Secretly Good at Visual In-Context Learning

Trevine Oorloff, Vishwanath Sindagi, Wele Gedara Chaminda Bandara, Ali Shafahi, Amin Ghiasi, Charan Prakash, Reza Ardekani

机构 * Apple(苹果公司) University of Maryland - College Park(马里兰大学-College Park)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09943 2025-08-14 cs.CV 79%

AST-n: A Fast Sampling Approach for Low-Dose CT Reconstruction using Diffusion Models

Tomás de la Sotta, José M. Saavedra, Héctor Henríquez, Violeta Chang, Aline Xavier

机构 * Universidad de los Andes(安第斯大学) Universidad de Santiago de Chile(圣地亚哥大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09847 2025-08-14 cs.CV 79%

Enhancing Diffusion Face Generation with Contrastive Embeddings and SegFormer Guidance

Dhruvraj Singh Rawat, Enggen Sherpa, Rishikesan Kirupanantha, Tin Hoang

机构 * University of Surrey, UK(英国萨里大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments 10 pages, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09709 2025-08-14 cs.CV 79%

MangaDiT: Reference-Guided Line Art Colorization with Hierarchical Attention in Diffusion Transformers

Qianru Qiu, Jiafeng Mao, Kento Masui, Xueting Wang

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Codes and benchmarks will be released soon

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09667 2025-08-14 cs.CV 79%

GSFixer: Improving 3D Gaussian Splatting with Reference-Guided Video Diffusion Priors

Xingyilang Yin, Qi Zhang, Jiahao Chang, Ying Feng, Qingnan Fan, Xi Yang, Chi-Man Pun, Huaqi Zhang, Xiaodong Cun

机构 * University of Macau(澳门大学) VIVO CUHKSZ(香港中文大学) Xidian University(西安电子科技大学) GVC Lab, Great Bay University(大澳大学GVC实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08173 2025-08-14 cs.CV eess.IV 79%

CD-TVD: Contrastive Diffusion for 3D Super-Resolution with Scarce High-Resolution Time-Varying Data

Chongke Bi, Xin Gao, Jiangkang Deng, Guan Li, Jun Han

机构 * College of Intelligence and Computing, Tianjin University(智能与计算学院,天津大学) Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心) University of Chinese Academy of Sciences(中国科学院大学) Division of Emerging Interdisciplinary Areas and Center for Ocean Research in Hong Kong and Macau (CORE), The Hong Kong University of Science and Technology(新兴交叉领域 division 和香港澳门海洋研究中心(CORE),香港科学与技术大学)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

Comments Accepted to IEEE VIS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01105 2025-08-14 cs.CV 79%

LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer

Yiren Song, Danze Chen, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学展示实验室)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.20771 2025-08-14 cs.CR cs.AI cs.CV cs.LG 79%

Towards Black-Box Membership Inference Attack for Diffusion Models

Jingwei Li, Jing Dong, Tianxing He, Jingzhao Zhang

机构 * Institute for Interdisciplinary Information Sciences, Tsinghua University(清华大学交叉信息研究院) The Chinese University of Hong Kong(香港中文大学) Shanghai Qizhi Institute(上海启智研究院)

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09444 2025-08-14 cs.RO cs.CV 79%

DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation

Haoxiang Shi, Xiang Deng, Zaijing Li, Gongwei Chen, Yaowei Wang, Liqiang Nie

专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09924 2025-08-14 cond-mat.stat-mech 78%

Active Particle Diffusion in Convection Roll Arrays

Pulak Kumar Ghosh, Fabio Marchesoni, Yunyun Li, Franco Nori

专题命中 扩散模型 :diffusion(title,abstract)

Journal ref Phys. Chem. Chem. Phys. 23, 11944(2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.24042 2025-08-14 cs.LG cs.NA math.NA math.ST stat.ML stat.TH 78%

Faster Diffusion Models via Higher-Order Approximation

Gen Li, Yuchen Zhou, Yuting Wei, Yuxin Chen

机构 * Department of Statistics, Chinese University of Hong Kong(香港中文大学统计学系) Department of Statistics, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校统计学系) Department of Statistics and Data Science, the Wharton School, University of Pennsylvania(宾夕法尼亚大学沃顿商学院统计学与数据科学系)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.00586 2025-08-14 cs.RO cs.LG 78%

ParkDiffusion: Heterogeneous Multi-Agent Multi-Modal Trajectory Prediction for Automated Parking using Diffusion Models

Jiarong Wei, Niclas Vödisch, Anna Rehr, Christian Feist, Abhinav Valada

机构 * Department of Computer Science, University of Freiburg(弗赖堡大学计算机科学系) CARIAD SE(CARIAD公司)

专题命中 扩散模型 :diffusion(title,abstract)

Comments IROS 2025 Camera-Ready Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09192 2025-08-14 cs.LG cs.AI 78%

Diffusion LLMs Can Do Faster-Than-AR Inference via Discrete Diffusion Forcing

Xu Wang, Chenkai Xu, Yijie Jin, Jiachun Jin, Hao Zhang, Zhijie Deng

机构 * Shanghai Jiao Tong University(上海交通大学) University of California San Diego(加州大学圣地亚哥分校) Shanghai University(上海大学)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09164 2025-08-14 cs.LG 78%

Generating Feasible and Diverse Synthetic Populations Using Diffusion Models

Min Tang, Peng Lu, Qing Feng

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01006 2025-08-14 cs.LG stat.ML 78%

Underdamped Diffusion Bridges with Applications to Sampling

Denis Blessing, Julius Berner, Lorenz Richter, Gerhard Neumann

机构 * Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院) NVIDIA(NVIDIA公司) Zuse Institute Berlin(柏林泽尼克研究所) dida Datenschmiede GmbH(dida数据安全公司) FZI Research Center for Information Technology(信息技术研究中心)

专题命中 扩散模型 :diffusion(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12724 2025-08-14 cs.RO 78%

Responsive Noise-Relaying Diffusion Policy: Responsive and Efficient Visuomotor Control

Zhuoqun Chen, Xiu Yuan, Tongzhou Mu, Hao Su

机构 * UC San Diego(UC圣地亚哥大学)

专题命中 扩散模型 :diffusion(title,abstract)

Comments Project website: https://rnr-dp.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04684 2025-08-14 q-bio.PE cond-mat.stat-mech physics.bio-ph physics.data-an q-bio.QM 78%

Strong anomalous diffusion for free-ranging birds

Ohad Vilk, Motti Charter, Sivan Toledo, Eli Barkai, Ran Nathan

专题命中 扩散模型 :diffusion(title,abstract)

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.05290 2025-08-14 cs.LG eess.IV eess.SP 78%

Removing Structured Noise with Diffusion Models

Tristan S. W. Stevens, Hans van Gorp, Faik C. Meral, Junseob Shin, Jason Yu, Jean-Luc Robert, Ruud J. G. van Sloun

机构 * Department of Electrical Engineering, Eindhoven University of Technology(埃因霍温理工大学电子工程系) Philips Research North America(飞利浦美国研究公司)

专题命中 扩散模型 :diffusion(title,abstract)

Comments 20 pages, 8 figures, Transactions on Machine Learning Research

Journal ref Transactions on Machine Learning Research (2025): 2835-8856

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09936 2025-08-14 cs.CV cs.DL 57%

Quo Vadis Handwritten Text Generation for Handwritten Text Recognition?

Vittorio Pippi, Konstantina Nikolaidou, Silvia Cascianelli, George Retsinas, Giorgos Sfikas, Rita Cucchiara, Marcus Liwicki

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) Luleå University of Technology(吕勒奥技术大学) National Technical University of Athens(雅典国家技术大学) University of West Attica(西阿提卡大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

Comments Accepted at ICCV Workshop VisionDocs

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09858 2025-08-14 cs.CV 57%

HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics

Weiqi Li, Zehao Zhang, Liang Lin, Guangrun Wang

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09746 2025-08-14 cs.CV cs.AI 57%

Region-to-Region: Enhancing Generative Image Harmonization with Adaptive Regional Injection

Zhiqiu Zhang, Dongqi Fan, Mingjie Wang, Qiang Tang, Jian Yang, Zili Yi

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.19860 2025-08-14 cs.CV 57%

CoherenDream: Boosting Holistic Text Coherence in 3D Generation via Multimodal Large Language Models Feedback

Chenhan Jiang, Yihan Zeng, Dit-Yan Yeung

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏