arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-07-28 至 2025-07-28 共收录 45 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 29 篇

2408.09364 2025-07-28 math.PR 50%

Birth-death processes are time-changed Feller's Brownian motions

Liping Li

专题命中 扩散模型 :diffusion(abstract)

Journal ref Stochastic Processes and their Applications, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16642 2025-07-28 math.AP 50%

A stable-compact method for qualitative properties of semilinear elliptic equations

Henri Berestycki, Cole Graham

专题命中 扩散模型 :diffusion(abstract)

Comments 54 pages

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 可控生成 4 篇

2503.15557 2025-07-28 cs.GR cs.CV cs.RO 62%

Motion Synthesis with Sparse and Flexible Keyjoint Control

Inwoo Hwang, Jinseok Bae, Donggeun Lim, Young Min Kim

机构 * ECE, Seoul National University(电子工程系,首尔国立大学)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV、cs.GR

Comments Accepted to ICCV 2025. Project Page: http://inwoohwang.me/SFControl

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19071 2025-07-28 cs.CV 57%

Cross-Subject Mind Decoding from Inaccurate Representations

Yangyang Xu, Bangzhen Liu, Wenqi Shao, Yong Du, Shengfeng He, Tingting Zhu

机构 * The University of Oxford(牛津大学) South China University of Technology(华南理工大学) Shanghai AI Lab(上海人工智能实验室) Ocean University of China(中国海洋大学) Singapore Management University(新加坡管理大学)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15265 2025-07-28 cs.CV cs.AI cs.CR 57%

Blind Spot Navigation: Evolutionary Discovery of Sensitive Semantic Concepts for LVLMs

Zihao Pan, Yu Tong, Weibin Wu, Jingyi Wang, Lifeng Chen, Zhe Zhao, Jiajia Wei, Yitong Qiao, Zibin Zheng

专题命中 可控生成 :text-to-image(abstract);分类 cs.CV

Comments The paper needs major revisions, so it is being withdrawn

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13807 2025-07-28 cs.CV 57%

MagicDrive-V2: High-Resolution Long Video Generation for Autonomous Driving with Adaptive Control

Ruiyuan Gao, Kai Chen, Bo Xiao, Lanqing Hong, Zhenguo Li, Qiang Xu

机构 * CUHK(中文大学) HKUST(香港科技大学) Huawei Cloud(华为云) Huawei Noah’s Ark Lab(华为诺亚实验室)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments ICCV 2025 camera-ready version, Project Website: https://flymin.github.io/magicdrive-v2/

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 图像修复 2 篇

2507.19138 2025-07-28 eess.IV cs.CV 79%

RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution

Weisong Zhao, Jingkai Zhou, Xiangyu Zhu, Weihua Chen, Xiao-Yu Zhang, Zhen Lei, Fan Wang

专题命中 图像修复 :diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09476 2025-07-28 physics.flu-dyn cs.LG 50%

Mean flow data assimilation using physics-constrained Graph Neural Networks

M. Quattromini, M. A. Bucci, S. Cherubini, O. Semeraro

机构 * Dipartimento di Meccanica, Matematica e Management Politecnico di Bari(巴里理工学院机械、数学与管理系) LISN-CNRS Université Paris-Saclay(巴黎萨克雷大学LISN-CNRS) Digital Sciences & Technologies Department SafranTech(SafranTech数字科学与技术部门)

专题命中 图像修复 :inpainting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 个性化与一致性 1 篇

2505.04963 2025-07-28 cs.CV 83%

ViCTr: Vital Consistency Transfer for Pathology Aware Image Synthesis

Onkar Susladkar, Gayatri Deshmukh, Yalcin Tur, Gorkhem Durak, Ulas Bagci

机构 * Northwestern University(西北大学) Stanford University(斯坦福大学)

专题命中 个性化与一致性 :image synthesis(title,abstract);diffusion(abstract);分类 cs.CV

Comments Accepted in ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 图像生成评测 3 篇

2507.19002 2025-07-28 cs.CV 83%

Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment

Ying Ba, Tianyu Zhang, Yalong Bai, Wenyi Mo, Tao Liang, Bing Su, Ji-Rong Wen

机构 * Gaoling School of Artificial Intelligence(中关村人工智能学院) Renmin University of China(中国人民大学) Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大模型与智能治理重点实验室) Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程研究中心,教育部) iN2X

专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12612 2025-07-28 cs.CL cs.CR 82%

T2ISafety: Benchmark for Assessing Fairness, Toxicity, and Privacy in Image Generation

Lijun Li, Zhelun Shi, Xuhao Hu, Bowen Dong, Yiran Qin, Xihui Liu, Lu Sheng, Jing Shao

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Beihang University(北航) Harbin Institute of Technology(哈尔滨工业大学) Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)) The University of Hong Kong(香港大学)

专题命中 图像生成评测 :image generation(title);text-to-image(abstract);diffusion(abstract)

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11493 2025-07-28 cs.CV 79%

GIE-Bench: Towards Grounded Evaluation for Text-Guided Image Editing

Yusu Qian, Jiasen Lu, Tsu-Jui Fu, Xinze Wang, Chen Chen, Yinfei Yang, Wenze Hu, Zhe Gan

机构 * Apple(苹果公司)

专题命中 图像生成评测 :image editing(title,abstract);分类 cs.CV

Comments Project page: https://sueqian6.github.io/GIE-Bench-web/

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 效率与蒸馏 2 篇

2507.18192 2025-07-28 cs.CV 70%

TeEFusion: Blending Text Embeddings to Distill Classifier-Free Guidance

Minghao Fu, Guo-Hua Wang, Xiaohao Chen, Qing-Guo Chen, Zhao Xu, Weihua Luo, Kaifu Zhang

机构 * School of Artificial Intelligence, Nanjing University(南京大学人工智能学院) National Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家实验室) Alibaba International Digital Commerce Group(阿里巴巴国际数字商业集团)

专题命中 效率与蒸馏 :text-to-image(abstract);image synthesis(abstract);分类 cs.CV

Comments Accepted by ICCV 2025. The code is publicly available at https://github.com/AIDC-AI/TeEFusion

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19330 2025-07-28 physics.app-ph cond-mat.mtrl-sci 50%

Efficient integration of self-assembled organic monolayer tunnel barriers in large area pinhole-free magnetic tunnel junctions

Maryam S. Dehaghani, Sophie Guézo Aguinet, Arnaud Le Pottier, Soraya Ababou-Girard, Rozenn Bernard, Sylvain Tricot, Philippe Schieffer, Bruno Lépine, Francine Solal, Pascal Turban

专题命中 效率与蒸馏 :diffusion(abstract)

Comments 25 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

7. 其他图像生成 1 篇

2507.18640 2025-07-28 cs.HC cs.AI cs.CV 57%

How good are humans at detecting AI-generated images? Learnings from an experiment

Thomas Roca, Anthony Cintron Roman, Jehú Torres Vega, Marcelo Duarte, Pengce Wang, Kevin White, Amit Misra, Juan Lavista Ferres

机构 * Microsoft AI for Good Lab(微软AI for Good实验室)

专题命中 其他图像生成 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏