arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-07-29 至 2025-07-29 共收录 13 信号源:cs.CV, cs.GR, cs.MM

1. 可控生成 13 篇

2507.19939 2025-07-29 cs.CV 88%

LLMControl: Grounded Control of Text-to-Image Diffusion-based Synthesis with Multimodal LLMs

Jiaze Wang, Rui Chen, Haowang Cui

机构 * Tianjin University(天津大学)

专题命中 可控生成 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20855 2025-07-29 cs.CV 57%

Compositional Video Synthesis by Temporal Object-Centric Learning

Adil Kaan Akan, Yucel Yemez

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments 12+21 pages, submitted to IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI), currently under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19950 2025-07-29 cs.CV cs.AI 57%

RARE: Refine Any Registration of Pairwise Point Clouds via Zero-Shot Learning

Chengyu Zheng, Jin Huang, Honghua Chen, Mingqiang Wei

机构 * College of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大学计算机科学与技术学院) Shenzhen Research Institute, Nanjing University of Aeronautics and Astronautics(南京航空航天大学深圳研究院) School of Data Science, Lingnan University(岭南大学数据科学学院)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21771 2025-07-29 cs.CV 57%

A Unified Image-Dense Annotation Generation Model for Underwater Scenes

Hongkai Lin, Dingkang Liang, Zhenghao Qi, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学)

专题命中 可控生成 :text-to-image(abstract);分类 cs.CV

Comments Accepted by CVPR 2025. The code is available at https://github.com/HongkLin/TIDE

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08714 2025-07-29 cs.CV cs.AI 57%

Versatile Multimodal Controls for Expressive Talking Human Animation

Zheng Qin, Ruobing Zheng, Yabing Wang, Tianqi Li, Zixin Zhu, Sanping Zhou, Ming Yang, Le Wang

机构 * National Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Xi'an Jiaotong University, Ant Group(人机混合增强智能国家级实验室,西安交通大学,蚂蚁集团) National Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Xi'an Jiaotong University(人机混合增强智能国家级实验室,西安交通大学) University at Buffalo(布法罗大学)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Accepted by ACM MM2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06438 2025-07-29 cs.CV 57%

Qffusion: Controllable Portrait Video Editing via Quadrant-Grid Attention Learning

Maomao Li, Lijian Lin, Yunfei Liu, Ye Zhu, Yu Li

机构 * International Digital Economy Academy (IDEA)(国际数字经济学院)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments 19 pages

Journal ref TVCG 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19644 2025-07-29 math.OC cs.NA math.NA 50%

Hierarchical clustering and dimensional reduction for optimal control of large-scale agent-based models

Angela Monti, Fasma Diele, Dante Kalise

专题命中 可控生成 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20868 2025-07-29 physics.flu-dyn 50%

Passive control of wing tip vortices through a grooved-tip design

Junchen Tan, Shūji Ōtomo, Ignazio Maria Viola, Yabin Liu

专题命中 可控生成 :diffusion(abstract)

Comments Submitted to Experiments in Fluids

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19910 2025-07-29 eess.SP 50%

Toward Dual-Functional LAWN: Control-Aware System Design for Aerodynamics-Aided UAV Formations

Jun Wu, Weijie Yuan, Qingqing Cheng, Haijia Jin

专题命中 可控生成 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09328 2025-07-29 math.DS 50%

Spatiotemporal SEIQR Epidemic Modeling with Optimal Control for Vaccination, Treatment, and Social Measures

Achraf Zinihi, Matthias Ehrhardt, Moulay Rchid Sidi Ammi

专题命中 可控生成 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09524 2025-07-29 cs.RO 50%

FlowNav: Combining Flow Matching and Depth Priors for Efficient Navigation

Samiran Gode, Abhijeet Nayak, Débora N. P. Oliveira, Michael Krawez, Cordelia Schmid, Wolfram Burgard

机构 * Artificial Intelligence and Robotics Lab, Department of Computer Science and Artificial Intelligence, University of Technology Nuremberg(人工智能与机器人实验室,计算机科学与人工智能系,图腾大学) Inria, Ecole Normale Supérieure, CNRS, PSL Research University(法国国家信息与自动化研究所,巴黎高等师范学校,国家科学研究中心,巴黎综合理工研究学院)

专题命中 可控生成 :diffusion(abstract)

Comments Accepted to IROS'25. Previous version accepted at CoRL 2024 workshop on Learning Effective Abstractions for Planning (LEAP) and workshop on Differentiable Optimization Everywhere: Simulation, Estimation, Learning, and Control

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03655 2025-07-29 cs.LG cs.AI 50%

Geometric Representation Condition Improves Equivariant Molecule Generation

Zian Li, Cai Zhou, Xiyuan Wang, Xingang Peng, Muhan Zhang

机构 * Institute for Artificial Intelligence, Peking University, Beijing, China(北京大学人工智能研究院) School of Intelligence Science and Technology, Peking University, Beijing, China(北京大学智能科学与技术学院) Department of Electrical Engineering and Computer Science, Massachusetts Institute of Technology, Cambridge, MA, USA(麻省理工学院电子工程与计算机科学系) Department of Automation, Tsinghua University, Beijing, China(清华大学自动化系)

专题命中 可控生成 :diffusion(abstract)

Comments Accepted to ICML 2025 as a Spotlight Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09722 2025-07-29 cs.LG cs.SY eess.SY stat.ML 50%

The Pitfalls of Imitation Learning when Actions are Continuous

Max Simchowitz, Daniel Pfrommer, Ali Jadbabaie

机构 * CMU(卡内基梅隆大学) MIT(麻省理工学院)

专题命中 可控生成 :diffusion(abstract)

Comments 98 pages, 2 figures, updated proof sketch

详情

展开后加载摘要…

URL PDF HTML 收藏