arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-11-18 至 2025-11-18 共收录 14 信号源:cs.CV, cs.GR, cs.MM

1. 可控生成 14 篇

2511.11693 2025-11-18 cs.AI cs.CR cs.CV cs.LG 87%

Value-Aligned Prompt Moderation via Zero-Shot Agentic Rewriting for Safe Image Generation

Xin Zhao, Xiaojun Chen, Bingshan Liu, Zeyao Liu, Zhendong Zhao, Xiaoyan Gu

专题命中 可控生成 :image generation(title,abstract);text-to-image(abstract);diffusion(abstract);generative vision(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10825 2025-11-18 cs.CV 79%

OmniVDiff: Omni Controllable Video Diffusion for Generation and Understanding

Dianbing Xi, Jiepeng Wang, Yuanzhi Liang, Xi Qiu, Yuchi Huo, Rui Wang, Chi Zhang, Xuelong Li

专题命中 可控生成 :diffusion(title,abstract);分类 cs.CV

Comments Accepted by AAAI 2026. Our project page: https://tele-ai.github.io/OmniVDiff/

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12032 2025-11-18 cs.CV 79%

Improved Masked Image Generation with Knowledge-Augmented Token Representations

Guotao Liang, Baoquan Zhang, Zhiyuan Wen, Zihao Han, Yunming Ye

专题命中 可控生成 :image generation(title,abstract);分类 cs.CV

Comments AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15222 2025-11-18 cs.AI cs.CV cs.MA 70%

See it. Say it. Sorted: Agentic System for Compositional Diagram Generation

Hantao Zhang, Jingyang Liu, Ed Li

机构 * Yale University(耶鲁大学) University of Edinburgh(爱丁堡大学)

专题命中 可控生成 :image generation(abstract);diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00298 2025-11-18 cs.CV 70%

AniMer+: Unified Pose and Shape Estimation Across Mammalia and Aves via Family-Aware Transformer

Liang An, Jin Lyu, Li Lin, Pujin Cheng, Yebin Liu, Xiaoying Tang

机构 * Department of Electronic and Electrical Engineering, Southern University of Science and Technology, Shenzhen, China(南方科技大学电子与电气工程系) Department of Automation, Tsinghua University, Beijing, China(清华大学自动化系) Jiaxing Research Institute, Southern University of Science and Technology, Jiaxing, China(南方科技大学嘉兴研究所) Department of Electrical and Electronic Engineering, the University of Hong Kong, Hong Kong, China(香港大学电子与电气工程系)

专题命中 可控生成 :image generation(abstract);diffusion(abstract);分类 cs.CV

Comments Accepted to TPAMI2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13684 2025-11-18 cs.CV cs.LG 57%

Training-Free Multi-View Extension of IC-Light for Textual Position-Aware Scene Relighting

Jiangnan Ye, Jiedong Zhuang, Lianrui Mu, Wenjie Zheng, Jiaqi Hu, Xingze Zou, Jing Wang, Haoji Hu

机构 * College of Information Science(信息科学学院) Electronic Engineering, Zhejiang University, Hangzhou, 310027, Zhejiang, China(电子工程系,浙江大学,杭州,310027,浙江,中国)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Submitting for Neurocomputing

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13191 2025-11-18 cs.CV 57%

Birth of a Painting: Differentiable Brushstroke Reconstruction

Ying Jiang, Jiayin Lu, Yunuo Chen, Yumeng He, Kui Wu, Yin Yang, Chenfanfu Jiang

专题命中 可控生成 :image generation(abstract);分类 cs.CV

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26196 2025-11-18 cs.CV 57%

Sketch2PoseNet: Efficient and Generalized Sketch to 3D Human Pose Prediction

Li Wang, Yiyu Zhuang, Yanwen Wang, Xun Cao, Chuan Guo, Xinxin Zuo, Hao Zhu

机构 * Nanjing University(南京大学) Snap Inc.(Snap公司) Concordia University(Concordia大学)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments SIGGRAPH Asia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19946 2025-11-18 cs.CV 57%

SCALAR: Scale-wise Controllable Visual Autoregressive Learning

Ryan Xu, Dongyang Jin, Yancheng Bai, Rui Lan, Xu Duan, Lei Sun, Xiangxiang Chu

机构 * Alibaba Group(阿里巴巴集团)

专题命中 可控生成 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15095 2025-11-18 cs.CV 57%

VistaDepth: Improving far-range Depth Estimation with Spectral Modulation and Adaptive Reweighting

Mingxia Zhan, Li Zhang, Yingjie Wang, Xiaomeng Chu, Beibei Wang, Yanyong Zhang

机构 * Hefei University of Technology(合肥工业大学) University of Science and Technology of China(中国科学技术大学) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(人工智能研究院,合肥综合性国家科学中心)

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12030 2025-11-18 cs.CV 57%

VPHO: Joint Visual-Physical Cue Learning and Aggregation for Hand-Object Pose Estimation

Jun Zhou, Chi Xu, Kaifeng Tang, Yuting Ge, Tingrui Guo, Li Cheng

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments 14 pages, 9 figures, extended version of the AAAI 2026 paper "VPHO: Joint Visual-Physical Cue Learning and Aggregation for Hand-Object Pose Estimation"

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.06421 2025-11-18 cs.CV cs.AI cs.LG 57%

Using Self-Supervised Auxiliary Tasks to Improve Fine-Grained Facial Representation

Mahdi Pourmirzaei, Gholam Ali Montazer, Farzaneh Esmaili

专题命中 可控生成 :inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26367 2025-11-18 math.AP math-ph math.MP math.SP 50%

Competition of small targets in planar domains: from Dirichlet to Robin and Steklov boundary condition

Denis S. Grebenkov, Michael J. Ward

专题命中 可控生成 :diffusion(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11909 2025-11-18 q-fin.MF 50%

Modeling and Stabilizing Financial Systemic Risk Using Optimal Control Theory

Jiacheng Wu

专题命中 可控生成 :diffusion(abstract)

Comments 21 pages

详情

展开后加载摘要…

URL PDF HTML 收藏