arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 4233 信号源:cs.CV, cs.GR, cs.MM

1. 可控生成 4233 篇

2402.17214 2024-07-11 cs.CV 57%

CharacterGen: Efficient 3D Character Generation from Single Images with Multi-View Pose Canonicalization

Hao-Yang Peng, Jia-Peng Zhang, Meng-Hao Guo, Yan-Pei Cao, Shi-Min Hu

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.06304 2024-07-10 cs.CV cs.AI cs.CL 57%

VIMI: Grounding Video Generation through Multi-modal Instruction

Yuwei Fang, Willi Menapace, Aliaksandr Siarohin, Tsai-Shien Chen, Kuan-Chien Wang, Ivan Skorokhodov, Graham Neubig, Sergey Tulyakov

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.03980 2024-07-10 cs.CV 57%

Continual Learning for LiDAR Semantic Segmentation: Class-Incremental and Coarse-to-Fine strategies on Sparse Data

Elena Camuffo, Simone Milani

专题命中 可控生成 :inpainting(abstract);分类 cs.CV

Journal ref International Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16849 2024-07-09 cs.CV 57%

Sync4D: Video Guided Controllable Dynamics for Physics-Based 4D Generation

Zhoujie Fu, Jiacheng Wei, Wenhao Shen, Chaoyue Song, Xiaofeng Yang, Fayao Liu, Xulei Yang, Guosheng Lin

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Our project page: https://sync4dphys.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14868 2024-07-08 cs.CV cs.AI cs.LG cs.RO 57%

Generative Camera Dolly: Extreme Monocular Dynamic Novel View Synthesis

Basile Van Hoorick, Rundi Wu, Ege Ozguroglu, Kyle Sargent, Ruoshi Liu, Pavel Tokmakov, Achal Dave, Changxi Zheng, Carl Vondrick

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Accepted to ECCV 2024. Project webpage is available at: https://gcd.cs.columbia.edu/

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.03355 2024-07-08 cs.CV cs.AI cs.LG 57%

SegGen: Supercharging Segmentation Models with Text2Mask and Mask2Img Synthesis

Hanrong Ye, Jason Kuen, Qing Liu, Zhe Lin, Brian Price, Dan Xu

专题命中 可控生成 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.01456 2024-07-04 cs.CV 57%

ColorizeDiffusion: Adjustable Sketch Colorization with Reference Image and Text

Dingkun Yan, Liang Yuan, Erwin Wu, Yuma Nishioka, Issei Fujishiro, Suguru Saito

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.19796 2024-07-01 eess.IV cs.CV 57%

Comprehensive Generative Replay for Task-Incremental Segmentation with Concurrent Appearance and Semantic Forgetting

Wei Li, Jingyang Zhang, Pheng-Ann Heng, Lixu Gu

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Accepted by MICCAI24

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18570 2024-06-28 cs.HC cs.AI cs.CV 57%

It's a Feature, Not a Bug: Measuring Creative Fluidity in Image Generators

Aditi Ramaswamy, Melane Navaratnarajah, Hana Chockler

专题命中 可控生成 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18198 2024-06-27 cs.CV 57%

VDG: Vision-Only Dynamic Gaussian for Driving Simulation

Hao Li, Jingfeng Li, Dingwen Zhang, Chenming Wu, Jieqi Shi, Chen Zhao, Haocheng Feng, Errui Ding, Jingdong Wang, Junwei Han

专题命中 可控生成 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17541 2024-06-26 cs.CV 57%

Principal Component Clustering for Semantic Segmentation in Synthetic Data Generation

Felix Stillger, Frederik Hasecke, Tobias Meisen

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments This is a technical report for a submission to the CVPR "SyntaGen - Harnessing Generative Models for Synthetic Visual Datasets" workshop challenge. The report is already uploaded to the workshop's homepage https://syntagen.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03443 2024-06-25 cs.CV stat.ML 57%

Data-driven Crop Growth Simulation on Time-varying Generated Images using Multi-conditional Generative Adversarial Networks

Lukas Drees, Dereje T. Demie, Madhuri R. Paul, Johannes Leonhardt, Sabine J. Seidel, Thomas F. Döring, Ribana Roscher

专题命中 可控生成 :image generation(abstract);分类 cs.CV

Comments 26 pages, 16 figures, code available at https://github.com/luked12/crop-growth-cgan

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.11293 2024-06-25 cs.CV eess.IV 57%

Beyond the Field-of-View: Enhancing Scene Visibility and Perception with Clip-Recurrent Transformer

Hao Shi, Qi Jiang, Kailun Yang, Xiaoting Yin, Ze Wang, Kaiwei Wang

专题命中 可控生成 :inpainting(abstract);分类 cs.CV

Comments Accepted to IEEE Transactions on Intelligent Vehicles (T-IV). The source code and dataset are made publicly available at https://github.com/MasterHow/FlowLens

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13897 2024-06-21 cs.CV 57%

CLAY: A Controllable Large-scale Generative Model for Creating High-quality 3D Assets

Longwen Zhang, Ziyu Wang, Qixuan Zhang, Qiwei Qiu, Anqi Pang, Haoran Jiang, Wei Yang, Lan Xu, Jingyi Yu

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Project page: https://sites.google.com/view/clay-3dlm Video: https://youtu.be/YcKFp4U2Voo

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.15253 2024-06-21 cs.CV cs.NA math.NA physics.data-an 57%

Seeing the World through an Antenna's Eye: Reception Quality Visualization Using Incomplete Technical Signal Information

Leif Bergerhoff

专题命中 可控生成 :inpainting(abstract);分类 cs.CV

Comments 5 pages, to be published in the conference proceedings of the European Signal Processing Conference (EUSIPCO) 2024, camera-ready version adding information on the space and time complexity, as well as the dataset size

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08801 2024-06-18 cs.CV 57%

Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Mingwang Xu, Hui Li, Qingkun Su, Hanlin Shang, Liwei Zhang, Ce Liu, Jingdong Wang, Yao Yao, Siyu Zhu

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments 20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17095 2024-06-18 cs.CV cs.AI 57%

Emergent Open-Vocabulary Semantic Segmentation from Off-the-shelf Vision-Language Models

Jiayun Luo, Siddhesh Khandelwal, Leonid Sigal, Boyang Li

专题命中 可控生成 :text-to-image(abstract);分类 cs.CV

Comments Accepted to CVPR 2024; Earlier version of this paper contained an unintentional error stemming from a bug in the code. This version corrects this error, which had to do with filtering of class names. In consultation with CVPR Program Chairs it was suggested errata be submitted as the updated (fixed) code reinforced original findings (albeit with slightly different final numbers)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08249 2024-06-13 cs.CV cs.LG 57%

Dataset Enhancement with Instance-Level Augmentations

Orest Kupyn, Christian Rupprecht

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13254 2024-06-13 cs.CV cs.AI cs.CL cs.LG 57%

CounterCurate: Enhancing Physical and Semantic Visio-Linguistic Compositional Reasoning via Counterfactual Examples

Jianrui Zhang, Mu Cai, Tengyang Xie, Yong Jae Lee

专题命中 可控生成 :image generation(abstract);分类 cs.CV

Comments 15 pages, 6 figures, 12 tables, Project Page: https://countercurate.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.07207 2024-06-12 cs.CV 57%

GALA3D: Towards Text-to-3D Complex Scene Generation via Layout-guided Generative Gaussian Splatting

Xiaoyu Zhou, Xingjian Ran, Yajiao Xiong, Jinlin He, Zhiwei Lin, Yongtao Wang, Deqing Sun, Ming-Hsuan Yang

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04983 2024-06-10 cs.CV 57%

CityCraft: A Real Crafter for 3D City Generation

Jie Deng, Wenhao Chai, Junsheng Huang, Zhonghan Zhao, Qixuan Huang, Mingyan Gao, Jianshu Guo, Shengyu Hao, Wenhao Hu, Jenq-Neng Hwang, Xi Li, Gaoang Wang

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments 20 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01900 2024-06-10 cs.CV 57%

Follow-Your-Emoji: Fine-Controllable and Expressive Freestyle Portrait Animation

Yue Ma, Hongyu Liu, Hongfa Wang, Heng Pan, Yingqing He, Junkun Yuan, Ailing Zeng, Chengfei Cai, Heung-Yeung Shum, Wei Liu, Qifeng Chen

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Project Page: https://follow-your-emoji.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04340 2024-06-07 cs.CV 57%

GLACE: Global Local Accelerated Coordinate Encoding

Fangjinhua Wang, Xudong Jiang, Silvano Galliani, Christoph Vogel, Marc Pollefeys

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Large-scale visual localization with a single optimizable MLP. CVPR 2024. Code: https://github.com/cvg/glace. Project page: https://xjiangan.github.io/glace

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03866 2024-06-07 cs.CV 57%

LLplace: The 3D Indoor Scene Layout Generation and Editing via Large Language Model

Yixuan Yang, Junru Lu, Zixiang Zhao, Zhen Luo, James J. Q. Yu, Victor Sanchez, Feng Zheng

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01349 2024-06-07 cs.CV 57%

Unleashing Generalization of End-to-End Autonomous Driving with Controllable Long Video Generation

Enhui Ma, Lijun Zhou, Tao Tang, Zhan Zhang, Dong Han, Junpeng Jiang, Kun Zhan, Peng Jia, Xianpeng Lang, Haiyang Sun, Di Lin, Kaicheng Yu

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Project Page: https://westlake-autolab.github.io/delphi.github.io/, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18610 2024-06-07 cs.CV 57%

DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB Image

Daoyi Gao, Dávid Rozenberszki, Stefan Leutenegger, Angela Dai

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments SIGGRAPH 2024, Project page: https://daoyig.github.io/DiffCAD/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.02509 2024-06-05 cs.CV 57%

CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation

Dejia Xu, Weili Nie, Chao Liu, Sifei Liu, Jan Kautz, Zhangyang Wang, Arash Vahdat

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

Comments Project page: https://ir1d.github.io/CamCo/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14781 2024-06-04 cs.CV 57%

Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance

Shenhao Zhu, Junming Leo Chen, Zuozhuo Dai, Qingkun Su, Yinghui Xu, Xun Cao, Yao Yao, Hao Zhu, Siyu Zhu

专题命中 可控生成 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.19854 2024-05-31 cs.CV 57%

RTGen: Generating Region-Text Pairs for Open-Vocabulary Object Detection

Fangyi Chen, Han Zhang, Zhantao Yang, Hao Chen, Kai Hu, Marios Savvides

专题命中 可控生成 :inpainting(abstract);分类 cs.CV

Comments Technical report

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.18801 2024-05-30 cs.CV 57%

SketchTriplet: Self-Supervised Scenarized Sketch-Text-Image Triplet Generation

Zhenbei Wu, Qiang Wang, Jie Yang

专题命中 可控生成 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏