arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86714 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3486 篇

2202.12165 2022-08-22 cs.CV 57%

Transformers in Medical Image Analysis: A Review

Kelei He, Chen Gan, Zhuoyuan Li, Islem Rekik, Zihao Yin, Wen Ji, Yang Gao, Qian Wang, Junfeng Zhang, Dinggang Shen

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.08284 2022-08-18 eess.IV cs.CV cs.LG 57%

Novel Deep Learning Approach to Derive Cytokeratin Expression and Epithelium Segmentation from DAPI

Felix Jakob Segerer, Katharina Nekolla, Lorenz Rognoni, Ansh Kapil, Markus Schick, Helen Angell, Günter Schmidt

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Short Paper - MIDL2022 (Medical Imaging with Deep Learning)

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.08207 2022-08-18 cs.CV cs.AI 57%

Time flies by: Analyzing the Impact of Face Ageing on the Recognition Performance with Synthetic Data

Marcel Grimmer, Haoyu Zhang, Raghavendra Ramachandra, Kiran Raja, Christoph Busch

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.13105 2022-08-10 eess.IV cs.CV 57%

A Novel Framework for Image-to-image Translation and Image Compression

Fei Yang, Yaxing Wang, Luis Herranz, Yongmei Cheng, Mikhail Mozerov

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 14 pages, 15 figures, accepted by Neurocomputing

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.11075 2022-07-25 cs.CV 57%

RealFlow: EM-based Realistic Optical Flow Dataset Generation from Videos

Yunhui Han, Kunming Luo, Ao Luo, Jiangyu Liu, Haoqiang Fan, Guiming Luo, Shuaicheng Liu

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments ECCV 2022 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.08238 2022-07-22 cs.CV cs.LG 57%

AXM-Net: Implicit Cross-Modal Feature Alignment for Person Re-identification

Ammarah Farooq, Muhammad Awais, Josef Kittler, Syed Safwan Khalid

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments AAAI-2022 (Oral Paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.02250 2022-07-07 cs.CV eess.IV 57%

Array Camera Image Fusion using Physics-Aware Transformers

Qian Huang, Minghao Hu, David Jones Brady

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.11404 2022-06-24 cs.CV cs.AI cs.LG 57%

The ArtBench Dataset: Benchmarking Generative Models with Artworks

Peiyuan Liao, Xiuyu Li, Xihui Liu, Kurt Keutzer

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments The first two authors contributed equally to this work. The code and data are available at https://github.com/liaopeiyuan/artbench

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.04503 2022-06-10 cs.CV cs.AI 57%

cycle text2face: cycle text-to-face gan via transformers

Faezeh Gholamrezaie, Mohammad Manthouri

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15868 2022-06-01 cs.CV cs.CL cs.LG 57%

CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Wenyi Hong, Ming Ding, Wendi Zheng, Xinghan Liu, Jie Tang

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.13294 2022-05-27 cs.CV eess.IV eess.SP 57%

Analytical Interpretation of Latent Codes in InfoGAN with SAR Images

Zhenpeng Feng, Milos Dakovic, Hongbing Ji, Mingzhe Zhu, Ljubisa Stankovic

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 13 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.07441 2022-05-23 cs.CV cs.CL cs.IR 57%

COTS: Collaborative Two-Stream Vision-Language Pre-Training Model for Cross-Modal Retrieval

Haoyu Lu, Nanyi Fei, Yuqi Huo, Yizhao Gao, Zhiwu Lu, Ji-Rong Wen

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Accepted by CVPR2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.08101 2022-05-17 cs.CV cs.IR 57%

ARTEMIS: Attention-based Retrieval with Text-Explicit Matching and Implicit Similarity

Ginger Delmas, Rafael Sampaio de Rezende, Gabriela Csurka, Diane Larlus

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Published in ICLR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.06126 2022-05-13 cs.CL cs.CV cs.LG 57%

One Model, Multiple Modalities: A Sparsely Activated Approach for Text, Sound, Image, Video and Code

Yong Dai, Duyu Tang, Liangxin Liu, Minghuan Tan, Cong Zhou, Jingquan Wang, Zhangyin Feng, Fan Zhang, Xueyu Hu, Shuming Shi

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.05922 2022-05-13 cs.CV 57%

Ray Priors through Reprojection: Improving Neural Radiance Fields for Novel View Extrapolation

Jian Zhang, Yuanqing Zhang, Huan Fu, Xiaowei Zhou, Bowen Cai, Jinchi Huang, Rongfei Jia, Binqiang Zhao, Xing Tang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.00273 2022-05-06 cs.LG cs.CV 57%

StyleGAN-XL: Scaling StyleGAN to Large Diverse Datasets

Axel Sauer, Katja Schwarz, Andreas Geiger

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments To appear in SIGGRAPH 2022. Project Page: https://sites.google.com/view/stylegan-xl/

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.01879 2022-05-05 cs.CV cs.AI 57%

Learning Two-Stream CNN for Multi-Modal Age-related Macular Degeneration Categorization

Weisen Wang, Xirong Li, Zhiyan Xu, Weihong Yu, Jianchun Zhao, Dayong Ding, Youxin Chen

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted by IEEE Journal of Biomedical and Health Informatics (J-BHI)

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.12797 2022-04-28 cs.GR 57%

Towards Quantum Ray Tracing

Luís Paulo Santos, Thomas Bashford-Rogers, João Barbosa, Paul Navrátil

专题命中 文生图 :image synthesis(abstract);分类 cs.GR

Comments 27 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.12007 2022-04-28 eess.IV cs.CV physics.med-ph 57%

Assessing the ability of generative adversarial networks to learn canonical medical image statistics

Varun A. Kelkar, Dimitrios S. Gotsis, Frank J. Brooks, Prabhat KC, Kyle J. Myers, Rongping Zeng, Mark A. Anastasio

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.08075 2022-04-28 cs.CL cs.CV 57%

Things not Written in Text: Exploring Spatial Commonsense from Visual Signals

Xiao Liu, Da Yin, Yansong Feng, Dongyan Zhao

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted by ACL 2022 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.05864 2022-04-26 cs.CV cs.AI 57%

Human Silhouette and Skeleton Video Synthesis through Wi-Fi signals

Danilo Avola, Marco Cascio, Luigi Cinque, Alessio Fagioli, Gian Luca Foresti

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Journal ref International Journal of Neural Systems, 2022, 2250015

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.09268 2022-04-21 cs.LG cs.CL cs.CV cs.IR 57%

Uncertainty-based Cross-Modal Retrieval with Probabilistic Representations

Leila Pishdad, Ran Zhang, Konstantinos G. Derpanis, Allan Jepson, Afsaneh Fazly

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments 13 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.08339 2022-04-19 cs.CV 57%

Migrating Face Swap to Mobile Devices: A lightweight Framework and A Supervised Training Solution

Haiming Yu, Hao Zhu, Xiangju Lu, Junhui Liu

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to IEEE International Conference on Multimedia and Expo 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.03547 2022-04-08 eess.IV cs.CV physics.med-ph 57%

Evaluating Procedures for Establishing Generative Adversarial Network-based Stochastic Image Models in Medical Imaging

Varun A. Kelkar, Dimitrios S. Gotsis, Frank J. Brooks, Kyle J. Myers, Prabhat KC, Rongping Zeng, Mark A. Anastasio

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Published in SPIE Medical Imaging 2022: Image Perception, Observer Performance, and Technology Assessment

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.01960 2022-04-06 cs.CV cs.AI stat.ML 57%

FaceSigns: Semi-Fragile Neural Watermarks for Media Authentication and Countering Deepfakes

Paarth Neekhara, Shehzeen Hussain, Xinqiao Zhang, Ke Huang, Julian McAuley, Farinaz Koushanfar

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 13 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.03959 2022-04-01 cs.LG cs.CV 57%

Generative Flows with Invertible Attentions

Rhea Sanjay Sukthanker, Zhiwu Huang, Suryansh Kumar, Radu Timofte, Luc Van Gool

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15006 2022-03-30 cs.CV 57%

TL-GAN: Improving Traffic Light Recognition via Data Synthesis for Autonomous Driving

Danfeng Wang, Xin Ma, Xiaodong Yang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.09824 2022-03-21 cs.CV cs.LG eess.AS 57%

Cross-Modal Perceptionist: Can Face Geometry be Gleaned from Voices?

Cho-Ying Wu, Chin-Cheng Hsu, Ulrich Neumann

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to CVPR 2022. Project page: https://choyingw.github.io/works/Voice2Mesh/index.html. This version supersedes arXiv:2104.10299

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.09771 2022-03-21 cs.CV 57%

Beyond a Video Frame Interpolator: A Space Decoupled Learning Approach to Continuous Image Transition

Tao Yang, Peiran Ren, Xuansong Xie, Xiansheng Hua, Lei Zhang

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.03638 2022-03-09 eess.IV cs.AI cs.CV cs.LG 57%

Unsupervised Image Registration Towards Enhancing Performance and Explainability in Cardiac And Brain Image Analysis

Chengjia Wang, Guang Yang, Giorgos Papanastasiou

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments 38 pages, 7 figures, will be published in Sensors journal by MDPI

详情

展开后加载摘要…

URL PDF HTML 收藏