arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 86714 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3486 篇

2409.20340 2024-10-30 eess.IV cs.AI cs.CV cs.LG 57%

Enhancing GANs with Contrastive Learning-Based Multistage Progressive Finetuning SNN and RL-Based External Optimization

Osama Mustafa

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.04306 2024-10-29 cs.CV cs.AI cs.LG 57%

Effectiveness Assessment of Recent Large Vision-Language Models

Yao Jiang, Xinyu Yan, Ge-Peng Ji, Keren Fu, Meijun Sun, Huan Xiong, Deng-Ping Fan, Fahad Shahbaz Khan

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Accepted by Visual Intelligence

Journal ref Visual Intelligence, 2024, Vol. 2, article no. 17

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21001 2024-10-28 cs.CV cs.AI cs.LG 57%

GABInsight: Exploring Gender-Activity Binding Bias in Vision-Language Models

Ali Abdollahi, Mahdi Ghaznavi, Mohammad Reza Karimi Nejad, Arash Mari Oriyad, Reza Abbasi, Ali Salesi, Melika Behjati, Mohammad Hossein Rohban, Mahdieh Soleymani Baghshah

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Journal ref Volume 392 of ECAI 2024, Pages 729 - 736

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.06131 2024-10-28 cs.CV 57%

Generative AI meets 3D: A Survey on Text-to-3D in AIGC Era

Chenghao Li, Chaoning Zhang, Joseph Cho, Atish Waghwase, Lik-Hang Lee, Francois Rameau, Yang Yang, Sung-Ho Bae, Choong Seon Hong

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17959 2024-10-24 eess.IV cs.CV cs.LG 57%

Medical Imaging Complexity and its Effects on GAN Performance

William Cagas, Chan Ko, Blake Hsiao, Shryuk Grandhi, Rishi Bhattacharya, Kevin Zhu, Michael Lam

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted to ACCV, Workshop on Generative AI for Synthetic Medical Data

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16820 2024-10-23 cs.CV 57%

AttriPrompter: Auto-Prompting with Attribute Semantics for Zero-shot Nuclei Detection via Visual-Language Pre-trained Models

Yongjian Wu, Yang Zhou, Jiya Saiyin, Bingzheng Wei, Maode Lai, Jianzhong Shou, Yan Xu

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments This article has been accepted for publication in a future issue of IEEE Transactions on Medical Imaging (TMI), but has not been fully edited. Content may change prior to final publication. Citation information: DOI: https://doi.org/10.1109/TMI.2024.3473745 . Code: https://github.com/wuyongjianCODE/AttriPrompter

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14749 2024-10-22 cs.LG cs.CV 57%

CFTS-GAN: Continual Few-Shot Teacher Student for Generative Adversarial Networks

Munsif Ali, Leonardo Rossi, Massimo Bertozzi

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08534 2024-10-22 cs.CV eess.IV 57%

Quality Prediction of AI Generated Images and Videos: Emerging Trends and Opportunities

Abhijay Ghildyal, Yuanhan Chen, Saman Zadtootaghaj, Nabajeet Barman, Alan C. Bovik

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments "The abstract field cannot be longer than 1,920 characters", the abstract appearing here is slightly shorter than that in the PDF file

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.09977 2024-10-15 eess.IV cs.CV 57%

Unified Framework for Histopathology Image Augmentation and Classification via Generative Models

Meng Li, Chaoyi Li, Can Peng, Brian C. Lovell

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02613 2024-10-04 cs.CV cs.AI cs.CL 57%

NL-Eye: Abductive NLI for Images

Mor Ventura, Michael Toker, Nitay Calderon, Zorik Gekhman, Yonatan Bitton, Roi Reichart

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.00905 2024-10-02 cs.CV 57%

Removing Distributional Discrepancies in Captions Improves Image-Text Alignment

Yuheng Li, Haotian Liu, Mu Cai, Yijun Li, Eli Shechtman, Zhe Lin, Yong Jae Lee, Krishna Kumar Singh

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.13788 2024-10-01 cs.CV cs.AI 57%

AnyPattern: Towards In-context Image Copy Detection

Wenhao Wang, Yifan Sun, Zhentao Tan, Yi Yang

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments The project is publicly available at https://anypattern.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15781 2024-09-25 cs.CV 57%

Training Data Attribution: Was Your Model Secretly Trained On Data Created By Mine?

Likun Zhang, Hao Wu, Lingcui Zhang, Fengyuan Xu, Jin Cao, Fenghua Li, Ben Niu

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12244 2024-09-20 cs.CV cs.AI cs.LG 57%

Sparks of Artificial General Intelligence(AGI) in Semiconductor Material Science: Early Explorations into the Next Frontier of Generative AI-Assisted Electron Micrograph Analysis

Sakhinana Sagar Srinivas, Geethan Sannidhi, Sreeja Gangasani, Chidaksh Ravuru, Venkataramana Runkana

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Published at Deployable AI (DAI) Workshop at AAAI-2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11923 2024-09-19 cs.CV 57%

Agglomerative Token Clustering

Joakim Bruslund Haurum, Sergio Escalera, Graham W. Taylor, Thomas B. Moeslund

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments ECCV 2024. Project webpage at https://vap.aau.dk/atc/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02530 2024-09-18 cs.CV cs.AI 57%

Manipulating and Mitigating Generative Model Biases without Retraining

Jordan Vice, Naveed Akhtar, Richard Hartley, Ajmal Mian

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Accepted to ECCV 2024 WS: Workshop on critical evaluation of generative models and their impact on society (CEGIS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09361 2024-09-17 cs.LG cs.CV stat.ML 57%

Beta-Sigma VAE: Separating beta and decoder variance in Gaussian variational autoencoder

Seunghwan Kim, Seungkyu Lee

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted for ICPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.01936 2024-09-04 cs.CV cs.LG 57%

Optimizing CLIP Models for Image Retrieval with Maintained Joint-Embedding Alignment

Konstantin Schall, Kai Uwe Barthel, Nico Hezel, Klaus Jung

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.05783 2024-09-04 cs.CY cs.AI cs.CL cs.CV 57%

A Survey on Responsible Generative AI: What to Generate and What Not

Jindong Gu

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments 77 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.13906 2024-08-27 cs.CV cs.AI cs.LG 57%

ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models

Yeji Park, Deokyeong Lee, Junsuk Choe, Buru Chang

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments First two authors contributed equally. Source code is available at https://github.com/yejipark-m/ConVis

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.07588 2024-08-19 eess.IV cs.CV cs.LG stat.ML 57%

Uncertainty Quantification using Variational Inference for Biomedical Image Segmentation

Abhinav Sagar

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08541 2024-08-15 cs.CV 57%

Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation

Zhengyuan Yang, Jianfeng Wang, Linjie Li, Kevin Lin, Chung-Ching Lin, Zicheng Liu, Lijuan Wang

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments ECCV 2024; Project page at https://idea2img.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07315 2024-08-13 cs.CV 57%

NVS-Adapter: Plug-and-Play Novel View Synthesis from a Single Image

Yoonwoo Jeong, Jinwoo Lee, Chiheon Kim, Minsu Cho, Doyup Lee

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments [ECCV2024] Project Page: https://postech-cvlab.github.io/nvsadapter/

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.02231 2024-08-06 cs.CV 57%

REVISION: Rendering Tools Enable Spatial Fidelity in Vision-Language Models

Agneet Chatterjee, Yiran Luo, Tejas Gokhale, Yezhou Yang, Chitta Baral

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Accepted to ECCV 2024. Project Page : https://agneetchatterjee.com/revision/

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00352 2024-08-02 cs.CV 57%

Autonomous LLM-Enhanced Adversarial Attack for Text-to-Motion

Honglei Miao, Fan Ma, Ruijie Quan, Kun Zhan, Yi Yang

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20268 2024-07-31 cs.CV cs.LG eess.IV 57%

Utilizing Generative Adversarial Networks for Image Data Augmentation and Classification of Semiconductor Wafer Dicing Induced Defects

Zhining Hu, Tobias Schlosser, Michael Friedrich, André Luiz Vieira e Silva, Frederik Beuth, Danny Kowerko

专题命中 文生图 :image synthesis(abstract);分类 cs.CV

Comments Accepted for: 2024 IEEE 29th International Conference on Emerging Technologies and Factory Automation (ETFA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.15992 2024-07-19 cs.CV cs.CL 57%

BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval

Yinda Chen, Che Liu, Xiaoyu Liu, Rossella Arcucci, Zhiwei Xiong

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.14000 2024-07-19 cs.CV 57%

Real-time 3D-aware Portrait Editing from a Single Image

Qingyan Bai, Zifan Shi, Yinghao Xu, Hao Ouyang, Qiuyu Wang, Ceyuan Yang, Xuan Wang, Gordon Wetzstein, Yujun Shen, Qifeng Chen

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments ECCV 2024 camera-ready version. Project page: https://github.com/EzioBy/3dpe

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01832 2024-07-19 cs.CV cs.AI cs.LG 57%

SynthCLIP: Are We Ready for a Fully Synthetic CLIP Training?

Hasan Abed Al Kader Hammoud, Hani Itani, Fabio Pizzati, Philip Torr, Adel Bibi, Bernard Ghanem

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02192 2024-07-18 cs.CV 57%

DiverseDream: Diverse Text-to-3D Synthesis with Augmented Text Embedding

Uy Dieu Tran, Minh Luu, Phong Ha Nguyen, Khoi Nguyen, Binh-Son Hua

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments Accepted to ECCV 2024. Project page: https://diversedream.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏