arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 3490 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3490 篇

2501.09488 2025-02-04 astro-ph.IM astro-ph.HE astro-ph.SR 50%

Mining the time axis with TRON. I. Millisecond pulsars in Omega Centauri, Terzan 5 and 47 Tucanae detected through MeerKAT interferometric imaging

Oleg M. Smirnov, Ian Heywood, Marisa Geyer, Talon Myburgh, Cyril Tasse, Jonathan S. Kenyon, Simon J. Perkins, James Dawson, Hertzog L. Bester, Joe S. Bright, Buntu Ngcebetsha, Nadeem Oozeer, Victoria G. G. Samboco, Isaac Sihlangu, Carmen Choza, Andrew P. V. Siemion

专题命中 文生图 :image synthesis(abstract)

Comments 6 pages, 8 figures, published in MNRAS Letters

Journal ref Monthly Notices of the Royal Astronomical Society Letters (2025), 538, 1, L62-L68

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06781 2025-01-27 cs.AI 50%

Eliza: A Web3 friendly AI Agent Operating System

Shaw Walters, Sam Gao, Shakker Nerd, Feng Da, Warren Williams, Ting-Chien Meng, Amie Chow, Hunter Han, Frank He, Allen Zhang, Ming Wu, Timothy Shen, Maxwell Hu, Jerry Yan

专题命中 文生图 :text-to-image(abstract)

Comments 20 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10383 2025-01-22 cs.CY cs.HC 50%

The Generative AI Ethics Playbook

Jessie J. Smith, Wesley Hanwen Deng, William H. Smith, Maarten Sap, Nicole DeCario, Jesse Dodge

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.17181 2025-01-07 cs.CL cs.LG 50%

Unsupervised Text Embedding Space Generation Using Generative Adversarial Networks for Text Synthesis

Jun-Min Lee, Tae-Bin Ha

专题命中 文生图 :image synthesis(abstract)

Comments NEJLT accpeted

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03287 2024-12-05 cs.AI 50%

Integrating Generative AI into Art Therapy: A Technical Showcase

Yannis Valentin Schmutz, Tetiana Kravchenko, Souhir Ben Souissi, Mascha Kurpicz-Briki

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08946 2024-11-08 eess.IV 50%

Color Agnostic Cross-Spectral Disparity Estimation

Frank Sippel, Nils Genser, Hannah Och, Jürgen Seiler, André Kaup

专题命中 文生图 :image synthesis(abstract)

Journal ref 2024 IEEE International Conference on Acoustics, Speech and Signal Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.19561 2024-11-05 cs.AI cs.CL 50%

Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models

Venkat Venkatasubramanian, Arijit Chakraborty

专题命中 文生图 :image synthesis(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17032 2024-10-23 cs.AI 50%

Insights on Disagreement Patterns in Multimodal Safety Perception across Diverse Rater Groups

Charvi Rastogi, Tian Huey Teh, Pushkar Mishra, Roma Patel, Zoe Ashwood, Aida Mostafazadeh Davani, Mark Diaz, Michela Paganini, Alicia Parrish, Ding Wang, Vinodkumar Prabhakaran, Lora Aroyo, Verena Rieser

专题命中 文生图 :text-to-image(abstract)

Comments 20 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19242 2024-10-17 cs.CL 50%

SciDoc2Diagrammer-MAF: Towards Generation of Scientific Diagrams from Documents guided by Multi-Aspect Feedback Refinement

Ishani Mondal, Zongxia Li, Yufang Hou, Anandhavelu Natarajan, Aparna Garimella, Jordan Boyd-Graber

专题命中 文生图 :text-to-image(abstract)

Comments Code and data available at https://github.com/Ishani-Mondal/SciDoc2DiagramGeneration

Journal ref Empirical Methods in Natural Language Processing 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08549 2024-10-16 cs.LG 50%

Score Neural Operator: A Generative Model for Learning and Generalizing Across Multiple Probability Distributions

Xinyu Liao, Aoyang Qin, Jacob Seidman, Junqi Wang, Wei Wang, Paris Perdikaris

专题命中 文生图 :image synthesis(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.06768 2024-10-10 cs.HC 50%

Patterns of Creativity: How User Input Shapes AI-Generated Visual Diversity

Maria-Teresa De Rosa Palmini, Eva Cetinic

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04280 2024-10-08 cs.HC 50%

The Visualization JUDGE : Can Multimodal Foundation Models Guide Visualization Design Through Visual Perception?

Matthew Berger, Shusen Liu

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.09247 2024-10-07 astro-ph.CO astro-ph.IM 50%

Weak lensing scattering transform: dark energy and neutrino mass sensitivity

Sihao Cheng, Brice Ménard

专题命中 文生图 :image synthesis(abstract)

Comments 9 pages, 6 figures. Accepted to MNRAS

Journal ref Monthly Notices of the Royal Astronomical Society, Volume 507, Issue 1, 2021, pp. 1012-1020

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.08323 2024-09-18 cs.CY cs.AI 50%

Mapping the Ethics of Generative AI: A Comprehensive Scoping Review

Thilo Hagendorff

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01499 2024-08-06 q-fin.ST cs.LG 50%

NeuralFactors: A Novel Factor Learning Approach to Generative Modeling of Equities

Achintya Gopal

专题命中 文生图 :text-to-image(abstract)

Comments 9 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14779 2024-08-06 cs.CY cs.AI cs.HC 50%

Do Generative AI Models Output Harm while Representing Non-Western Cultures: Evidence from A Community-Centered Approach

Sourojit Ghosh, Pranav Narayanan Venkit, Sanjana Gautam, Shomir Wilson, Aylin Caliskan

专题命中 文生图 :text-to-image(abstract)

Comments This is the pre-peer reviewed version, which has been accepted at the 7th AAAI ACM Conference on AI, Ethics, and Society, Oct. 21, 2024, California, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.02591 2024-08-01 cs.LG cs.AI 50%

Unveiling the Potential of AI for Nanomaterial Morphology Prediction

Ivan Dubrovsky, Andrei Dmitrenko, Aleksei Dmitrenko, Nikita Serov, Vladimir Vinogradov

专题命中 文生图 :text-to-image(abstract)

Journal ref Proceedings of the 41 st International Conference on Machine Learning. PMLR 235, 2024, 11957--11978

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12793 2024-07-31 cs.CL 50%

ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Team GLM, :, Aohan Zeng, Bin Xu, Bowen Wang, Chenhui Zhang, Da Yin, Dan Zhang, Diego Rojas, Guanyu Feng, Hanlin Zhao, Hanyu Lai, Hao Yu, Hongning Wang, Jiadai Sun, Jiajie Zhang, Jiale Cheng, Jiayi Gui, Jie Tang, Jing Zhang, Jingyu Sun, Juanzi Li, Lei Zhao, Lindong Wu, Lucen Zhong, Mingdao Liu, Minlie Huang, Peng Zhang, Qinkai Zheng, Rui Lu, Shuaiqi Duan, Shudan Zhang, Shulin Cao, Shuxun Yang, Weng Lam Tam, Wenyi Zhao, Xiao Liu, Xiao Xia, Xiaohan Zhang, Xiaotao Gu, Xin Lv, Xinghan Liu, Xinyi Liu, Xinyue Yang, Xixuan Song, Xunkai Zhang, Yifan An, Yifan Xu, Yilin Niu, Yuantao Yang, Yueyan Li, Yushi Bai, Yuxiao Dong, Zehan Qi, Zhaoyu Wang, Zhen Yang, Zhengxiao Du, Zhenyu Hou, Zihan Wang

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.11844 2024-06-19 cs.CY cs.AI 50%

Prompting the E-Brushes: Users as Authors in Generative AI

Yiyang Mei

专题命中 文生图 :text-to-image(abstract)

Journal ref International Journal of Law, Ethics, and Technology 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08447 2024-06-12 physics.med-ph 50%

Generating Synthetic Computed Tomography for Radiotherapy: SynthRAD2023 Challenge Report

Evi M. C. Huijben, Maarten L. Terpstra, Arthur Jr. Galapon, Suraj Pai, Adrian Thummerer, Peter Koopmans, Manya Afonso, Maureen van Eijnatten, Oliver Gurney-Champion, Zeli Chen, Yiwen Zhang, Kaiyi Zheng, Chuanpu Li, Haowen Pang, Chuyang Ye, Runqi Wang, Tao Song, Fuxin Fan, Jingna Qiu, Yixing Huang, Juhyung Ha, Jong Sung Park, Alexandra Alain-Beaudoin, Silvain Bériault, Pengxin Yu, Hongbin Guo, Zhanyao Huang, Gengwan Li, Xueru Zhang, Yubo Fan, Han Liu, Bowen Xin, Aaron Nicolson, Lujia Zhong, Zhiwei Deng, Gustav Müller-Franzes, Firas Khader, Xia Li, Ye Zhang, Cédric Hémon, Valentin Boussot, Zhihao Zhang, Long Wang, Lu Bai, Shaobin Wang, Derk Mus, Bram Kooiman, Chelsea A. H. Sargeant, Edward G. A. Henderson, Satoshi Kondo, Satoshi Kasai, Reza Karimzadeh, Bulat Ibragimov, Thomas Helfer, Jessica Dafflon, Zijie Chen, Enpei Wang, Zoltan Perko, Matteo Maspero

专题命中 文生图 :image synthesis(abstract)

Comments Preprint submitted to Medical Image Analysis

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.03972 2024-06-11 cs.IR 50%

Category-Oriented Representation Learning for Image to Multi-Modal Retrieval

Zida Cheng, Chen Ju, Shuai Xiao, Xu Chen, Zhonghua Zhai, Xiaoyi Zeng, Weilin Huang, Junchi Yan

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.02523 2024-06-05 cs.RO cs.AI cs.LG 50%

RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots

Soroush Nasiriany, Abhiram Maddukuri, Lance Zhang, Adeet Parikh, Aaron Lo, Abhishek Joshi, Ajay Mandlekar, Yuke Zhu

专题命中 文生图 :text-to-image(abstract)

Comments RSS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.04629 2024-05-30 eess.IV cs.AI physics.med-ph 50%

ResNCT: A Deep Learning Model for the Synthesis of Nephrographic Phase Images in CT Urography

Syed Jamal Safdar Gardezi, Lucas Aronson, Peter Wawrzyn, Hongkun Yu, E. Jason Abel, Daniel D. Shapiro, Meghan G. Lubner, Joshua Warner, Giuseppe Toia, Lu Mao, Pallavi Tiwari, Andrew L. Wentland

专题命中 文生图 :image synthesis(abstract)

Comments 10 pages, 5 Figures,2 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17346 2024-05-28 cs.LG cs.AI 50%

Prompt Optimization with Human Feedback

Xiaoqiang Lin, Zhongxiang Dai, Arun Verma, See-Kiong Ng, Patrick Jaillet, Bryan Kian Hsiang Low

专题命中 文生图 :text-to-image(abstract)

Comments Preprint, 18 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16119 2024-05-28 cs.LG eess.IV 50%

Method and Software Tool for Generating Artificial Databases of Biomedical Images Based on Deep Neural Networks

Oleh Berezsky, Petro Liashchynskyi, Oleh Pitsun, Grygoriy Melnyk

专题命中 文生图 :image synthesis(abstract)

Comments CEUR Workshop Proceedings (CEUR-WS.org). IDDM'2023: 6th International Conference on Informatics & Data-Driven Medicine, November 17 - 19, 2023, Bratislava, Slovakia

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.15406 2024-05-27 physics.flu-dyn cs.LG 50%

A Misleading Gallery of Fluid Motion by Generative Artificial Intelligence

Ali Kashefi

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14794 2024-05-27 cs.HC 50%

RetAssist: Facilitating Vocabulary Learners with Generative Images in Story Retelling Practices

Qiaoyi Chen, Siyu Liu, Kaihui Huang, Xingbo Wang, Xiaojuan Ma, Junkai Zhu, Zhenhui Peng

专题命中 文生图 :text-to-image(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09734 2024-05-17 cs.CY 50%

Attention is All You Want: Machinic Gaze and the Anthropocene

Liam Magee, Vanicka Arora

专题命中 文生图 :text-to-image(abstract)

Comments 19 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07553 2024-05-14 cs.AI cs.CL 50%

Hijacking Context in Large Multi-modal Models

Joonhyun Jeong

专题命中 文生图 :text-to-image(abstract)

Comments Technical Report. Preprint

Journal ref ICLR 2024 Workshop on Reliable and Responsible Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04420 2024-05-06 cs.LG q-bio.NC 50%

BrainSCUBA: Fine-Grained Natural Language Captions of Visual Cortex Selectivity

Andrew F. Luo, Margaret M. Henderson, Michael J. Tarr, Leila Wehbe

专题命中 文生图 :image synthesis(abstract)

Comments ICLR 2024. Project page: https://www.cs.cmu.edu/~afluo/BrainSCUBA

详情

展开后加载摘要…

URL PDF HTML 收藏