arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12228 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12228 篇

2303.12786 2024-04-10 cs.CV 78%

FeatureNeRF: Learning Generalizable NeRFs by Distilling Foundation Models

Jianglong Ye, Naiyan Wang, Xiaolong Wang

专题命中 其他LLM :foundation model(title,abstract)

Comments Project page: https://jianglongye.com/featurenerf/

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04781 2024-04-09 cs.RO 78%

Unifying Foundation Models with Quadrotor Control for Visual Tracking Beyond Object Categories

Alessandro Saviolo, Pratyaksh Rao, Vivek Radhakrishnan, Jiuhong Xiao, Giuseppe Loianno

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02706 2024-04-04 cs.HC 78%

Unblind Text Inputs: Predicting Hint-text of Text Input in Mobile Apps via LLM

Zhe Liu, Chunyang Chen, Junjie Wang, Mengzhuo Chen, Boyu Wu, Yuekai Huang, Jun Hu, Qing Wang

专题命中 其他LLM :LLM(title,abstract)

Comments Accepted by the 2024 CHI Conference on Human Factors in Computing Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.11178 2024-03-22 cs.CV 78%

Active Prompt Learning in Vision Language Models

Jihwan Bang, Sumyeong Ahn, Jae-Gil Lee

专题命中 其他LLM :language model(title,abstract)

Comments accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12082 2024-03-20 cs.CL cs.AI cs.LG 78%

The Boy Who Survived: Removing Harry Potter from an LLM is harder than reported

Adam Shostack

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI、cs.LG

Comments 2 pages, 4 pages of appendix. Comment on arXiv:2310.02238

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.05657 2024-03-19 cs.CV 78%

Tag2Text: Guiding Vision-Language Model via Image Tagging

Xinyu Huang, Youcai Zhang, Jinyu Ma, Weiwei Tian, Rui Feng, Yuejie Zhang, Yaqian Li, Yandong Guo, Lei Zhang

专题命中 其他LLM :language model(title,abstract)

Comments Accepted by ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07451 2024-03-19 cs.RO 78%

Daily Assistive View Control Learning of Low-Cost Low-Rigidity Robot via Large-Scale Vision-Language Model

Kento Kawaharazuka, Naoaki Kanazawa, Yoshiki Obinata, Kei Okada, Masayuki Inaba

专题命中 其他LLM :language model(title,abstract)

Comments accepted at Humanoids2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.09985 2024-03-04 cs.HC 78%

Prompting for Discovery: Flexible Sense-Making for AI Art-Making with Dreamsheets

Shm Garanganao Almeda, J. D. Zamfirescu-Pereira, Kyu Won Kim, Pradeep Mani Rathnam, Bjoern Hartmann

专题命中 其他LLM :prompting(title);LLM(abstract)

Comments 13 pages, 14 figures, currently under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.08866 2024-01-18 cs.CY 78%

Foundation Models in Augmentative and Alternative Communication: Opportunities and Challenges

Ambra Di Paola, Serena Muraro, Roberto Marinelli, Christian Pilato

专题命中 其他LLM :foundation model(title,abstract)

Comments This work is intended to be inclusive, fostering collaboration rather than competition in this social activity. For this reason, we invite researchers to contribute with comments and suggestions that will be included (and acknowledged) in any future versions of this work

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16426 2023-12-27 cs.RO 78%

QwenGrasp: A Usage of Large Vision-Language Model for Target-Oriented Grasping

Xinyu Chen, Jian Yang, Zonghan He, Haobin Yang, Qi Zhao, Yuhui Shi

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.12479 2023-12-21 cs.CV 78%

Zero-shot Building Attribute Extraction from Large-Scale Vision and Language Models

Fei Pan, Sangryul Jeon, Brian Wang, Frank Mckenna, Stella X. Yu

专题命中 其他LLM :language model(title,abstract)

Comments Accepted to WACV 2024, Project Page: https://sites.google.com/view/zobae/home

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09740 2023-12-18 cs.RO 78%

VITA: A Multi-modal LLM-based System for Longitudinal, Autonomous, and Adaptive Robotic Mental Well-being Coaching

Micol Spitale, Minja Axelsson, Hatice Gunes

专题命中 其他LLM :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.00500 2023-12-07 cs.CV 78%

Self-Supervised Open-Ended Classification with Small Visual Language Models

Mohammad Mahdi Derakhshani, Ivona Najdenkoska, Cees G. M. Snoek, Marcel Worring, Yuki M. Asano

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16090 2023-11-28 cs.CV 78%

Self-correcting LLM-controlled Diffusion Models

Tsung-Han Wu, Long Lian, Joseph E. Gonzalez, Boyi Li, Trevor Darrell

专题命中 其他LLM :LLM(title,abstract)

Comments 16 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.11004 2023-11-21 q-bio.QM 78%

A Foundation Model for Cell Segmentation

Uriah Israel, Markus Marks, Rohit Dilip, Qilin Li, Morgan Schwartz, Elora Pradhan, Edward Pao, Shenyi Li, Alexander Pearson-Goulart, Pietro Perona, Georgia Gkioxari, Ross Barnowski, Yisong Yue, David Van Valen

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.08245 2023-11-15 cs.CV 78%

TENT: Connect Language Models with IoT Sensors for Zero-Shot Activity Recognition

Yunjiao Zhou, Jianfei Yang, Han Zou, Lihua Xie

专题命中 其他LLM :language model(title,abstract)

Comments Preprint manuscript in submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.03785 2023-11-09 cs.DB 78%

Zelda: Video Analytics using Vision-Language Models

Francisco Romero, Caleb Winston, Johann Hauswald, Matei Zaharia, Christos Kozyrakis

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01747 2023-11-08 cs.CV 78%

UMDFood: Vision-language models boost food composition compilation

Peihua Ma, Yixin Wu, Ning Yu, Yang Zhang, Michael Backes, Qin Wang, Cheng-I Wei

专题命中 其他LLM :language model(title,abstract)

Comments 13 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11430 2023-10-26 cs.AI cs.CL cs.IR cs.LG 78%

TELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex Tasks

Shubhra Kanti Karmaker Santu, Dongji Feng

专题命中 其他LLM :LLM(title);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.06343 2023-10-26 cs.CV 78%

Incorporating Structured Representations into Pretrained Vision & Language Models Using Scene Graphs

Roei Herzig, Alon Mendelson, Leonid Karlinsky, Assaf Arbelle, Rogerio Feris, Trevor Darrell, Amir Globerson

专题命中 其他LLM :language model(title,abstract)

Comments EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.09112 2023-10-24 cs.CL cs.AI cs.LG 78%

Differentially Private Natural Language Models: Recent Advances and Future Directions

Lijie Hu, Ivan Habernal, Lei Shen, Di Wang

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.09199 2023-10-19 cs.CV 78%

PaLI-3 Vision Language Models: Smaller, Faster, Stronger

Xi Chen, Xiao Wang, Lucas Beyer, Alexander Kolesnikov, Jialin Wu, Paul Voigtlaender, Basil Mustafa, Sebastian Goodman, Ibrahim Alabdulmohsin, Piotr Padlewski, Daniel Salz, Xi Xiong, Daniel Vlasic, Filip Pavetic, Keran Rong, Tianli Yu, Daniel Keysers, Xiaohua Zhai, Radu Soricut

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.10125 2023-10-17 cs.CV 78%

Few-shot Action Recognition with Captioning Foundation Models

Xiang Wang, Shiwei Zhang, Hangjie Yuan, Yingya Zhang, Changxin Gao, Deli Zhao, Nong Sang

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.07866 2023-09-21 cs.CV 78%

Gradient constrained sharpness-aware prompt learning for vision-language models

Liangchen Liu, Nannan Wang, Dawei Zhou, Xinbo Gao, Decheng Liu, Xi Yang, Tongliang Liu

专题命中 其他LLM :language model(title,abstract)

Comments 19 pages 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.01528 2023-09-07 cs.RO 78%

Recognition of Heat-Induced Food State Changes by Time-Series Use of Vision-Language Model for Cooking Robot

Naoaki Kanazawa, Kento Kawaharazuka, Yoshiki Obinata, Kei Okada, Masayuki Inaba

专题命中 其他LLM :language model(title,abstract)

Comments Accepted at IAS18-2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13355 2023-08-28 cs.HC 78%

WorldSmith: Iterative and Expressive Prompting for World Building with a Generative AI

Hai Dang, Frederik Brudy, George Fitzmaurice, Fraser Anderson

专题命中 其他LLM :prompting(title,abstract)

Comments User Interface Software and Technology 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11095 2023-08-17 eess.AS cs.AI cs.CL cs.LG cs.SD 78%

Prompting the Hidden Talent of Web-Scale Speech Models for Zero-Shot Task Generalization

Puyuan Peng, Brian Yan, Shinji Watanabe, David Harwath

专题命中 其他LLM :prompting(title);分类 cs.CL、cs.AI、cs.LG

Comments Interspeech 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.07078 2023-08-15 cs.CV 78%

ICPC: Instance-Conditioned Prompting with Contrastive Learning for Semantic Segmentation

Chaohui Yu, Qiang Zhou, Zhibin Wang, Fan Wang

专题命中 其他LLM :prompting(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14331 2023-07-27 cs.CV 78%

Visual Instruction Inversion: Image Editing via Visual Prompting

Thao Nguyen, Yuheng Li, Utkarsh Ojha, Yong Jae Lee

专题命中 其他LLM :prompting(title,abstract)

Comments Project page: https://thaoshibe.github.io/visii/

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.02297 2023-05-04 cs.CV 78%

Making the Most of What You Have: Adapting Pre-trained Visual Language Models in the Low-data Regime

Chuhan Zhang, Antoine Miech, Jiajun Shen, Jean-Baptiste Alayrac, Pauline Luc

专题命中 其他LLM :language model(title,abstract)

Comments Tech Report

详情

展开后加载摘要…

URL PDF HTML 收藏