arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12229 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12229 篇

2409.14605 2024-09-25 eess.SY cs.SY 78%

First Field Trial of LLM-Powered AI Agent for Lifecycle Management of Autonomous Driving Optical Networks

Xiaomin Liu, Qizhi Qiu, Yihao Zhang, Yuming Cheng, Lilin Yi, Weisheng Hu, Qunbi Zhuge

专题命中 其他LLM :LLM(title,abstract)

Comments Version submitted to ECOC PDP 2024 on September 6th

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.14083 2024-09-24 cs.CV 78%

SURf: Teaching Large Vision-Language Models to Selectively Utilize Retrieved Information

Jiashuo Sun, Jihai Zhang, Yucheng Zhou, Zhaochen Su, Xiaoye Qu, Yu Cheng

专题命中 其他LLM :language model(title,abstract)

Comments 19 pages, 9 tables, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12304 2024-09-24 cs.RO 78%

MAGIC-VFM: Meta-learning Adaptation for Ground Interaction Control with Visual Foundation Models

Elena Sorina Lupu, Fengze Xie, James A. Preiss, Jedidiah Alindogan, Matthew Anderson, Soon-Jo Chung

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11494 2024-09-19 eess.AS cs.SD 78%

M-BEST-RQ: A Multi-Channel Speech Foundation Model for Smart Glasses

Yufeng Yang, Desh Raj, Ju Lin, Niko Moritz, Junteng Jia, Gil Keren, Egor Lakomkin, Yiteng Huang, Jacob Donley, Jay Mahadeokar, Ozlem Kalinli

专题命中 其他LLM :foundation model(title,abstract)

Comments In submission to IEEE ICASSP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11214 2024-09-18 eess.AS cs.SD 78%

Ideal-LLM: Integrating Dual Encoders and Language-Adapted LLM for Multilingual Speech-to-Text

Hongfei Xue, Wei Ren, Xuelong Geng, Kun Wei, Longhao Li, Qijie Shao, Linju Yang, Kai Diao, Lei Xie

专题命中 其他LLM :LLM(title,abstract)

Comments 5 pages, 3 figures, submitted to ICASSP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09338 2024-09-17 cs.SI cs.HC 78%

What you say or how you say it? Predicting Conflict Outcomes in Real and LLM-Generated Conversations

Priya Ronald D'Costa, Evan Rowbotham, Xinlan Emily Hu

专题命中 其他LLM :LLM(title);language model(abstract)

Comments Submitted to the NeurIPS 2024 Workshop on Behavioral ML

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09276 2024-09-17 cs.RO 78%

Visuo-Tactile Zero-Shot Object Recognition with Vision-Language Model

Shiori Ueda, Atsushi Hashimoto, Masashi Hamaya, Kazutoshi Tanaka, Hideo Saito

专题命中 其他LLM :language model(title,abstract)

Comments 9 pages, 9 figures, accepted to IROS2024, project page: https://omron-sinicx.github.io/visuo-tactile-recognition/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18197 2024-09-12 cs.CV 78%

Human-Free Automated Prompting for Vision-Language Anomaly Detection: Prompt Optimization with Meta-guiding Prompt Scheme

Pi-Wei Chen, Jerry Chun-Wei Lin, Jia Ji, Feng-Hao Yeh, Zih-Ching Chen, Chao-Chun Chen

专题命中 其他LLM :prompting(title);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.06090 2024-09-11 q-bio.BM 78%

AbGPT: De Novo Antibody Design via Generative Language Modeling

Desmond Kuan, Amir Barati Farimani

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.05413 2024-09-10 cs.CV cs.RO 78%

From Words to Poses: Enhancing Novel Object Pose Estimation with Vision Language Models

Tessa Pulli, Stefan Thalhammer, Simon Schwaiger, Markus Vincze

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14977 2024-09-10 cs.CV 78%

A Lost Opportunity for Vision-Language Models: A Comparative Study of Online Test-Time Adaptation for Vision-Language Models

Mario Döbler, Robert A. Marsden, Tobias Raichle, Bin Yang

专题命中 其他LLM :language model(title);foundation model(abstract)

Comments Accepted at ECCV 2024 OOD-CV Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.04508 2024-09-10 cs.HC 78%

Toward LLM-Powered Social Robots for Supporting Sensitive Disclosures of Stigmatized Health Conditions

Alemitu Bezabih, Shadi Nourriz, C. Estelle Smith

专题命中 其他LLM :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.03525 2024-09-06 cs.CV 78%

FrozenSeg: Harmonizing Frozen Foundation Models for Open-Vocabulary Segmentation

Xi Chen, Haosen Yang, Sheng Jin, Xiatian Zhu, Hongxun Yao

专题命中 其他LLM :foundation model(title,abstract)

Comments 14 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.16373 2024-08-30 cs.SD eess.AS 78%

Enabling Beam Search for Language Model-Based Text-to-Speech Synthesis

Zehai Tu, Guangyan Zhang, Yiting Lu, Adaeze Adigwe, Simon King, Yiwen Guo

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18915 2024-08-30 cs.RO cs.CV 78%

Manipulate-Anything: Automating Real-World Robots using Vision-Language Models

Jiafei Duan, Wentao Yuan, Wilbert Pumacay, Yi Ru Wang, Kiana Ehsani, Dieter Fox, Ranjay Krishna

专题命中 其他LLM :language model(title,abstract)

Comments Project page: https://robot-ma.github.io/. All supplementary material, prompts and code can be found on the project page

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.15377 2024-08-15 cs.CV 78%

InternVideo2: Scaling Foundation Models for Multimodal Video Understanding

Yi Wang, Kunchang Li, Xinhao Li, Jiashuo Yu, Yinan He, Chenting Wang, Guo Chen, Baoqi Pei, Ziang Yan, Rongkun Zheng, Jilan Xu, Zun Wang, Yansong Shi, Tianxiang Jiang, Songze Li, Hongjie Zhang, Yifei Huang, Yu Qiao, Yali Wang, Limin Wang

专题命中 其他LLM :foundation model(title,abstract)

Comments a technical report about video understanding (accepted to ECCV2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.16696 2024-07-24 cs.CV 78%

PartGLEE: A Foundation Model for Recognizing and Parsing Any Objects

Junyi Li, Junfeng Wu, Weizhi Zhao, Song Bai, Xiang Bai

专题命中 其他LLM :foundation model(title,abstract)

Comments Accepted by ECCV2024, homepage: https://provencestar.github.io/PartGLEE-Vision/

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09829 2024-07-16 cs.RO 78%

VLMPC: Vision-Language Model Predictive Control for Robotic Manipulation

Wentao Zhao, Jiaming Chen, Ziyu Meng, Donghui Mao, Ran Song, Wei Zhang

专题命中 其他LLM :language model(title,abstract)

Comments Accepted by RSS2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08156 2024-07-12 cs.CV 78%

AddressCLIP: Empowering Vision-Language Models for City-wide Image Address Localization

Shixiong Xu, Chenghao Zhang, Lubin Fan, Gaofeng Meng, Shiming Xiang, Jieping Ye

专题命中 其他LLM :language model(title,abstract)

Comments Accepted at ECCV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06634 2024-07-10 cs.CR 78%

Stealing Part of a Production Language Model

Nicholas Carlini, Daniel Paleka, Krishnamurthy Dj Dvijotham, Thomas Steinke, Jonathan Hayase, A. Feder Cooper, Katherine Lee, Matthew Jagielski, Milad Nasr, Arthur Conmy, Itay Yona, Eric Wallace, David Rolnick, Florian Tramèr

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.13534 2024-07-08 cs.HC 78%

Prompting AI Art: An Investigation into the Creative Skill of Prompt Engineering

Jonas Oppenlaender, Rhema Linder, Johanna Silvennoinen

专题命中 其他LLM :prompting(title,abstract)

Comments 42 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17998 2024-06-27 cs.CV 78%

Changen2: Multi-Temporal Remote Sensing Generative Change Foundation Model

Zhuo Zheng, Stefano Ermon, Dongjun Kim, Liangpei Zhang, Yanfei Zhong

专题命中 其他LLM :foundation model(title,abstract)

Comments The enhanced extension of our ICCV 2023 (Changen)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04452 2024-06-10 cs.HC 78%

Revisiting Human Information Foraging: Adaptations for LLM-based Chatbots

Sruti Srinivasa Ragavan, Mohammad Amin Alipour

专题命中 其他LLM :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.20305 2024-05-31 cs.CV 78%

Can't make an Omelette without Breaking some Eggs: Plausible Action Anticipation using Large Video-Language Models

Himangi Mittal, Nakul Agarwal, Shao-Yuan Lo, Kwonjoon Lee

专题命中 其他LLM :language model(title,abstract)

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.19638 2024-05-31 cs.CV 78%

Learning Robust Correlation with Foundation Model for Weakly-Supervised Few-Shot Segmentation

Xinyang Huang, Chuang Zhu, Kebin Liu, Ruiying Ren, Shengjie Liu

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07472 2024-05-24 cs.CV 78%

GaussianVTON: 3D Human Virtual Try-ON via Multi-Stage Gaussian Splatting Editing with Image Prompting

Haodong Chen, Yongle Huang, Haojian Huang, Xiangsheng Ge, Dian Shao

专题命中 其他LLM :prompting(title,abstract)

Comments On-going work

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.12979 2024-05-22 cs.CV 78%

OmniGlue: Generalizable Feature Matching with Foundation Model Guidance

Hanwen Jiang, Arjun Karpur, Bingyi Cao, Qixing Huang, Andre Araujo

专题命中 其他LLM :foundation model(title,abstract)

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09164 2024-05-16 quant-ph 78%

Rapidly Achieving Chemical Accuracy with Quantum Computing Enforced Language Model

Honghui Shang, Xiongzhi Zeng, Ming Gong, Yangju Wu, Shaojun Guo, Haoran Qian, Chen Zha, Zhijie Fan, Kai Yan, Xiaobo Zhu, Zhenyu Li, Yi Luo, Jian-Wei Pan, Jinlong Yang

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05346 2024-05-10 cs.CV 78%

VLM-PL: Advanced Pseudo Labeling Approach for Class Incremental Object Detection via Vision-Language Model

Junsu Kim, Yunhoe Ku, Jihyeon Kim, Junuk Cha, Seungryul Baek

专题命中 其他LLM :language model(title,abstract)

Comments Accept to CVPRW2024 (CLvision). The camera-ready version of the manuscript

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.19321 2024-04-19 hep-th gr-qc 78%

Chaotic LLM billiards

David Berenstein, Elliot Maderazo, Robinson Mancilla, Anayeli Ramirez

专题命中 其他LLM :LLM(title,abstract)

Comments 18 pages, 8 figures, uses JHEP. v2: Typos corrected, references added

详情

展开后加载摘要…

URL PDF HTML 收藏