arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2508.03566 2025-08-06 cs.CV 71%

SAM2-UNeXT: An Improved High-Resolution Baseline for Adapting Foundation Models to Downstream Segmentation Tasks

Xinyu Xiong, Zihuang Wu, Lei Zhang, Lei Lu, Ming Li, Guanbin Li

机构 * Sun Yat-sen University(中山大学) Jiangxi Normal University(江西师范大学) Hainan University(海南大学) Shandong Inspur Database Technology Co., Ltd(山东 Inspur 数据库技术有限公司)

专题命中 其他LLM :foundation model(title)

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18889 2025-07-28 cs.AR cs.DC cs.NI 71%

RailX: A Flexible, Scalable, and Low-Cost Network Architecture for Hyper-Scale LLM Training Systems

Yinxiao Feng, Tiancheng Chen, Yuchen Wei, Siyuan Shen, Shiju Wang, Wei Li, Kaisheng Ma, Torsten Hoefler

专题命中 其他LLM :LLM(title)

Comments 25 pages, 21 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11470 2025-07-16 cs.HC 71%

REVA: Supporting LLM-Generated Programming Feedback Validation at Scale Through User Attention-based Adaptation

Xiaohang Tang, Sam Wong, Zicheng He, Yalong Yang, Yan Chen

专题命中 其他LLM :LLM(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16813 2025-06-23 cs.CE 71%

Integrating Traditional Technical Analysis with AI: A Multi-Agent LLM-Based Approach to Stock Market Forecasting

Michał Wawer, Jarosław A. Chudziak

专题命中 其他LLM :LLM(title)

Comments 12 pages, 8 figures, 1 table. This is the accepted version of the paper presented at the 17th International Conference on Agents and Artificial Intelligence (ICAART 2025), Porto, Portugal

Journal ref Proceedings of the 17th International Conference on Agents and Artificial Intelligence - Volume 1 (ICAART 2025), pages 100-111. SciTePress, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13282 2025-06-17 cs.CV 71%

Anomaly Object Segmentation with Vision-Language Models for Steel Scrap Recycling

Daichi Tanaka, Takumi Karasawa, Shu Takenouchi, Rei Kawakami

机构 * Institute Science of Tokyo(东京科学研究所)

专题命中 其他LLM :language model(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.21026 2025-06-02 cs.CR cs.PL cs.SE 71%

Artemis: Toward Accurate Detection of Server-Side Request Forgeries through LLM-Assisted Inter-Procedural Path-Sensitive Taint Analysis

Yuchen Ji, Ting Dai, Zhichao Zhou, Yutian Tang, Jingzhu He

专题命中 其他LLM :LLM(title)

Comments Full version of paper accepted by OOPSLA '25

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16954 2025-05-23 cs.HC 71%

Cracking Aegis: An Adversarial LLM-based Game for Raising Awareness of Vulnerabilities in Privacy Protection

Jiaying Fu, Yiyang Lu, Zehua Yang, Fiona Nah, RAY LC

专题命中 其他LLM :LLM(title)

Comments 24 pages, In Designing Interactive Systems Conference (DIS 25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15189 2025-04-23 cs.HC 71%

LACE: Controlled Image Prompting and Iterative Refinement with GenAI for Professional Visual Art Creators

Yenkai Huang, Ning Zheng

专题命中 其他LLM :prompting(title)

Comments Accepted at the GenAICHI Workshop at CHI 2025 (4 pages) For comprehensive methods, analysis, and extended discussions, please refer to the full-length version at [arXiv:2504.14827]

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18701 2025-04-23 cs.HC 71%

LLM-Driven Optimization of HTML Structure to Support Screen Reader Navigation

Yaman Yu, Bektur Ryskeldiev, Ayaka Tsutsui, Matthew Gillingham, Yang Wang

专题命中 其他LLM :LLM(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13173 2024-12-18 cs.CV 71%

Locate n' Rotate: Two-stage Openable Part Detection with Foundation Model Priors

Siqi Li, Xiaoxue Chen, Haoyu Cheng, Guyue Zhou, Hao Zhao, Guanzhong Tian

专题命中 其他LLM :foundation model(title)

Comments ACCV 2024 Oral, Project: https://github.com/lisiqi-zju/MOPD

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09219 2024-11-15 cs.CV 71%

Harnessing Vision Foundation Models for High-Performance, Training-Free Open Vocabulary Segmentation

Yuheng Shi, Minjing Dong, Chang Xu

专题命中 其他LLM :foundation model(title)

Comments 12 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.03354 2024-11-12 cs.CR cs.NI 71%

LLM-based Continuous Intrusion Detection Framework for Next-Gen Networks

Frederic Adjewa, Moez Esseghir, Leila Merghem-Boulahia

专题命中 其他LLM :LLM(title)

Comments 8 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19967 2024-10-01 cs.CV 71%

Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function

Chenyi Zhuang, Ying Hu, Pan Gao

专题命中 其他LLM :language model(title)

Comments Accepted to NeurIPS 2024. Code is available at https://github.com/I2-Multimedia-Lab/Magnet

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.13148 2024-09-23 cs.CV 71%

UniTabNet: Bridging Vision and Language Models for Enhanced Table Structure Recognition

Zhenrong Zhang, Shuhang Liu, Pengfei Hu, Jiefeng Ma, Jun Du, Jianshu Zhang, Yu Hu

专题命中 其他LLM :language model(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.06643 2024-09-11 cs.HC 71%

Strategic management analysis: from data to strategy diagram by LLM

Richard Brath, Adam Bradley, David Jonker

专题命中 其他LLM :LLM(title)

Comments NLVIZ Workshop at IEEE VIZ 2024. 7 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.05817 2024-09-10 cs.CV cs.HC 71%

VFA: Vision Frequency Analysis of Foundation Models and Human

Mohammad-Javad Darvishi-Bayazi, Md Rifat Arefin, Jocelyn Faubert, Irina Rish

专题命中 其他LLM :foundation model(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03842 2024-08-13 q-bio.BM 71%

PepMLM: Target Sequence-Conditioned Generation of Therapeutic Peptide Binders via Span Masked Language Modeling

Tianlai Chen, Madeleine Dumas, Rio Watson, Sophia Vincoff, Christina Peng, Lin Zhao, Lauren Hong, Sarah Pertsemlidis, Mayumi Shaepers-Cheu, Tian Zi Wang, Divya Srijay, Connor Monticello, Pranay Vure, Rishab Pulugurta, Kseniia Kholina, Shrey Goel, Matthew P. DeLisa, Ray Truant, Hector C. Aguilar, Pranam Chatterjee

专题命中 其他LLM :language model(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.06573 2024-07-10 cs.SE 71%

LLM for Mobile: An Initial Roadmap

Daihang Chen, Yonghui Liu, Mingyi Zhou, Yanjie Zhao, Haoyu Wang, Shuai Wang, Xiao Chen, Tegawendé F. Bissyandé, Jacques Klein, Li Li

专题命中 其他LLM :LLM(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.10448 2024-06-18 eess.AS cs.SD 71%

AVR: Synergizing Foundation Models for Audio-Visual Humor Detection

Sarthak Sharma, Orchid Chetia Phukan, Drishti Singh, Arun Balaji Buduru, Rajesh Sharma

专题命中 其他LLM :foundation model(title)

Comments Accepted to INTERSPEECH 2024 Show & Tell Demonstrations

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13532 2024-05-24 cs.CV 71%

What Makes Good Few-shot Examples for Vision-Language Models?

Zhaojun Guo, Jinghui Lu, Xuejing Liu, Rui Zhao, ZhenXing Qian, Fei Tan

专题命中 其他LLM :language model(title)

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.01926 2024-05-06 cs.CV 71%

Auto-Encoding Morph-Tokens for Multimodal LLM

Kaihang Pan, Siliang Tang, Juncheng Li, Zhaoyu Fan, Wei Chow, Shuicheng Yan, Tat-Seng Chua, Yueting Zhuang, Hanwang Zhang

专题命中 其他LLM :LLM(title)

Comments Accepted by ICML 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09766 2024-03-18 cs.CV 71%

An Image Is Worth 1000 Lies: Adversarial Transferability across Prompts on Vision-Language Models

Haochen Luo, Jindong Gu, Fengyuan Liu, Philip Torr

专题命中 其他LLM :language model(title)

Comments Accepted to ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05538 2024-03-13 cs.CY cs.HC cs.SE 71%

Explaining Code Examples in Introductory Programming Courses: LLM vs Humans

Arun-Balajiee Lekshmi-Narayanan, Priti Oli, Jeevan Chapagain, Mohammad Hassany, Rabin Banjade, Peter Brusilovsky, Vasile Rus

专题命中 其他LLM :LLM(title)

Comments 3 tables; 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.08108 2023-09-18 cs.SD eess.AS 71%

Foundation Model Assisted Automatic Speech Emotion Recognition: Transcribing, Annotating, and Augmenting

Tiantian Feng, Shrikanth Narayanan

专题命中 其他LLM :foundation model(title)

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.03465 2023-09-01 eess.AS 71%

A generative framework for conversational laughter: Its 'language model' and laughter sound synthesis

Hiroki Mori, Shunya Kimura

专题命中 其他LLM :language model(title)

Comments Submitted to INTERSPEECH

Journal ref Proc. Interspeech 2023 (2023) 3372-3376

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14587 2023-07-28 q-bio.BM 71%

Artificial intelligence-aided protein engineering: from topological data analysis to deep protein language models

Yuchi Qiu, Guo-Wei Wei

专题命中 其他LLM :language model(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.09026 2023-07-19 cs.CV 71%

ActionPrompt: Action-Guided 3D Human Pose Estimation With Text and Pose Prompting

Hongwei Zheng, Han Li, Bowen Shi, Wenrui Dai, Botao Wan, Yu Sun, Min Guo, Hongkai Xiong

专题命中 其他LLM :prompting(title)

Comments 6 pages, 4 figures, 2023ICME

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.03369 2023-03-10 cs.CV 71%

Multimodal Prompting with Missing Modalities for Visual Recognition

Yi-Lun Lee, Yi-Hsuan Tsai, Wei-Chen Chiu, Chen-Yu Lee

专题命中 其他LLM :prompting(title)

Comments Accepted by CVPR 2023. Codes are available at https://github.com/YiLunLee/Missing_aware_prompts

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.07202 2023-02-09 cs.CY 71%

Mask and Cloze: Automatic Open Cloze Question Generation using a Masked Language Model

Shoya Matsumori, Kohei Okuoka, Ryoichi Shibata, Minami Inoue, Yosuke Fukuchi, Michita Imai

专题命中 其他LLM :language model(title)

Comments 14 pages, 8 figures

Journal ref IEEE Access, vol. 11, pp. 9835-9850, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.09302 2021-02-03 eess.IV cs.GR 71%

Michelson Holography: Dual-SLM Holography with Camera-in-the-loop Optimization

Suyeon Choi, Jonghyun Kim, Yifan Peng, Gordon Wetzstein

专题命中 其他LLM :SLM(title)

详情

展开后加载摘要…

URL PDF HTML 收藏