arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12637 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12637 篇

2509.19125 2025-09-24 cs.CL 87%

Context-Aware Hierarchical Taxonomy Generation for Scientific Papers via LLM-Guided Multi-Aspect Clustering

Kun Zhu, Lizi Liao, Yuxuan Gu, Lei Huang, Xiaocheng Feng, Bing Qin

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted to EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10621 2025-07-16 cs.CR cs.AI cs.CY cs.GT 87%

Game Theory Meets LLM and Agentic AI: Reimagining Cybersecurity for the Age of Intelligent Threats

Quanyan Zhu

机构 * Quanyan Zhu

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23663 2025-07-01 cs.CV cs.LG 87%

On the Domain Robustness of Contrastive Vision-Language Models

Mario Koddenbrock, Rudolf Hoffmann, David Brodmann, Erik Rodner

机构 * KI-Werkstatt/Fachbereich 2, University of Applied Sciences Berlin(柏林应用技术大学人工智能工作室/第二学系)

专题命中 领域大模型 :language model(title,abstract);LLM(abstract);large language model(abstract);foundation model(abstract)

Comments Deepbench is available at https://github.com/ml-lab-htw/deepbench

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19107 2025-07-01 cs.HC cs.AI 87%

Improving Student-AI Interaction Through Pedagogical Prompting: An Example in Computer Science Education

Ruiwei Xiao, Xinying Hou, Runlong Ye, Majeed Kazemitabaar, Nicholas Diana, Michael Liut, John Stamper

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Michigan(密歇根大学) University of Toronto(多伦多大学) Colgate University(科尔盖特大学) University of Toronto Mississauga(多伦多大学密西根分校)

专题命中 领域大模型 :prompting(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

Comments Under review for Elsevier Journal. Journal policy allows submitting as preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05069 2025-06-10 cs.IR cs.AI 87%

Reason-to-Recommend: Using Interaction-of-Thought Reasoning to Enhance LLM Recommendation

Keyu Zhao, Fengli Xu, Yong Li

机构 * Department of Electronic Engineering, Tsinghua University(电子工程系,清华大学)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17604 2025-04-24 cs.AI 87%

OmniScience: A Domain-Specialized LLM for Scientific Reasoning and Discovery

Vignesh Prabhakar, Md Amirul Islam, Adam Atanas, Yao-Ting Wang, Joah Han, Aastha Jhunjhunwala, Rucha Apte, Robert Clark, Kang Xu, Zihan Wang, Kai Liu

机构 * SES AI(SES人工智能公司) NVIDIA(英伟达公司)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);instruction tuning(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17784 2025-03-25 cs.AI 87%

MEPNet: Medical Entity-balanced Prompting Network for Brain CT Report Generation

Xiaodan Zhang, Yanzhao Shi, Junzhong Ji, Chengxin Zheng, Liangqiong Qu

专题命中 领域大模型 :prompting(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

Comments AAAI 2025 Oral Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05010 2025-03-10 cs.CL 87%

Leveraging Domain Knowledge at Inference Time for LLM Translation: Retrieval versus Generation

Bryan Li, Jiaming Luo, Eleftheria Briakou, Colin Cherry

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19410 2025-02-27 cs.HC cs.AI 87%

Less or More: Towards Glanceable Explanations for LLM Recommendations Using Ultra-Small Devices

Xinru Wang, Mengjie Yu, Hannah Nguyen, Michael Iuzzolino, Tianyi Wang, Peiqi Tang, Natasha Lynova, Co Tran, Ting Zhang, Naveen Sendhilnathan, Hrvoje Benko, Haijun Xia, Tanya Jonker

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.17630 2025-02-13 cs.IR cs.CL 87%

Uncertainty Quantification and Decomposition for LLM-based Recommendation

Wonbin Kweon, Sanghwan Jang, SeongKu Kang, Hwanjo Yu

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments WWW 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13630 2025-01-22 cs.LG 87%

UniGraph: Learning a Unified Cross-Domain Foundation Model for Text-Attributed Graphs

Yufei He, Yuan Sui, Xiaoxin He, Bryan Hooi

专题命中 领域大模型 :foundation model(title,abstract);large language model(abstract);language model(abstract);instruction tuning(abstract)

Comments KDD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02552 2025-01-07 cs.CL cs.CV 87%

Multi-LLM Collaborative Caption Generation in Scientific Documents

Jaeyoung Kim, Jongho Lee, Hong-Jun Choi, Ting-Yao Hsu, Chieh-Yang Huang, Sungchul Kim, Ryan Rossi, Tong Yu, Clyde Lee Giles, Ting-Hao 'Kenneth' Huang, Sungchul Choi

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted to AAAI 2025 AI4Research Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12075 2024-12-31 cs.CV cs.AI 87%

WeatherDG: LLM-assisted Diffusion Model for Procedural Weather Generation in Domain-Generalized Semantic Segmentation

Chenghao Qian, Yuhu Guo, Yuhong Mo, Wenjing Li

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08847 2024-12-13 cs.IR cs.LG 87%

MOPI-HFRS: A Multi-objective Personalized Health-aware Food Recommendation System with LLM-enhanced Interpretation

Zheyuan Zhang, Zehong Wang, Tianyi Ma, Varun Sameer Taneja, Sofia Nelson, Nhi Ha Lan Le, Keerthiram Murugesan, Mingxuan Ju, Nitesh V Chawla, Chuxu Zhang, Yanfang Ye

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10237 2024-09-04 cs.CV cs.CL 87%

Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models

Songtao Jiang, Tuo Zheng, Yan Zhang, Yeying Jin, Li Yuan, Zuozhu Liu

专题命中 领域大模型 :language model(title,abstract);LLM(abstract);large language model(abstract);instruction tuning(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21264 2024-08-15 cs.CL 87%

Model Attribution in LLM-Generated Disinformation: A Domain Generalization Approach with Supervised Contrastive Learning

Alimohammad Beigi, Zhen Tan, Nivedh Mudiam, Canyu Chen, Kai Shu, Huan Liu

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments 10 pages, 2 figures, accepted at DSAA 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.21170 2024-08-01 cs.CL cs.HC 87%

Decomposed Prompting to Answer Questions on a Course Discussion Board

Brandon Jaipersaud, Paul Zhang, Jimmy Ba, Andrew Petersen, Lisa Zhang, Michael R. Zhang

专题命中 领域大模型 :prompting(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

Comments 6 pages. Published at International Conference on Artificial Intelligence in Education 2023. Code repository: https://github.com/brandonjaipersaud/piazza-qabot-gpt

Journal ref In: Artificial Intelligence in Education. AIED 2023. Communications in Computer and Information Science, vol 1831. Springer, Cham

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.18992 2024-07-30 cs.AI 87%

Towards Automated Solution Recipe Generation for Industrial Asset Management with LLM

Nianjun Zhou, Dhaval Patel, Shuxin Lin, Fearghal O'Donncha

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17887 2024-07-01 cs.CL cs.IR 87%

JMLR: Joint Medical LLM and Retrieval Training for Enhancing Reasoning and Professional Question Answering Capability

Junda Wang, Zhichao Yang, Zonghai Yao, Hong Yu

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.07478 2024-03-13 cs.IR cs.LG 87%

Towards Graph Foundation Models for Personalization

Andreas Damianou, Francesco Fabbri, Paul Gigioli, Marco De Nadai, Alice Wang, Enrico Palumbo, Mounia Lalmas

专题命中 领域大模型 :foundation model(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.15328 2024-01-31 cs.CL 87%

Equipping Language Models with Tool Use Capability for Tabular Data Analysis in Finance

Adrian Theuma, Ehsan Shareghi

专题命中 领域大模型 :language model(title,abstract);LLM(abstract);large language model(abstract);SFT(abstract)

Comments Accepted to EACL2024; code, model and dataset are available at https://raven-lm.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.08939 2023-09-19 cs.IR cs.AI 87%

An Unified Search and Recommendation Foundation Model for Cold-Start Scenario

Yuqi Gong, Xichen Ding, Yehui Su, Kaiming Shen, Zhongyi Liu, Guannan Zhang

专题命中 领域大模型 :foundation model(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

Comments CIKM 2023,6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.03333 2023-08-21 cs.IR cs.AI 87%

Heterogeneous Knowledge Fusion: A Novel Approach for Personalized Recommendation via LLM

Bin Yin, Junjie Xie, Yu Qin, Zixiang Ding, Zhichao Feng, Xiang Li, Wei Lin

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);instruction tuning(abstract)

Comments Accepted at RecSys 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.04811 2023-06-09 cs.CV cs.AI 87%

Generative Text-Guided 3D Vision-Language Pretraining for Unified Medical Image Segmentation

Yinda Chen, Che Liu, Wei Huang, Sibo Cheng, Rossella Arcucci, Zhiwei Xiong

专题命中 领域大模型 :pretraining(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.03153 2023-04-07 cs.IR cs.CL 87%

Zero-Shot Next-Item Recommendation using Large Pretrained Language Models

Lei Wang, Ee-Peng Lim

专题命中 领域大模型 :language model(title,abstract);LLM(abstract);large language model(abstract);prompting(abstract)

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09839 2025-07-24 eess.SP cs.IT math.IT 87%

AI and Deep Learning for Terahertz Ultra-Massive MIMO: From Model-Driven Approaches to Foundation Models

Wentao Yu, Hengtao He, Shenghui Song, Jun Zhang, Linglong Dai, Lizhong Zheng, Khaled B. Letaief

专题命中 领域大模型 :foundation model(title,abstract);LLM(abstract);large language model(abstract);language model(abstract)

Comments 30 pages, 8 figures, 1 table, accepted by Engineering. Model-driven deep learning, CSI foundation models, and applications of LLMs are presented as three systematic research roadmaps for AI-enabled THz ultra-massive MIMO systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.20887 2026-08-24 cs.CL cs.AI 新提交 87%

KREL: Automatic Medical Coding via Knowledge-Guided Reasoning over Clinical Evidence with LLMs

KREL:基于大型语言模型对临床证据进行知识引导推理的自动医学编码方法

Xubin Chen, Yipeng Zhou, Wen Sun, Chengkai Huang, Xiaoming Fu, Quan Z. Sheng

机构 * The University of New South Wales(新南威尔士大学) University of Göttingen(哥廷根大学) Macquarie University(麦考瑞大学) Beijing Intelligent Decision Medical Technology Co. Ltd(北京智决医疗科技有限公司)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究提出KREL框架,将LLM与ICD编码指南知识结合解决自动医学编码问题,在基准数据集上性能优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.20369 2026-08-24 cs.CL cs.AI 新提交 87%

ASTAR: Automated induction of STAndardized radiology Reporting templates from large-scale clinical free-text corpora

ASTAR:从大规模临床自由文本语料库自动生成标准化放射学报告模板

Xinfeng Zhang, Mingxuan Liu, Yifei Chen, Juncheng Zhu, Kasidit Anmahapong, Yiming Huang, Yuan Zhang, Hongjia Yang, Yi Liao, Gang Ning, Haibo Qu, Qiyuan Tian

机构 * Tsinghua University(清华大学) Sichuan University(四川大学) University of California San Diego(加利福尼亚大学圣迭戈分校) Southeast University(东南大学)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 该研究针对放射学报告模板构建的人工瓶颈,提出基于LLM的ASTAR框架,从大规模临床自由文本语料库自动生成标准化模板,实验显示其性能优于专家构建的模板,大幅缩短模板开发时间。

Comments Accepted by MICCAI

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.20345 2026-08-24 cs.CL cs.AI cs.CY 新提交 87%

When Vocabulary Comprehension Fails Clinical Reasoning: Evaluating Therapy Bots' Safety Risks for Generation Alpha

当词汇理解失效于临床推理:评估面向阿尔法世代(Gen Alpha,2010-2024年出生)的治疗机器人的安全风险

Manisha Mehta, Virendra Mehta

机构 * Lynbrook High School(林布鲁克高中) University of Trento(特伦托大学)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本研究构建两个基准评估Claude、GPT-4o等LLM的治疗机器人安全风险,发现其存在10-14个百分点的词汇理解与临床风险校准差距,识别六种失败模式,提出四项监管与技术改进建议。

Comments 42 pages, 6 figures. Accepted at ACM FAccT '26

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.19206 2026-08-21 cs.CL cs.AI cs.MA 新提交 87%

Hallucination as a Feature, not a Defect: Evaluating a multi-agent architecture to transform speculative language-model outputs into testable scientific hypotheses

幻觉作为特征而非缺陷:评估一种将推测性语言模型输出转化为可检验科学假说的多智能体架构

Nicolas Rodriguez-Alvarez

机构 * IES Parquesol(帕尔奎索尔学院)

专题命中 领域大模型 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 该研究提出基于Rust的多智能体架构,通过生成与评估智能体的认识论摩擦循环将LLM的推测性输出转化为可检验假说,实验显示该架构在需经受严格约束时更具优势,且各架构在原创性等维度的平衡表现不同。

Comments 25 pages. Bilingual: full English version followed by the complete Spanish version. Includes an exploratory paired baseline and ablation study (6 conditions). Code and data: https://doi.org/10.5281/zenodo.20649714

详情

展开后加载摘要…

URL PDF HTML 收藏