arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12659 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12659 篇

2604.04020 2026-04-07 cs.CL cs.LG 88%

Unmasking Hallucinations: A Causal Graph-Attention Perspective on Factual Reliability in Large Language Models

揭示幻觉:从因果图注意力视角探讨大语言模型中的事实可靠性

Sailesh kiran kurra, Shiek Ruksana, Vishal Borusu

专题命中 领域大模型 :language model(title,abstract);large language model(title);LLM(abstract);分类 cs.CL、cs.LG

AI总结 本文提出因果图注意力网络框架,通过构建token级图来减少大语言模型中的幻觉问题,提升事实可靠性。

Comments Paper accepted for publication at IEEE International Conference on Emerging Computing and Intelligent Technologies 2026 (ICoECIT),5 Pages,5 figures,1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03004 2026-04-01 cs.CL cs.AI 88%

SemioLLM: Evaluating Large Language Models for Diagnostic Reasoning from Unstructured Clinical Narratives in Epilepsy

SemioLLM:评估用于癫痫诊断推理的大型语言模型在无结构临床叙述中的表现

Meghal Dani, Muthu Jeyanthi Prakash, Filip Rosa, Zeynep Akata, Stefanie Liebe

机构 * University of Tübingen(蒂宾根大学) Technical University of Munich(慕尼黑工业大学) University Clinic Tübingen(蒂宾根大学医院) Hertie Institute for Clinical Brain Research(赫蒂临床脑研究所) Excellence Cluster Machine Learning, Tübingen University(蒂宾根大学机器学习卓越集群) Hertie Institute for AI in Brain Health (Hertie AI)(赫蒂人工智能脑健康研究所)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文评估了八个大型语言模型在癫痫诊断任务中的表现,通过过滤和标准化 seizure 描述短语,将其映射到七个可能的癫痫发作起始区。结果显示,经过提示工程后,多数模型表现接近临床水平,但临床情境下的表现受语言上下文影响显著,需改进模型的可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05658 2026-03-31 cs.CL cs.AI 88%

Multilingual Medical Reasoning for Question Answering with Large Language Models

多语言医疗推理用于问答的大型语言模型

Pietro Ferrazzi, Aitor Soroa, Rodrigo Agerri

机构 * HiTZ Center - Ixa, University of the Basque Country EHU(HiTZ中心 - Ixa,巴斯克大学EHU)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于维基百科医疗知识生成多语言推理轨迹的方法,通过检索增强生成技术生成英文、意大利语和西班牙语的50万条轨迹,提升医疗问答性能,达到8B参数模型的最新水平。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.19467 2026-03-31 cs.CL cs.AI 88%

BRIDGE: Benchmarking Large Language Models for Understanding Real-world Clinical Practice Text

BRIDGE:用于评估大语言模型理解现实世界临床实践文本的基准测试

Jiageng Wu, Bowen Gu, Ren Zhou, Kevin Xie, Doug Snyder, Yixing Jiang, Valentina Carducci, Richard Wyss, Rishi J Desai, Emily Alsentzer, Leo Anthony Celi, Adam Rodman, Sebastian Schneeweiss, Jonathan H. Chen, Santiago Romero-Brufau, Kueiyu Joshua Lin, Jie Yang

机构 * Brigham and Women's Hospital(布莱根妇女医院) Harvard Medical School(哈佛医学院) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Massachusetts Institute of Technology(麻省理工学院) Mayo Clinic(梅奥诊所) Harvard T.H. Chan School of Public Health(哈佛大学陈曾熙公共卫生学院) Harvard University(哈佛大学) Stanford University(斯坦福大学) Beth Israel Deaconess Medical Center(贝斯以色列女执事医疗中心) Kempner Institute for the Study of Natural and Artificial Intelligence(肯普纳自然与人工智能研究所) Broad Institute of MIT and Harvard(博德研究所) Harvard Data Science Initiative(哈佛数据科学计划)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出BRIDGE基准测试,涵盖9种语言的87项任务,涵盖患者护理全过程的六个临床阶段和20种应用,评估95种LLM在不同推理策略下的性能差异,展示开源模型与专业模型的性能对比。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23678 2026-03-26 cs.CL cs.AI 88%

PLACID: Privacy-preserving Large language models for Acronym Clinical Inference and Disambiguation

PLACID:隐私保护的大语言模型用于缩写临床推断和消歧

Manjushree B. Aithal, Ph. D., Alexander Kotz, James Mitchell, Ph. D

机构 * Department of Biomedical Informatics, University of Colorado Anschutz(科罗拉多大学安舒茨分校生物医学信息学系)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出PLACID模型,通过本地部署的小参数模型实现隐私保护的临床缩写消歧,利用通用本地模型检测缩写并路由至领域特定的生物医学模型以提高扩展准确性。

Comments 10 pages, 2 figures, Under review AMIA Symposium

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23509 2026-03-26 cs.CL cs.AI cs.CR 88%

Internal Safety Collapse in Frontier Large Language Models

前沿大语言模型中的内部安全崩溃

Yutao Wu, Xiao Liu, Yifeng Gao, Xiang Zheng, Hanxun Huang, Yige Li, Cong Wang, Bo Li, Xingjun Ma, Yu-Gang Jiang

机构 * Deakin University(德克萨斯大学) Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究院) Shanghai Key Laboratory of Multimodal Embodied AI(上海多模态具身人工智能重点实验室) City University of Hong Kong(香港城市大学) The University of Melbourne(墨尔本大学) Singapore Management University(新加坡管理大学) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 研究揭示了前沿大语言模型中一种关键故障模式——内部安全崩溃,通过构建ISC-Bench测试集,发现模型在执行某些任务时会持续生成有害内容,且比传统劫持攻击更危险。

Comments 15 pages of the main text, qualitative examples of jailbreaks may be harmful in nature

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13401 2026-03-24 cs.CL cs.AI 88%

Levels of Analysis for Large Language Models

大型语言模型的分析层次

Alexander Y. Ku, Declan Campbell, Xuechunzi Bai, Jiayi Geng, Ryan Liu, Raja Marjieh, R. Thomas McCoy, Andrew Nam, Ilia Sucholutsky, Veniamin Veselovsky, Liyi Zhang, Jian-Qiao Zhu, Thomas L. Griffiths

机构 * Department of Psychology, Princeton University(普林斯顿大学心理学系) Princeton Neuroscience Institute, Princeton University(普林斯顿神经科学研究所) Department of Psychology, The University of Chicago(芝加哥大学心理学系) Department of Computer Science, Princeton University(普林斯顿大学计算机科学系) Department of Linguistics, Yale University(耶鲁大学语言学系) Princeton Laboratory for Artificial Intelligence, Princeton University(普林斯顿人工智能实验室) Center for Data Science, New York University(纽约大学数据科学中心)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出基于David Marr分析层次框架,利用认知科学方法理解大型语言模型的结构与行为,提供分析工具以应对AI理解难题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19948 2026-03-06 cs.CL cs.AI cs.CY cs.HC cs.MA 88%

Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming

评估大型语言模型在心理健康支持中的风险:一种用于自动化临床AI红队测试的框架

Ian Steenstra, Paola Pedrelli, Weiyan Shi, Stacy Marsella, Timothy W. Bickmore

机构 * Northeastern University(东北大学) Harvard Medical School(哈佛医学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种评估AI心理治疗师在心理健康支持中安全风险的框架,通过模拟测试发现AI在治疗中的潜在风险,并验证了交互式可视化工具的有效性。

Comments This paper is a condensed version of the first author's Ph.D. dissertation submitted to Northeastern University

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22157 2026-02-26 cs.CL cs.HC cs.LG 88%

Dynamic Personality Adaptation in Large Language Models via State Machines

通过状态机实现大语言模型的动态人格适应

Leon Pielage, Ole Hätscher, Mitja Back, Bernhard Marschall, Benjamin Risse

机构 * Institute for Geoinformatics, University of Münster(地理信息研究所,穆尔斯特大学) Faculty of Mathematics and Computer Science, University of Münster(数学与计算机科学学院,穆尔斯特大学) Department of Psychology, University of Münster(心理学系,穆尔斯特大学) Institute of Medical Education and Student Affairs, University of Münster(医学教育与学生事务研究所,穆尔斯特大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG

AI总结 本文提出通过状态机实现大语言模型动态人格适应的框架,通过模块化评分管道实现人格状态的动态调整,提升在复杂互动场景中的表现。

Comments 22 pages, 5 figures, submitted to ICPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11390 2026-02-24 cs.LG cs.AI 88%

Medical Interpretability and Knowledge Maps of Large Language Models

大语言模型的医疗可解释性与知识图谱

Razvan Marinescu, Victoria-Elisabeth Gruber, Diego Fajardo

机构 * Lumos AI

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

AI总结 研究通过四种技术分析大语言模型在医疗领域的可解释性,揭示医学知识在模型中的存储位置及处理机制,为后续医疗任务的模型优化提供指导。

Comments 29 pages, 34 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16715 2026-02-20 cs.AI cs.CL cs.SY eess.SY 88%

Retrieval Augmented (Knowledge Graph), and Large Language Model-Driven Design Structure Matrix (DSM) Generation of Cyber-Physical Systems

检索增强(知识图谱),以及由大型语言模型驱动的面向网络系统的设计结构矩阵(DSM)生成

H. Sinan Bank, Daniel R. Herber

机构 * Department of Systems Engineering, Colorado State University(系统工程系,科罗拉多州立大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文研究了利用大型语言模型、检索增强生成和图基RAG生成面向网络系统的设计结构矩阵,通过两个具体案例评估其在组件关系识别和生成方面的性能。

Comments 26 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.00869 2026-02-19 cs.CL cs.AI 88%

m1: Unleash the Potential of Test-Time Scaling for Medical Reasoning with Large Language Models

m1:通过大语言模型的测试时缩放释放医疗推理的潜力

Xiaoke Huang, Juncheng Wu, Hui Liu, Xianfeng Tang, Yuyin Zhou

机构 * UC Santa Cruz(加州大学圣克ruz分校) Amazon Research(亚马逊研究院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出m1方法,通过测试时缩放提升医疗推理能力,发现最佳推理预算和医学知识不足是关键瓶颈。

Comments 17 pages; 7 figures; Data, code, and models: https://github.com/UCSC-VLAA/m1 ; Accepted by ML4H'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11965 2026-02-13 cs.LG cs.AI 88%

Manifold-Aware Temporal Domain Generalization for Large Language Models

面向流形的时域泛化用于大语言模型

Yiheng Yao, Zekun Cai, Xinyuan Song, Hiroki Hill Kobayashi, Xuan Song, Ryosuke Shibasaki, Liang Zhao

机构 * The University of Tokyo(东京大学) LocationMind Emory University(埃默里大学) Jilin University(吉林大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出MaT-LoRA,通过在低秩适应子空间内约束时间更新到共享低维流形,实现大语言模型的时间域泛化,提升模型的时间建模效率与泛化能力。

Comments 14 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.10937 2026-02-12 cs.AI cs.CL cs.MA 88%

SCALE: Towards Collaborative Content Analysis in Social Science with Large Language Model Agents and Human Intervention

SCALE:基于大语言模型代理和人工干预的社会科学研究内容分析

Chengshuai Zhao, Zhen Tan, Chau-Wai Wong, Xinyan Zhao, Tianlong Chen, Huan Liu

专题命中 领域大模型 :language model(title,abstract);large language model(title);LLM(abstract);分类 cs.CL、cs.AI

AI总结 SCALE通过大语言模型代理和人工干预,实现了社会科学中复杂内容分析的高效模拟与提升。

Comments Accepted by the Annual Meeting of the Association for Computational Linguistics (ACL) 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22042 2026-02-02 cs.CL cs.AI 88%

Emotions Where Art Thou: Understanding and Characterizing the Emotional Latent Space of Large Language Models

情感在哪里?:理解并表征大语言模型的情感潜在空间

Benjamin Reichman, Adar Avsian, Larry Heck

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 研究揭示了大语言模型中情感的潜在空间结构,通过方向编码和跨语言一致性,展示了对情感的可控表征与可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20674 2026-01-29 cs.CL cs.AI 88%

Harnessing Large Language Models for Precision Querying and Retrieval-Augmented Knowledge Extraction in Clinical Data Science

利用大语言模型实现临床数据科学中的精准查询与检索增强的知识提取

Juan Jose Rubio Jan, Jack Wu, Julia Ive

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本研究利用大语言模型在临床数据科学中实现精准查询和检索增强的知识提取,通过实验验证其在结构化数据查询和非结构化文本信息提取中的有效性。

Comments 11 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14268 2026-01-22 cs.CY cs.AI cs.CL 88%

Developmental trajectories of decision making and affective dynamics in large language models

大语言模型决策机制与情感动态的发展轨迹

Zhihao Wang, Yiyang Liu, Ting Wang, Zhiyuan Liu

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 研究通过对比不同大语言模型与人类在赌博任务中的表现,揭示了模型决策和情感动态的发展轨迹及其对AI伦理的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06875 2026-01-13 cs.AI cs.CL 88%

An Ubuntu-Guided Large Language Model Framework for Cognitive Behavioral Mental Health Dialogue

以Ubuntu为指导的认知行为大型语言模型框架用于认知行为心理健康对话

Sontaga G. Forane, Absalom E. Ezugwu, Kevin Igwe, Karen van den Berg

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出一种结合Ubuntu哲学与认知行为疗法的AI心理健康对话系统,旨在提升非洲语境下的文化敏感性和治疗效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06750 2026-01-13 cs.CV cs.AI cs.CL 88%

Benchmarking Egocentric Clinical Intent Understanding Capability for Medical Multimodal Large Language Models

医疗多模态大语言模型的视点临床意图理解能力基准测试

Shaonan Liu, Guo Yu, Xiaoling Luo, Shiyi Zheng, Wenting Chen, Jie Liu, Linlin Shen

机构 * Shenzhen University(深圳大学) Stanford University(斯坦福大学) City University of Hong Kong(香港城市大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出MedGaze-Bench,首个评估医疗多模态大语言模型视点临床意图理解能力的基准测试,通过三维意图框架和陷阱QA机制,揭示现有模型在手术、急救和诊断任务中对意图理解的不足。

Comments 16 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22533 2026-01-12 cs.CL cs.AI 88%

CliCARE: Grounding Large Language Models in Clinical Guidelines for Decision Support over Longitudinal Cancer Electronic Health Records

CliCARE: 在纵向癌症电子健康记录上基于临床指南 grounding 大型语言模型以支持决策

Dongchen Li, Jitao Liang, Wei Li, Xiaoyu Wang, Longbing Cao, Kun Yu

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 CliCARE 通过将纵向癌症 EHRs 转换为时序知识图谱并结合临床指南,为肿瘤科医生提供证据支持的决策支持。

Comments Accepted in AAAI Conference on Artificial Intelligence (AAAI-26, Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00797 2026-01-06 cs.CL cs.AI cs.CY cs.MA 88%

The Qualitative Laboratory: Theory Prototyping and Hypothesis Generation with Large Language Models

定性实验室:利用大语言模型进行社会学人设模拟与假设生成

Hugues Draelants

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出利用大语言模型进行社会学人设模拟,通过生成自然对话生成深入的定性假设,挑战传统方法的局限性。

Comments 26 pages, 3 tables. Manuscript submitted for peer-reviewed journal publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22738 2025-12-30 cs.CL cs.AI 88%

Harnessing Large Language Models for Biomedical Named Entity Recognition

利用大型语言模型进行生物医学命名实体识别

Jian Chen, Leilei Su, Cong Sun

机构 * Department of Data Science and Big Data Technology, Hainan University, Haikou 570228, China(数据科学与大数据技术学院,海南大学,海口570228,中国) Department of Mathematics, Hainan University, Haikou 570228, China(数学学院,海南大学,海口570228,中国) Department of Population Health Sciences, Weill Cornell Medicine, New York 10022, USA(流行病学与公共卫生科学学院,韦尔·柯尔医学中心,纽约10022,美国)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出BioSelectTune框架,通过高效的数据筛选方法提升生物医学命名实体识别的性能,实现优于现有模型的准确率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.21112 2025-12-17 cs.AI cs.CE cs.CL cs.CY cs.IR 88%

Optimizing Large Language Models for ESG Activity Detection in Financial Texts

优化大型语言模型以检测金融文本中的ESG活动

Mattia Birti, Andrea Maurino, Francesco Osborne

机构 * Department of Informatics, Systems and Communication, University of Milano-Bicocca(信息学、系统与通信系,米兰-比科卡大学) University of Milano-Bicocca(米兰-比科卡大学) The Open University(开放大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过微调优化大型语言模型,提升金融文本中ESG活动检测的准确性。

Comments Published in the Proceedings of the ACM International Conference on AI in Finance (ICAIF). ACM version

Journal ref Proceedings of the ACM International Conference on AI in Finance (ICAIF), 2024, ACM

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.13732 2025-12-12 cs.CL cond-mat.mtrl-sci cs.LG 88%

Enhancing Large Language Models with Domain-Specific Knowledge: The Case in Topological Materials

通过领域特定知识增强大型语言模型:拓扑材料案例

HuangChao Xu, Baohua Zhang, Zhong Jin, Tiannian Zhu, Quansheng Wu, Hongming Weng

机构 * Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心) University of Chinese Academy of Sciences(中国科学院大学) Institute of Physics, Chinese Academy of Sciences(中国科学院物理研究所)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG

AI总结 本文提出通过构建材料知识图谱与提示学习,开发专用对话系统TopoChat,以提升拓扑材料领域的信息检索与知识交互能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01812 2025-11-27 cs.CY cs.AI cs.CL 88%

From Text to Multimodality: Exploring the Evolution and Impact of Large Language Models in Medical Practice

从文本到多模态:探索大型语言模型在医疗实践中的演变与影响

Qian Niu, Keyu Chen, Ming Li, Pohsun Feng, Ziqian Bi, Lawrence KQ Yan, Yichao Zhang, Caitlyn Heqi Yin, Cheng Fei, Junyu Liu, Tianyang Wang, Yunze Wang, Silin Chen, Ming Liu, Benji Peng, Xinyuan Song, Ziyuan Qin, Riyang Bao, Zekun Jiang

机构 * Kyoto University(京都大学) Georgia Institute of Technology(佐治亚理工学院) National Taiwan Normal University(台湾师范大学) Indiana University(印第安纳大学) Hong Kong University of Science(香港科学大学) The University of Texas at Dallas(德克萨斯大学达拉斯分校) University of Wisconsin-Madison(威斯康星大学麦迪逊分校) Cornell University(康奈尔大学) University of Liverpool(利物浦大学) University of Edinburgh(爱丁堡大学) Zhejiang University(浙江大学) Purdue University(Purdue 大学) Emory University, Atlanta, GA, USA(埃默里大学) West China Biomedical Big Data Center, West China Hospital, Sichuan University, Chengdu, China(西京生物大数据中心,四川大学西京医院,成都,中国)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文探讨了多模态大型语言模型在医疗实践中的发展与影响,分析其在医疗影像、临床决策支持等领域的应用及面临的挑战。

Comments 12 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.02387 2025-11-27 cs.AI cs.CL 88%

Large Language Models and Cognitive Science: A Comprehensive Review of Similarities, Differences, and Challenges

大语言模型与认知科学:对相似性、差异性和挑战的全面综述

Qian Niu, Junyu Liu, Ziqian Bi, Pohsun Feng, Benji Peng, Keyu Chen, Ming Li, Lawrence KQ Yan, Yichao Zhang, Caitlyn Heqi Yin, Cheng Fei, Tianyang Wang, Yunze Wang, Silin Chen, Ming Liu, Ziyuan Qin, Riyang Bao, Xinyuan Song, Zekun Jiang

机构 * Kyoto University(京都大学) Indiana University(印第安纳大学) National Taiwan Normal University(台湾师范大学) Georgia Institute of Technology(佐治亚理工学院) Hong Kong University of Science(香港科学大学) The University of Texas at Dallas(德克萨斯大学达拉斯分校) University of Wisconsin-Madison(威斯康星大学麦迪逊分校) Cornell University(康奈尔大学) University of Liverpool, UK(利物浦大学) University of Edinburgh, UK(爱丁堡大学) Zhejiang University(浙江大学) Purdue University(普渡大学) Emory University(埃默里大学) West China Biomedical Big Data Center, West China Hospital, Sichuan University(四川大学西昌生物大数据中心、西昌医院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文综述了大语言模型与认知科学的相似性、差异性和挑战,探讨了LLMs在认知领域的应用及改进方法。

Comments 10 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09621 2025-11-26 cs.RO cs.AI cs.LG 88%

Interpretable Robot Control via Structured Behavior Trees and Large Language Models

通过结构化行为树和大语言模型实现可解释的机器人控制

Ingrid Maéva Chekam, Ines Pastor-Martinez, Ali Tourani, Jose Andres Millan-Romera, Laura Ribeiro, Pedro Miguel Bastos Soares, Holger Voos, Jose Luis Sanchez-Lopez

机构 * Automation and Robotics Research Group (ARG), Interdisciplinary Centre for Security, Reliability, and Trust (SnT), University of Luxembourg(自动化与机器人研究组(ARG)、安全、可靠性与信任跨学科中心(SnT)、卢森堡大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出结合大语言模型与结构化行为树的方法,实现机器人对自然语言指令的可解释执行,提升人机交互的实用性与适应性。

Comments 15 pages, 5 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09825 2025-11-25 cs.CL cs.AI 88%

GP-GPT: Large Language Model for Gene-Phenotype Mapping

GP-GPT:用于基因-表型映射的大型语言模型

Yanjun Lyu, Zihao Wu, Lu Zhang, Jing Zhang, Yiwei Li, Wei Ruan, Zhengliang Liu, Zeyu Zhang, Xiang Li, Rongjie Liu, Chao Huang, Wentao Li, Tianming Liu, Dajiang Zhu

机构 * Department of Computer Science and Engineering, The University of Texas at Arlington(计算机科学与工程系,德克萨斯理工大学 Arlington 分校) School of Computing, University of Georgia(计算学院,佐治亚大学) Department of Computer Science, Indiana University Indianapolis(计算机科学系,印第安纳大学 Indianapolis 分校) Department of Radiology, Massachusetts General Hospital and Harvard Medical School(放射科,麻省总医院和哈佛医学院) Department of Statistics, University of Georgia(统计系,佐治亚大学) Department of Epidemiology & Biostatistics, University of Georgia(流行病学与生物统计学系,佐治亚大学) Department of Environmental Health Science, University of Georgia(环境流行病学系,佐治亚大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 GP-GPT是一种专为基因-表型映射设计的大型语言模型,通过在基因组和医学遗传学领域进行微调,实现了对遗传信息的高效检索和分析,并在多个任务中超越了现有最先进的LLMs。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17012 2025-11-24 cs.CL cs.AI 88%

Supervised Fine Tuning of Large Language Models for Domain Specific Knowledge Graph Construction:A Case Study on Hunan's Historical Celebrities

为构建特定领域知识图谱对大型语言模型进行监督微调:以湖南历史名人案例研究

Junjie Hao, Chun Wang, Ying Qiao, Qiuyue Zuo, Qiya Song, Hua Ma, Xieping Gao

机构 * College of Information Science and Engineering(信息科学与工程学院) Hunan Provincial Key Laboratory of Philosophy and Social Sciences of Yuelushan Cultural and Digital Communication (Artificial Intelligence and International Communication AIIC)(湖南省级哲学社会科学岳麓文化与数字传播重点实验室(人工智能与国际传播AIIC)) College of Computer Science and Electronic Engineering(计算机科学与电子工程学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本研究通过监督微调提升大型语言模型在湖南历史名人领域知识提取能力,验证了参数高效方法在低资源环境下的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09613 2025-11-24 cs.CL cs.AI 88%

Task-Aligned Tool Recommendation for Large Language Models

面向大型语言模型的任务对齐工具推荐

Hang Gao, Yongfeng Zhang

机构 * Rutgers University(罗切斯特大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种以精度为导向的工具推荐方法,旨在为大型语言模型提供定制化的工具集,以提高解决复杂问题的效率。

Comments IJCNLP-AACL 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏