arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12635 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12635 篇

2605.28524 2026-05-28 cs.AI 90%

Let Relations Speak: An End-to-End LLM-GNN Soft Prompt Framework for Fraud Detection

让关系说话:面向欺诈检测的端到端LLM-GNN软提示框架

Zhixing Zuo, Huilin He, Jiasheng Wu, Dawei Cheng

机构 * School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院)

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出LGSPF框架,通过软提示桥接图结构与语义空间,并引入并行GNN编码器将多关系拓扑转化为图令牌,实现端到端优化,在欺诈检测中达到最优性能。

Comments 14 pages,3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27853 2026-05-28 cs.AI 90%

MolLingo: Molecule-Native Representations for LLM-Powered Scientific Agents

MolLingo:面向LLM驱动的科学智能体的分子原生表示

Thao Nguyen, Heng Ji

机构 * Siebel School of Computing and Data Science, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校Siebel计算与数据科学学院)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.AI

AI总结 提出MolLingo多智能体系统,通过共享内存协调文献、化学家和编排智能体,结合基于BRICS的片段枚举(BFE)表示方法,实现分子块级推理与编辑,在四个基准上优于前沿LLM和专用基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27333 2026-05-27 cs.CL 90%

FinHarness: An Inline Lifecycle Safety Harness for Finance LLM Agents

FinHarness:面向金融LLM代理的内联生命周期安全约束框架

Haoxuan Jia, Yang Liu, Bin Chong, Yingguang Yang, Yancheng Chen, Jiayu Liang, Qian Li, Hanning Lu, Kefu Xu, Hao Zheng, Chongyang Zhang, Hao Peng, Philip S. Yu

机构 * Peking University(北京大学) Nanyang Technological University(南洋理工大学) Tsinghua University(清华大学) University of Science and Technology of China(中国科学技术大学) University of Chinese Academy of Sciences(中国科学院大学) Soochow University(苏州大学) Beijing University of Posts and Telecommunications(北京邮电大学) University of Leeds(利兹大学) Fullive Innovation (Beijing) AI Technology Co., Ltd.(全维创新(北京)人工智能科技有限公司) Beihang University(北京航空航天大学) University of Illinois Chicago(伊利诺伊大学芝加哥分校)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.CL

AI总结 针对金融LLM代理在阻止提示诱导的未授权操作与批准合法多步骤业务流程之间的冲突,提出FinHarness内联安全约束框架,通过查询监控、工具监控和级联模块实现逐步骤风险评估与自适应验证,显著降低攻击成功率并保持良性批准率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27071 2026-05-27 cs.AI 90%

Traceable Knowledge Graph Reasoning Enables LLM-Assisted Decision Support for Industrial VOCs in the Steel Industry

可溯源知识图谱推理助力钢铁行业工业VOCs的LLM辅助决策支持

Changqing Su, Yu Ding, Zuhong Lin, Hongyu Liu, Xi He, Zheng Zeng, Liqing Li

机构 * Hunan Key Laboratory of Carbon Neutrality and Intelligent Energy, School of New Energy and Environment, Hunan University of Technology and Business(湖南碳中和与智能能源重点实验室,新能环境学院,湖南科技商务大学) Aerospace Kaitian Environmental Technology Co., Ltd.(航天凯天环境科技有限公司) School of Energy Science and Engineering, Central South University(能源科学与工程学院,中南大学)

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 针对钢铁行业VOCs治理知识分散、通用大模型易产生幻觉的问题,提出基于知识图谱增强的多智能体问答系统Chat-ISV,通过拓扑优化、多智能体路由和源回溯检索实现高可靠性决策支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.04940 2026-05-27 cs.AI 90%

ReVEL: Multi-Turn Reflective LLM-Guided Heuristic Evolution via Structured Performance Feedback

ReVEL:基于结构化性能反馈的多轮反思式LLM引导的启发式进化

Cuong Van Duc, Minh Nguyen Dinh Tuan, Tam Vu Duc, Tung Vu Duy, Son Nguyen Van, Hanh Nguyen Thi, Binh Huynh Thi Thanh

机构 * Hanoi University of Science and Technology(河内科学技术大学) Phenikaa University(Phenikaa大学)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.AI

AI总结 针对NP-hard组合优化问题的启发式设计,提出ReVEL框架,通过行为感知分组和多轮迭代细化,利用LLM和累积性能反馈联合优化启发式,实验表明优于现有LLM引导的进化基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.23916 2026-05-26 cs.IR cs.AI econ.GN q-fin.EC 90%

Agent-Facing Information Design in LLM Tool Registries

面向智能体的LLM工具注册表信息设计

Haochuan Kevin Wang

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.AI

AI总结 本研究首次系统性地分析了LLM工具注册表中广告式描述对智能体选择的影响,发现法律上允许的夸大宣传(如主观最高级表述)完全主导优化效果,而虚假声明无额外影响,并提出了分离选择导向与营销导向描述及智能体注意力质量分数等注册表设计建议。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00682 2026-05-26 cs.IR cs.AI 90%

RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment

RecGOAT: 用于LLM增强多模态推荐的图最优自适应传输与双语义对齐

Yuecheng Li, Hengwei Ju, Zeyu Song, Wei Yang, Chi Lu, Peng Jiang, Kun Gai

机构 * Fudan University(复旦大学) University of Southern California(南加州大学)

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 针对生成式语言模型表示与ID协同信号之间的语义异质性,提出基于图神经网络和最优传输理论的双粒度语义对齐框架RecGOAT,通过实例级和分布级对齐实现统一特征空间,理论证明其表示误差更低,实验达到最优性能。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.18890 2026-05-20 physics.soc-ph cs.AI cs.CY cs.MA 90%

Stop Drawing Scientific Claims from LLM Social Simulations Without Robustness Audits

不要在没有充分鲁棒性审计的情况下从LLM社会模拟中绘制科学结论

Jinyi Ye, Lei Cao, Ding Chen, Emilio Ferrara

机构 * Thomas Lord Department of Computer Science, University of Southern California(美国南加州大学计算机科学系汤姆·劳德部门) Marshall School of Business, University of Southern California(美国南加州大学马歇尔商学院)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.AI

AI总结 本文研究了从LLM社会模拟中得出的科学结论不应强于支持它们的鲁棒性审计,通过两个案例研究展示了小扰动如何影响模拟结果,并提出TRAILS框架以规范鲁棒性审计。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.08949 2026-05-18 cs.LG 90%

Muon-OGD: Muon-based Spectral Orthogonal Gradient Projection for LLM Continual Learning

Muon-OGD: 基于muon的谱范数正交梯度投影用于大语言模型持续学习

Binghang Lu, Zheyuan Deng, Runyu Zhang, Bing Hu, Yunhan Zhao, Yuan Tian, Changhong Mou, Guang Lin, Xiaomin Li

机构 * Purdue University(普渡大学) Brown University(布朗大学) Massachusetts Institute of Technology(麻省理工学院) University of California, Irvine(加州大学 Irvine 分校) Utah State University(犹他州立大学) Harvard University(哈佛大学)

专题命中 领域大模型 :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出Muon-OGD框架,通过谱范数约束优化和正交投影约束,改进LLM持续学习的稳定性-可塑性平衡,实验证明其在持续学习基准上优于传统方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.11532 2026-05-13 cs.AI 90%

Read, Grep, and Synthesize: Diagnosing Cross-Domain Seed Exposure for LLM Research Ideation

阅读、grep和合成:诊断跨领域种子暴露对LLM研究构想的影响

Yunju Choi, Min Song

机构 * Yonsei University, Seoul, Republic of Korea(延世大学,首尔,韩国)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.AI

AI总结 本文通过PaperGym框架研究跨领域检索对LLM构想系统的影响,发现种子暴露提升特定性,但检索效果不如随机多样种子控制,表明系统需进一步优化语义理解。

Comments 12 pages, 2 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.07723 2026-05-11 cs.DL cs.AI cs.CY physics.soc-ph 90%

LLM hallucinations in the wild: Large-scale evidence from non-existent citations

在现实世界中大型语言模型的幻觉:来自不存在引用的大规模证据

Zhenyue Zhao, Yihe Wang, Toby Stuart, Mathijs De Vaan, Paul Ginsparg, Yian Yin

机构 * Department of Information Science, Cornell University(信息科学系,康奈尔大学) Department of Sociology, University of California Los Angeles(社会学系,加州大学洛杉矶分校) Department of Computer Science and Technology, Tsinghua University(计算机科学与技术系,清华大学) Haas School of Business, University of California Berkeley(哈斯商学院,加州大学伯克利分校)

专题命中 领域大模型 :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究通过验证引用数据揭示LLM生成虚假引用的问题,发现2025年存在146932个虚假引用,且在AI应用快速发展的领域和语言特征显示AI辅助写作的论文中尤为严重,影响科学认可的公平性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27872 2026-05-01 cs.AI 90%

Modeling Clinical Concern Trajectories in Language Model Agents

语言模型代理中临床关注轨迹建模

Sukesh Subaharan, Venkatesan VS, Murugadasan P, Sivakumar D, Gautham N, Ganeshkumar M

专题命中 领域大模型 :LLM(summary_cn,abstract);language model(title,abstract);large language model(abstract);分类 cs.AI

AI总结 本文研究通过显式状态动态揭示临床关注轨迹,提出轻量级架构以生成连续升级压力信号,使LLM代理更符合临床实际需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00476 2026-04-28 cs.CL 90%

Remembering Unequally: Global and Disciplinary Bias in LLM Reconstruction of Scholarly Coauthor Lists

铭记不均:大型语言模型在重构学术合著者名单中的全球和学科偏见

Ghazal Kalhor, Afra Mashhadi

机构 * Computing and Software Systems, University of Washington, Bothell, WA, USA(华盛顿大学Bothell分校计算机与软件系统学院)

专题命中 领域大模型 :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究探讨了大型语言模型在重构学术合著者名单时的偏见问题,发现高引用学者更受青睐,但某些学科和地区表现更均衡,揭示了依赖LLM生成知识的风险与局限。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.17738 2026-04-21 cs.CL 90%

Mira-Embeddings-V1: Domain-Adapted Semantic Reranking for Recruitment via LLM-Synthesized Data

Mira-Embeddings-V1:基于LLM合成数据的领域适应语义重排序用于招聘

Zhaohua Liang, Zhilin Wang, Renjie Cao, Yining Zhang

机构 * OpenJobs AI

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.CL

AI总结 本文提出Mira-Embeddings-V1,通过LLM合成数据重塑嵌入空间并纠正边界混淆,提升招聘领域召回率和精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.16317 2026-04-21 cs.IR cs.AI 90%

Paper2Data: Large-Scale LLM Extraction and Metadata Structuring of Global Urban Data from Scientific Literature

Paper2Data: 大规模LLM提取和全球城市数据元数据结构化从科学文献

Runwen You, Tong Xia, Jingzhi Wang, Jiankun Zhang, Tengyao Tu, Jinghua Piao, Yi Chang, Yong Li

机构 * Jilin University(吉林大学) Zhongguancun Academy(中关村学院) Tsinghua University(清华大学)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.AI

AI总结 本文提出Paper2Data方法,通过大规模LLM驱动管道自动提取科学文献中的城市数据集并结构化元数据,构建了UrbanDataMiner平台,实现了全球城市数据的发现与系统化研究。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12385 2026-04-15 cs.CL 90%

From Myopic Selection to Long-Horizon Awareness: Sequential LLM Routing for Multi-Turn Dialogue

从短视选择到长周期意识:面向多轮对话的序列LLM路由

Jiarui Zhang, Xiangyu Liu, Yong Hu, Chaoyue Niu, Hang Zeng, Shaojie Tang, Fan Wu, Guihai Chen

机构 * School of Computer Science, Shanghai Jiao Tong University, Shanghai, China(上海交通大学计算机科学学院) WeChat, Tencent Inc, Beijing, China(腾讯北京研发中心) University at Buffalo, New York, United States(布法罗大学)

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出DialRouter,通过MCTS探索对话分支并学习轻量路由策略,实现多轮对话的高效路由,提升任务成功率并优化性能成本比。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06269 2025-11-11 cs.LG q-bio.QM 90%

LLM$^3$-DTI: A Large Language Model and Multi-modal data co-powered framework for Drug-Target Interaction prediction

Yuhao Zhang, Qinghong Guo, Qixian Chen, Liuwei Zhang, Hongyan Cui, Xiyi Chen

机构 * Polytechnic Institute(多技术学院) School of Pharmaceutical Sciences(药学学院) Innovation Center of Yangtze River Delta(长江三角洲创新中心) Agricultural Genomics Institute at Shenzhen(深圳农业基因组研究院) School of Public Health(公共卫生学院)

专题命中 领域大模型 :LLM(title,abstract);large language model(title);language model(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22287 2025-09-29 cs.RO cs.AI cs.HC 90%

Leveraging Large Language Models for Robot-Assisted Learning of Morphological Structures in Preschool Children with Language Vulnerabilities

Stina Sundstedt, Mattias Wingren, Susanne Hägglund, Daniel Ventus

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

Comments 12 pages, 2 figures, Preprint of: Sundstedt, S., Wingren, M., Hägglund, S. & Ventus, D. (2025). Leveraging Large Language Models for Robot-Assisted Learning of Morphological Structures in Preschool Children with Language Vulnerabilities. In: Stephanidis, C., Antona, M., Ntoa, S. & Salvendy, G. (eds.), Communications in Computer and Information Science, vol. 2523, pp. 415-425. Springer

Journal ref Communications in Computer and Information Science(2025). Stephanidis, C., Antona, M., Ntoa, S. & Salvendy, G. (eds.). p. 415-425 11 p. ( Communications in Computer and Information Science; vol. 2523)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.07653 2025-02-24 cs.AI cs.LO 90%

Large Language Models for Interpretable Mental Health Diagnosis

Brian Hyeongseok Kim, Chao Wang

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

Comments Accepted at AAAI 2025 Workshop on Large Language Models and Generative AI for Health (GenAI4Health)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.13013 2024-07-19 cs.CY cs.AI 90%

FernUni LLM Experimental Infrastructure (FLEXI) -- Enabling Experimentation and Innovation in Higher Education Through Access to Open Large Language Models

Torsten Zesch, Michael Hanses, Niels Seidel, Piush Aggarwal, Dirk Veiel, Claudia de Witt

专题命中 领域大模型 :LLM(title,abstract);large language model(title);language model(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14732 2026-08-17 cs.LG cs.AI cs.CV eess.IV 版本更新 90%

INFORM-CT: INtegrating LLMs and VLMs FOR Incidental Findings Management in Abdominal CT

INFORM-CT:整合LLM和VLM用于腹部CT的偶发发现管理

Idan Tankel, Nir Mazor, Rafi Brada, Christina LeBedis, Guy ben-Yosef

机构 * GE Healthcare Technology and Innovation Center(GE医疗技术与创新中心) Boston Medical Center(波士顿医疗中心)

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于LLM和VLM的计划-执行框架,用于提高腹部CT偶发发现的检测、分类和报告效率与精度,通过自动化流程提升临床应用效果。

Comments Spotlight presentation at the 9th International Conference on Medical Imaging with Deep Learning (MIDL) 2026 Additional code and implementation details available at https://idan-tankel.github.io/InformCT_ProjectPage/

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01826 2026-08-17 cs.CL cs.AI 版本更新 90%

Leveraging Few-Shot Learning and Large Language Models for Analyzing Blood Pressure Variations Across Biological Sex from Scientific Literature

利用小样本学习与大语言模型从科学文献中分析不同生物性别间的血压差异

Yuting Guo, Seyedeh Somayyeh Mousavi, Reza Sameni, Abeed Sarker

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本研究利用小样本学习、LLaMA3及GPT-3.5等大语言模型,从PubMed文献中提取血压相关信息,分析不同生物性别间的血压差异,生成可视化图表开展研究。

Comments Accepted by the journal of Computers in Biology and Medicine

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.09548 2026-08-12 cs.CL cs.AI cs.CY 版本更新 90%

ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models

ELBench:面向教育场景的大语言模型多维基准

Yilin Jiang, Xiaorong Zhu, Fei Tan, Zicheng Zhang, Kaiyi Huang, Yang Yu, Zexuan Fei, Yiming Luo, Keqian Li, Hao Hao, Guangtao Zhai, Aimin Zhou

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);post-training(abstract);分类 cs.CL、cs.AI

AI总结 ELBench是首个评估教育大模型四项核心要求的综合基准,评估发现通用模型综合表现相近但模块优势不同,中国模型在安全性模块领先,教育专用模型未在教育模块占优,高阶培养存在系统性盲点。

Comments 13 pages, 6 figures, 8 tables. Benchmark data: https://huggingface.co/datasets/ZeroLoss-Lab/ELBench

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.28098 2026-07-31 cs.AI cs.CL 新提交 90%

SciDataSailor: Deep Scientific Data Exploring

SciDataSailor:深度科学数据探索

Jiyong Rao, Yicheng Qiu, Chi Zhang, Chunfeng Song, Runkai Zhao

专题命中 领域大模型 :LLM(summary_cn,abstract);SFT(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 该研究针对科学数据仓库交互难题,提出SciDataSailor框架,以带特定机制的MCTS实现轨迹合成,构建了微调模型与含千余任务的评估基准,推动LLM智能体的科学数据探索能力。

Comments 63 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26670 2026-07-30 cs.DL cs.AI cs.CL cs.IR 新提交 90%

Scientific Knowledge Discovery in the Age of Large Language Models

大语言模型时代的科学知识发现

Eleni Adamidi, Serafeim Chatzopoulos, Thanasis Vergoulis

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL、cs.AI

AI总结 本章综述34篇应用生成式大语言模型的同行评审论文,针对学术文献检索与合格研究筛选任务,基于OpenAIRE Graph布尔搜索筛选文献,为科学知识发现提供更灵活方案。

Comments 21 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12393 2026-07-27 cs.CL cs.AI 版本更新 90%

MedKGent: A Large Language Model Agent Framework for Constructing Temporally Evolving Medical Knowledge Graph

MedKGent:用于构建随时间演变的医学知识图谱的大语言模型智能体框架

Duzhen Zhang, Zixiao Wang, Zhong-Zhi Li, Yahan Yu, Shuncheng Jia, Jiahua Dong, Haotian Xu, Xing Wu, Yingying Zhang, Tielin Zhang, Jie Yang, Xiuying Chen, Le Song

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德大学人工智能学院) University of Chinese Academy of Sciences(中国科学院大学) Kyoto University(京都大学) Tsinghua University(清华大学) East China Normal University(华东师范大学) Center for Excellence in Brain Science and Intelligence Technology(脑科学与智能技术卓越中心) Brigham and Women’s Hospital, Harvard Medical School(哈佛医学院布里特妇女医院) GenBio AI

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 研究针对医学文献增长带来的知识结构化挑战,引入MedKGent框架,利用PubMed摘要通过两个智能体每日增量构建医学知识图谱,经评估其三元组有效性高,能显著改进大语言模型在医学问答基准上的检索增强生成。

Comments Accepted by Npj Digital Medicine

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.12336 2026-07-15 cs.CL cs.AI cs.CY cs.ET cs.HC 新提交 90%

Evaluating Health Misinformation in Low-Resource Languages: Integrating Small Language Models with a Culturally-Sensitive Responsible NLP Framework (Bangla as a Case Study)

评估低资源语言中的健康错误信息:将小语言模型与文化敏感的负责任自然语言处理框架相结合(以孟加拉语为例)

Farnaz Farid, Raihan Alam, Al Al-Areqi, Farhad Ahamed, Muhammad Hassan Khan, Sadia Hossain, Irena Veljanova, Anika Tabassum Binte Hossain

机构 * Western Sydney University(西悉尼大学) Microsoft(微软公司) Excelsia College(埃克塞尔西亚学院) Faulconbridge Health Centre(福尔康布里奇健康中心)

专题命中 领域大模型 :language model(title,abstract);small language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 研究针对低资源语言中健康错误信息难检测问题,提出结合小语言模型与文化敏感的负责任自然语言处理框架,以孟加拉语为例进行实验,证明Phi-4表现优,还设计新框架,为评估低资源语言错误信息提供整体视角。

Comments 39 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06078 2026-06-26 cs.CY cs.AI cs.CL 90%

Simulating Students with Large Language Models: A Review of Architecture, Mechanisms, and Role Modelling in Education with Generative AI

用大型语言模型模拟学生:生成式AI在教育中的架构、机制与角色建模综述

Luis Marquez-Carpintero, Alberto Lopez-Sellers, Miguel Cazorla

机构 * Institute for Computer Research University of Alicante(计算机研究所阿利坎特大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI

AI总结 本文综述了利用大型语言模型模拟学生行为在教育中的应用,探讨了其在学习者建模、教学评估和教师培训中的潜力与挑战。

Journal ref Computer Science Review 62 (2026) 101008

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20929 2026-06-23 cs.CL cs.AI 新提交 90%

Peeking Inside LLMs: Leveraging Internal Artifacts of LLMs for Enhancing Reliability in Legal Classification

窥视LLM内部:利用LLM的内部构件增强法律分类的可靠性

Sudipta Santra, Debtanu Datta, Saptarshi Ghosh

机构 * Indian Institute of Technology Kharagpur(印度理工学院卡哈拉格普尔分校)

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 针对LLM在法律领域易产生错误或幻觉的问题,提出利用模型内部构件特征构建下游分类器检测预测正确性,在保释决策和法规违规预测任务上验证了有效性。

Comments Accepted at the International Workshop on Automated Semantic Analysis of Information in Law (ASAIL) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15314 2026-06-16 cs.LG cs.AI stat.ML 新提交 90%

LLMs on Tabular Data with Limited Semantics: Evidence from Industrial Car Retrofit Prediction

有限语义表格数据上的LLM:来自工业汽车改造预测的证据

Aina Vila Pons, Ioannis Tzachristas, Constantinos Antoniou

机构 * Technical University of Munich(慕尼黑工业大学) BMW Group(宝马集团)

专题命中 领域大模型 :LLM(title_cn,summary_cn);foundation model(abstract);prompting(abstract);分类 cs.AI、cs.LG

AI总结 研究在工业表格数据中,LLM(嵌入、直接分类、混合堆叠)与经典树集成方法的对比,发现LLM在语义受限时效果有限,但嵌入和混合方法仍有价值。

详情

展开后加载摘要…

URL PDF HTML 收藏