arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-05 至 2025-12-05 共收录 13 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 13 篇

2511.19730 2025-12-05 cs.LG cond-mat.mtrl-sci 90%

Training-Free Active Learning Framework in Materials Science with Large Language Models

材料科学中基于大语言模型的无需训练主动学习框架

Hongchen Wang, Rafael Espinosa Castañeda, Jay R. Werber, Yao Fehlis, Edward Kim, Jason Hattrick-Simpers

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)

AI总结 本文提出基于大语言模型的无需训练主动学习框架,通过两种提示策略在多个材料科学数据集中显著提升实验效率和搜索效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04273 2025-12-05 cs.SE cs.AI 89%

Quantitative Analysis of Technical Debt and Pattern Violation in Large Language Model Architectures

大语言模型架构中技术债务与模式违规的定量分析

Tyler Slater

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.AI

AI总结 本研究通过对比分析三种大语言模型在实现标准化微服务时的架构一致性与技术债务积累情况,揭示了开源模型在架构合规性方面的不足及潜在的技术债务风险。

Comments Under review at the Journal of Systems and Software (Special Issue on Impactful Software Architecture)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04228 2025-12-05 cs.AI 88%

Addressing Logical Fallacies In Scientific Reasoning From Large Language Models: Towards a Dual-Inference Training Framework

解决大型语言模型在科学推理中的逻辑谬误:一种双推理训练框架

Peter B. Walker, Hannah Davidson, Aiden Foster, Matthew Lienert, Thomas Pardue, Dale Russell

机构 * Intelligenesis LLC Uniformed Services University(美国武装部队服务大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文提出一种双推理训练框架,旨在解决大型语言模型在科学推理中的逻辑谬误问题,通过整合肯定生成与反事实否定,提升模型的鲁棒性和可解释性。

Comments 12 pages, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05066 2025-12-05 cs.LG cs.AI cs.CL 87%

Multi-LLM Collaboration for Medication Recommendation

多LLM协作用于药物推荐

Huascar Sanchez, Briland Hitaj, Jules Bergmann, Linda Briesemeister

机构 * Computer Science Laboratory, SRI International(SRI国际计算机科学实验室) University of Maryland St. Joseph Medical Center(马里兰大学圣约瑟夫医疗中心)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出基于LLM化学的多模型协作方法,通过增强互补性、稳定性和校准性,提高药物推荐的可靠性与可信度。

Comments 8 pages, 5 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09495 2025-12-05 cs.CL cs.LG 81%

Bridging Online Behavior and Clinical Insight: A Longitudinal LLM-based Study of Suicidality on YouTube Reveals Novel Digital Markers

弥合在线行为与临床洞察:一项基于长期LLM研究的YouTube自杀性行为揭示了新的数字标记

Ilanit Sobol, Shir Lissak, Refael Tikochinski, Tal Nakash, Anat Brunstein Klomek, Eyal Fruchter, Roi Reichart

机构 * Technion – Israel Institute of Technology(技术学院–以色列理工学院)

专题命中 领域大模型 :LLM(title,abstract);分类 cs.CL、cs.LG

AI总结 本研究通过分析YouTube自杀尝试者的语言模式,结合LLM和临床专家方法,揭示了新的数字标记,揭示了自杀行为与心理健康之间的关联。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13238 2025-12-05 cs.LG cs.AI cs.CL cs.CY 80%

Computational Measurement of Political Positions: A Review of Text-Based Ideal Point Estimation Algorithms

政治立场的计算测量:基于文本的理想点估计算法综述

Patrick Parschan, Charlott Jakob

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文综述了基于文本的理想点估计算法,分析了四种方法家族的优缺点,为应用研究提供指导。

Comments 46 pages, 8 figures, 2 tables, accepted for publication in Quality & Quantity

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04210 2025-12-05 cs.AI cs.CL cs.CY 79%

Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment

通过迭代偏好对齐平衡医疗AI助手的安全性与帮助性

Huy Nghiem, Swetasudha Panda, Devashish Khatwani, Huy V. Nguyen, Krishnaram Kenthapadi, Hal Daumé

机构 * University of Maryland(马里兰大学) Oracle Labs(Oracle实验室) Oracle Health AI(Oracle健康AI)

专题命中 领域大模型 :large language model(abstract);language model(abstract);preference optimization(abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过迭代偏好对齐框架提升医疗AI助手的安全性与帮助性,通过KTO和DPO优化改进模型,提升有害查询检测性能并平衡安全与效用。

Comments ML4H 2025 Proceedings, Best Paper Award

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09414 2025-12-05 cs.SE 75%

Multi-agent Assisted Automatic Test Generation for Java JSON Libraries

多智能体辅助的Java JSON库自动测试生成

Sinan Wang, Zhiyuan Zhong, Shaojin Wen, Yepang Liu

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 JsonATG通过多智能体系统生成Java JSON库的多样化测试用例,有效提升测试覆盖率并发现多个bug。

Comments In the 32nd Asia-Pacific Software Engineering Conference (APSEC 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04716 2025-12-05 physics.flu-dyn cs.AI 74%

Towards an AI Fluid Scientist: LLM-Powered Scientific Discovery in Experimental Fluid Mechanics

迈向人工智能流体科学家:基于LLM的实验流体力学中的科学发现

Haodong Feng, Lugang Ye, Dixia Fan

机构 * Westlake University(西湖大学)

专题命中 领域大模型 :LLM(title);分类 cs.AI

AI总结 本文提出基于LLM的AI流体科学家框架,实现自主实验流程,提升流体力学研究效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04834 2025-12-05 cs.AI cs.CL cs.IR 73%

Are LLMs Truly Multilingual? Exploring Zero-Shot Multilingual Capability of LLMs for Information Retrieval: An Italian Healthcare Use Case

LLMs真的多语言吗?探索LLMs在信息检索中的零样本多语言能力:一个意大利医疗应用案例

Vignesh Kumar Kembu, Pierandrea Morandini, Marta Bianca Maria Ranzini, Antonino Nocera

机构 * Department of Electrical, Computer and Biomedical Engineering(电气、计算机与生物医学工程系) University of Pavia(帕维亚大学) IRCCS Humanitas Research Hospital(人类itas研究医院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文探讨了开源多语言LLMs在零样本条件下处理意大利语电子健康记录信息提取的能力,发现部分模型在泛化能力上存在不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04113 2025-12-05 cs.CY cs.AI cs.HC cs.LG 73%

AI-Enabled grading with near-domain data for scaling feedback with human-level accuracy

基于近域数据的AI评阅:实现人类水平准确性的反馈扩展

Shyam Agarwal, Ali Moghimi, Kevin C. Haudek

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于近域数据的AI评阅方法,实现人类水平的准确反馈,优于现有机器学习模型和大型语言模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04116 2025-12-05 cs.CY 67%

Mapping the Probabilistic AI Ecosystem in Criminal Justice in England and Wales

映射英格兰和威尔士刑事司法系统中的概率AI生态系统

Evdoxia Taka, Temitope Lawal, Muffy Calder, Michele Sevegnani, Kyriakos Kotsoglou, Elizabeth McClory-Tiarks, Marion Oswald

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 本文研究了英格兰和威尔士刑事司法系统中概率AI工具的映射,分析其应用现状、潜在影响及未来发展方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04520 2025-12-05 cs.CV 50%

Boundary-Aware Test-Time Adaptation for Zero-Shot Medical Image Segmentation

面向边界的测试时适应用于零样本医学图像分割

Chenlin Xu, Lei Zhang, Lituan Wang, Xinyu Pu, Pengfei Ma, Guangwu Qian, Zizhou Wang, Yan Wang

机构 * School of Computer Science, Sichuan University(四川大学计算机学院) Pittsburgh Institute, Sichuan University(四川大学匹兹堡学院) Institute of High Performance Computing, Agency for Science, Technology and Research (A*STAR)(科技研究局高性能计算研究所)

专题命中 领域大模型 :foundation model(abstract)

AI总结 BA-TTA-SAM通过测试时适应提升SAM在医学图像零样本分割中的性能,采用高斯提示注入和边界感知注意力对齐机制,实现12.4%的Dice分数提升。

详情

展开后加载摘要…

URL PDF HTML 收藏