arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-15 至 2026-01-15 共收录 14 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 14 篇

2512.07890 2026-01-15 cs.MA cs.AI cs.LG stat.ME stat.ML 86%

CrowdLLM: Building LLM-Based Digital Populations Augmented with Generative Models

CrowdLLM: 构建基于大语言模型的数字人群并整合生成模型

Ryan Feng Lin, Keyu Tian, Hanming Zheng, Congjing Zhang, Li Zeng, Shuai Huang

机构 * Department of Industrial and Systems Engineering, University of Washington(华盛顿大学工业与系统工程系) Department of Data Science, City University of Hong Kong(香港城市大学数据科学系)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 CrowdLLM通过整合预训练大语言模型和生成模型,提升数字人群的多样性和保真度,适用于社交模拟、众包等多领域应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08858 2026-01-15 cs.SE cs.AI cs.CY 85%

Adaptive Trust Metrics for Multi-LLM Systems: Enhancing Reliability in Regulated Industries

多LLM系统的自适应信任度量:在受监管行业中的可靠性增强

Tejaswini Bollikonda

机构 * Independent Researcher(独立研究者)

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出了一种用于多LLM系统的自适应信任度量框架,通过分析系统行为和不确定性评估,提升受监管行业中的AI可靠性与安全性。

Comments 8 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16841 2026-01-15 cs.CV cs.AI cs.LG 81%

Fair Foundation Models for Medical Image Analysis: Challenges and Perspectives

医疗图像分析中的公平基础模型:挑战与展望

Dilermando Queiroz, Anderson Carlos, André Anjos, Lilian Berton

机构 * Federal University of São Paulo(巴西圣保罗联邦大学) Federal Institute of Goiás(哥亚斯联邦理工学院) Idiap Research Institute(Idiap研究机构)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文探讨了医疗图像分析中公平基础模型的挑战与前景,强调通过系统性干预和政策参与实现公平性,以推动医疗技术的普及。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09398 2026-01-15 cs.CL cs.AI cs.LG 80%

Ability Transfer and Recovery via Modularized Parameters Localization

通过模块化参数定位实现能力迁移与恢复

Songyao Jin, Kun Zhou, Wenqi Li, Peng Wang, Biwei Huang

机构 * University of California San Diego(加州大学圣地亚哥分校)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 ACT通过激活差异局部化能力相关通道,实现能力迁移与恢复,提升多语言数学和科学推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09449 2026-01-15 cs.CV 78%

PrivLEX: Detecting legal concepts in images through Vision-Language Models

PrivLEX:通过视觉-语言模型检测图像中的法律概念

Darya Baranouskaya, Andrea Cavallaro

机构 * EPFL(苏黎世联邦理工学院) Idiap Research Institute(Idiap研究机构)

专题命中 领域大模型 :language model(title,abstract)

AI总结 PrivLEX通过视觉-语言模型实现图像中法律概念的检测,无需显式标签即可进行可解释的隐私分类。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08998 2026-01-15 cs.SE 78%

On the Flakiness of LLM-Generated Tests for Industrial and Open-Source Database Management Systems

关于LLM生成测试在工业和开源数据库管理系统中的不稳定性

Alexander Berndt, Thomas Bach, Rainer Gemulla, Marcus Kessel, Sebastian Baltes

专题命中 领域大模型 :LLM(title,abstract)

AI总结 研究发现LLM生成的测试在工业和开源数据库管理系统中存在不稳定性问题,主要源于对不确定顺序的依赖,并强调了为LLM提供定制上下文的重要性。

Comments 12 pages, 5 tables, 3 figures, accepted at the 48th International Conference on Software Engineering: Software Engineering in Practice (ICSE SEIP 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23903 2026-01-15 cs.CV 78%

Scaling Remote Sensing Foundation Models: Data Domain Tradeoffs at the Peta-Scale

在百亿级数据规模下扩展遥感基础模型:在百亿级数据规模下的领域权衡

Charith Wickrema, Eliza Mace, Hunter Brown, Heidys Cabrera, Nick Krall, Matthew O'Neill, Shivangi Sarkar, Lowell Weissman, Eric Hughes, Guido Zarrella

机构 * The MITRE Corporation(MITRE公司)

专题命中 领域大模型 :foundation model(title,abstract)

AI总结 研究探讨了在百亿级数据规模下扩展遥感基础模型的挑战与权衡,通过训练更大规模的视觉Transformer主干,揭示了数据受限与模型参数受限之间的关系,并提出了优化数据收集和计算资源的策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09036 2026-01-15 cs.CL cs.IR 77%

SpectraQuery: A Hybrid Retrieval-Augmented Conversational Assistant for Battery Science

SpectraQuery: 一种用于电池科学的混合检索增强型对话助手

Sreya Vangara, Jagjit Nanda, Yan-Kai Tzeng, Eric Darve

机构 * Mechanical Engineering, Stanford University(斯坦福大学机械工程系) Applied Energy Division, SLAC National Accelerator Laboratory(SLAC国家加速器实验室应用能源部门) Institute of Computational and Mathematical Engineering, Stanford University(斯坦福大学计算与数学工程研究所)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SpectraQuery通过结合结构化数据库和文献语料库,为电池科学提供检索增强型对话助手,有效整合数值证据与机理解释,提升科学推理效率。

Comments 11 pages, 8 figures, appendix included

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09280 2026-01-15 cs.CL cs.AI 73%

ReGraM: Region-First Knowledge Graph Reasoning for Medical Question Answering

ReGraM:面向医疗问答的区域优先知识图谱推理

Chaerin Lee, Sohee Park, Hyunsik Na, Daseon Choi

机构 * Department of Software, Soongsil University(软件系,顺斯大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 ReGraM通过区域优先的知识图谱推理框架,提升医疗问答的准确性和一致性。

Comments 18 pages, 2 figures. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09470 2026-01-15 physics.ed-ph cs.AI 70%

Personalized Multimodal Feedback Using Multiple External Representations: Strategy Profiles and Learning in High School Physics

基于多种外部表征的个性化反馈:策略配置与高中物理学习中的学习

Natalia Revenga-Lozano, Karina E. Avila, Steffen Steinert, Matthias Schweinberger, Clara E. Gómez-Pérez, Jochen Kuhn, Stefan Küchemann

机构 * Chair of Physics Education, Faculty of Physics, Ludwig-Maximilians-Universität München (LMU Munich)(物理教育系主任,物理学院,慕尼黑路易斯-马克西姆利安大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究了多种外部表征与个性化反馈在高中物理学习中的整合效果,发现详细多表征反馈对学习成绩有积极影响,且学习者根据表征能力选择不同反馈策略。

Comments Keywords: Adaptive Feedback, Multimodal Learning, Multiple External Representations, Physics Education, Science Education, Representational Competences, Intelligent Tutoring Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08988 2026-01-15 cs.AI 70%

ART: Action-based Reasoning Task Benchmarking for Medical AI Agents

ART:面向医疗AI代理的基于动作的推理任务基准测试

Ananya Mantravadi, Shivali Dalmia, Abhishek Mukherji

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 ART通过挖掘真实EHR数据创建挑战性任务,评估医疗AI代理在阈值评估、时间聚合和条件逻辑方面的表现,揭示其推理弱点,推动更可靠的临床决策支持系统发展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07229 2026-01-15 cs.HC cs.AI cs.IR 70%

DiSCo: Making Absence Visible in Intelligent Summarization Interfaces

DiSCo:在智能摘要界面中使缺失内容可见

Eran Fainman, Hagit Ben Shoshan, Adir Solomon, Osnat Mokryn

机构 * University of Haifa(海法大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 DiSCo通过对比领域期望揭示摘要中的缺失内容,提升智能摘要界面的透明度和决策支持能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.16198 2026-01-15 cs.CL 70%

Towards Efficient Patient Recruitment for Clinical Trials: Application of a Prompt-Based Learning Model

迈向临床试验高效患者招募:基于提示学习模型的应用

Mojdeh Rahmanian, Seyed Mostafa Fakhrahmad, Seyedeh Zahra Mousavi

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究提出基于提示学习模型的高效患者招募方法,利用SNOMED CT本体实现医疗文本的精准分类,提升了临床试验招募效率。

Journal ref 2025;31(4):367-377

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08901 2026-01-15 cs.IR cs.AI cs.CL cs.LG 67%

Navigating Ideation Space: Decomposed Conceptual Representations for Positioning Scientific Ideas

导航构想空间:分解的概念表示用于定位科学思想

Yuexi Shen, Minqian Liu, Dawei Zhou, Lifu Huang

机构 * Virginia Tech(弗吉尼亚理工大学) University of California, Davis(加州大学戴维斯分校) University of California, Santa Barbara(加州大学圣巴巴拉分校)

专题命中 领域大模型 :LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出构想空间,通过分解科学知识的三个维度,实现科学思想的定位与评估,提升文献检索和创新性判断的效率与准确性。

Comments 21 pages, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏