arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-17 至 2025-12-17 共收录 18 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 18 篇

2505.16097 2025-12-17 cs.AI 89%

Developing Large Language Models for Clinical Research Using One Million Clinical Trials

利用一百万项临床试验开发大型语言模型用于临床研究

Zifeng Wang, Jiacheng Lin, Qiao Jin, Junyi Gao, Jathurshan Pradeepkumar, Pengcheng Jiang, Zhiyong Lu, Jimeng Sun

机构 * Usher Institute, Edinburgh Medical School, University of Edinburgh(埃丁堡大学埃丁堡医学院usher研究所) Health Data Research UK(英国健康数据研究) School of Computing and Data Science, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校计算与数据科学学院) Division of Intramural Research, National Library of Medicine, National Institutes of Health(国家卫生研究院国家医学图书馆内部研究部) Keiji AI

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出TrialPanorama,通过整合一百万项临床试验数据,训练出在八个关键临床研究任务中表现优于通用大语言模型的8B LLM。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10852 2025-12-17 cond-mat.mtrl-sci cs.CL cs.DB 89%

MatTools: Benchmarking Large Language Models for Materials Science Tools

MatTools: 评估大型语言模型在材料科学工具中的基准测试

Siyu Liu, Bo Hu, Beilin Ye, Jiamin Xu, David J. Srolovitz, Tongqi Wen

机构 * Center for Structural Materials, Department of Mechanical Engineering, The University of Hong Kong(香港大学机械工程系结构材料中心) Materials Innovation Institute for Life Sciences(生命科学创新研究所)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 MatTools通过代码生成和执行评估LLM在材料科学工具中的能力,揭示通用模型优于专业模型,AI理解AI,及简单方法更优的结论。

Comments 27 pages, 23 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09554 2025-12-17 cs.IR 89%

Mixture-of-RAG: Integrating Text and Tables with Large Language Models

混合RAG:利用大语言模型整合文本和表格

Chi Zhang, Qiyang Chen, Mengqi Zhang

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract)

AI总结 MixRAG通过三阶段框架整合文本和表格,提升异构文档检索性能,实现混合模态文档接地的最新成果。

Comments Accepted to SIGKDD 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.21112 2025-12-17 cs.AI cs.CE cs.CL cs.CY cs.IR 88%

Optimizing Large Language Models for ESG Activity Detection in Financial Texts

优化大型语言模型以检测金融文本中的ESG活动

Mattia Birti, Andrea Maurino, Francesco Osborne

机构 * Department of Informatics, Systems and Communication, University of Milano-Bicocca(信息学、系统与通信系,米兰-比科卡大学) University of Milano-Bicocca(米兰-比科卡大学) The Open University(开放大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出通过微调优化大型语言模型,提升金融文本中ESG活动检测的准确性。

Comments Published in the Proceedings of the ACM International Conference on AI in Finance (ICAIF). ACM version

Journal ref Proceedings of the ACM International Conference on AI in Finance (ICAIF), 2024, ACM

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02503 2025-12-17 cs.CV 87%

Adapting General-Purpose Foundation Models for X-ray Ptychography in Low-Data Regimes

为低数据情形下的X射线衍射成像适应通用基础模型

Robinson Umeike, Neil Getty, Yin Xiangyu, Yi Jiang

机构 * The University of Alabama(阿拉巴马大学) Argonne National Laboratory(阿贡国家实验室)

专题命中 领域大模型 :foundation model(title,abstract);language model(abstract);SFT(abstract);prompting(abstract)

AI总结 本文提出PtychoBench基准,通过比较SFT和ICL策略,在低数据环境下优化X射线衍射成像任务的模型适应性,发现任务模态决定最佳专门化路径。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14657 2025-12-17 cs.SD 82%

Adapting Speech Language Model to Singing Voice Synthesis

将语音语言模型适应于歌唱语音合成

Yiwen Zhao, Jiatong Shi, Jinchuan Tian, Yuxun Tang, Jiarui Hai, Jionghao Han, Shinji Watanabe

机构 * Carnegie Mellon University(卡内基梅隆大学) Renmin University of China(中国人民大学) John Hopkins University(约翰霍普金斯大学)

专题命中 领域大模型 :language model(title,abstract);SLM(abstract)

AI总结 本文提出将预训练语音语言模型适应于歌唱语音合成,通过多流语言模型预测、条件流匹配生成梅尔频谱图及梅尔到波形 vocoder 实现,验证了模型在歌唱语音合成任务中的有效性。

Comments Accepted by NeurIPS 2025 workshop AI for Music

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07584 2025-12-17 cs.LG 79%

MIRA: Medical Time Series Foundation Model for Real-World Health Data

MIRA:面向真实世界健康数据的医学时间序列基础模型

Hao Li, Bowen Deng, Chang Xu, Zhiyuan Feng, Viktor Schlegel, Yu-Hao Huang, Yizheng Sun, Jingyuan Sun, Kailai Yang, Yiyao Yu, Jiang Bian

机构 * Microsoft Research(微软研究院) University of Manchester(曼彻斯特大学) Peking University(北京大学) Tsinghua University(清华大学) Nanjing University(南京大学) Imperial Global Singapore, Imperial College London(帝国理工学院伦敦分校)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.LG

AI总结 MIRA是一种专为医学时间序列预测设计的基础模型,通过连续时间旋转位置编码、频率特定混合专家层和连续动态外推块,实现对不规则时间间隔、异质采样率和缺失值的高效处理,提升医疗时间序列数据的预测精度和跨机构迁移能力。

Comments NeurIPS 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18054 2025-12-17 cs.IR cs.AI cs.LG 79%

A Knowledge Graph-based Retrieval-Augmented Generation Framework for Algorithm Selection in the Facility Layout Problem

基于知识图谱的检索增强生成框架用于设施布局问题中的算法选择

Nikhil N S, Bilal Muhammed, Soban Babu Beemaraj, Amol Dilip Joshi

机构 * Indian Institute of Science(印度科学学院) TCS Research and Innovation(塔塔咨询公司研究与创新)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于知识图谱的检索增强生成框架,用于自动推荐设施布局问题中的算法选择,通过多维检索机制和大语言模型实现数据驱动的算法推荐。

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20047 2025-12-17 cs.CV eess.IV 78%

Med3DVLM: An Efficient Vision-Language Model for 3D Medical Image Analysis

Med3DVLM: 一种高效的视觉-语言模型用于3D医学图像分析

Yu Xin, Gorkem Can Ates, Kuang Gong, Wei Shao

机构 * University of Florida(佛罗里达大学)

专题命中 领域大模型 :language model(title,abstract)

AI总结 Med3DVLM通过三个创新提出,实现了高效的3D医学图像分析,显著提升了图像-文本检索、报告生成和视觉问答的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01115 2025-12-17 cs.AI cs.MA econ.TH physics.soc-ph 77%

Exploring Network-Knowledge Graph Duality: A Case Study in Agentic Supply Chain Risk Analysis

探索网络-知识图谱二元性:代理供应链风险分析的案例研究

Evan Heus, Rick Bookstaber, Dhruv Sharma

机构 * University of California, Berkeley(加州大学伯克利分校)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于LLM的代理框架,利用网络与知识图谱的二元性,通过图遍历和上下文壳技术实现供应链风险分析的实时可解释生成。

Comments Accepted to the 2nd Workshop on LLMs and Generative AI in Finance: International Conference on AI in Finance(ICAIF) 2025;7 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14429 2025-12-17 cs.AI cs.SE 70%

Seismology modeling agent: A smart assistant for geophysical researchers

地震建模代理:为地球物理研究者设计的智能助手

Yukun Ren, Siwei Yu, Kai Chen, Jianwei Ma

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于LLMs的SPECFEM智能交互工作流程,通过MCP协议降低地震模拟门槛,提升地球物理研究的自动化与可重复性。

Comments 26 pages, 15 figures. Code available at https://github.com/RenYukun1563/specfem-mcp

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13747 2025-12-17 cs.CV cs.AI 70%

Why Text Prevails: Vision May Undermine Multimodal Medical Decision Making

为何文本占上风:视觉可能损害多模态医疗决策制定

Siyuan Dai, Lunxiao Li, Kun Zhao, Eardi Lila, Paul K. Crane, Heng Huang, Dongkuan Xu, Haoteng Tang, Liang Zhan

机构 * University of Texas Rio Grande Valley(德克萨斯大学里奥格兰德谷大学) University of Pittsburgh(匹兹堡大学) NC State University(北卡罗来纳州立大学) University of Washington(华盛顿大学) University of Maryland(马里兰大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究发现文本推理在医疗多模态决策中优于多模态输入,提出三种策略以提升多模态医疗决策能力。

Comments Accepted by ICDM 2025 the Workshop on Synergy of AI and Multimodal Biomedical Data Mining

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14321 2025-12-17 cs.MA 67%

Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT Consultations

多智能体医疗决策共识矩阵系统:一种用于肿瘤多学科会诊的智能协作框架

Xudong Han, Xianglun Gao, Xiaoyi Qu, Zhenyu Yu

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 多智能体医疗决策共识矩阵系统通过智能体协作提升肿瘤多学科会诊的决策质量和效率。

Comments 14 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09605 2025-12-17 eess.IV cs.AI cs.LG q-bio.QM 62%

TomoGraphView: 3D Medical Image Classification with Omnidirectional Slice Representations and Graph Neural Networks

TomoGraphView: 基于全方位切片表示和图神经网络的3D医学图像分类

Johannes Kiechle, Stefan M. Fischer, Daniel M. Lang, Cosmin I. Bercea, Matthew J. Nyflot, Lina Felsner, Julia A. Schnabel, Jan C. Peeken

机构 * School of Computation, Information and Technology, Technical University of Munich(计算信息技术学院,慕尼黑技术大学) Department of Radiation Oncology, TUM School of Medicine, TUM University Hospital rechts der Isar, Technical University of Munich(放射肿瘤学系,TUM医学院,TUM大学医院rechts der Isar,慕尼黑技术大学) Institute of Machine Learning in Biomedical Imaging, Helmholtz Munich(生物医学影像机器学习研究所,海德堡慕尼黑) Institute of Radiation Medicine, Helmholtz Munich(放射医学研究所,海德堡慕尼黑) School of Biomedical Engineering and Imaging Sciences, King's College London(生物医学工程与成像科学学院,伦敦国王学院) Department of Radiation Oncology, University of Washington(放射肿瘤学系,华盛顿大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML))

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 TomoGraphView通过整合全方位体积切片与图神经网络,解决3D医学图像分类中切片方向限制和空间一致性问题。

Comments Preprint submitted to Medical Image Analysis (MedIA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18638 2025-12-17 cs.CV cs.AI 57%

Learning neuroimaging models from health system-scale data

从健康系统级数据中学习神经影像模型

Yiwei Lyu, Samir Harake, Asadur Chowdury, Soumyanil Banerjee, Rachel Gologorsky, Shixuan Liu, Anna-Katharina Meissner, Akshay Rao, Chenhui Zhao, Akhil Kondepudi, Cheng Jiang, Xinhai Hou, Rushikesh S. Joshi, Volker Neuschmelting, Ashok Srinivasan, Dawn Kleindorfer, Brian Athey, Vikas Gulani, Aditya Pandey, Honglak Lee, Todd Hollon

机构 * University of Michigan Computer Science and Engineering(密歇根大学计算机科学与工程系) University of Michigan Neurosugery(密歇根大学神经外科) University of Cologne Neurosugery(科隆大学神经外科) University of Michigan Radiology(密歇根大学放射学) University of Michigan Neurology(密歇根大学神经病学) University of Michigan Computational Medicine and Bioinformatics(密歇根大学计算医学与生物信息学)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 Prima是一款基于健康系统级数据训练的视觉语言模型,用于神经影像学,实现了高准确率的诊断和临床决策支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03189 2025-12-17 cs.CV 50%

Dynamic Prompt Generation for Interactive 3D Medical Image Segmentation Training

动态提示生成用于交互式3D医学图像分割训练

Tidiane Camaret Ndir, Alexander Pfefferle, Robin Tibor Schirrmeister

机构 * Medical Physics, Department of Diagnostic and Interventional Radiology, Medical Center—University of Freiburg, Faculty of Medicine, University of Freiburg(医学物理,诊断与介入放射科,弗赖堡大学医学中心,医学院,弗赖堡大学) University of Freiburg(弗赖堡大学) ELLIS Institute Tübingen(图宾根ELLIS研究所)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本文提出一种动态提示生成方法,结合内容感知自适应裁剪优化图像编码器,提升交互式3D医学图像分割的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14026 2025-12-17 cs.CV 50%

Unleashing the Power of Image-Tabular Self-Supervised Learning via Breaking Cross-Tabular Barriers

通过打破跨表格障碍释放图像-表格自监督学习的潜力

Yibing Fu, Yunpeng Zhao, Zhitao Zeng, Cheng Chen, Yueming Jin

专题命中 领域大模型 :pretraining(abstract)

AI总结 本文提出CITab框架,通过跨表格学习提升医学图像与表格数据的多模态自监督学习效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02677 2025-12-17 eess.IV cs.CV 50%

Multimodal Deep Learning for Stroke Prediction and Detection using Retinal Imaging and Clinical Data

基于视网膜成像和临床数据的多模态深度学习用于中风预测与检测

Saeed Shurrab, Aadim Nepal, Terrence J. Lee-St. John, Nicola G. Ghazi, Bartlomiej Piechowski-Jozwiak, Farah E. Shamout

机构 * Division of Engineering, New York University Abu Dhabi(纽约大学阿布扎赫尔分校工程系) Institute for Healthier Living Abu Dhabi(阿布扎赫尔健康生活研究所) Eye Institute at Cleveland Clinic Abu Dhabi(阿布扎赫尔克利夫兰医学中心眼科研究所) Canberra Hospital(堪培拉医院)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本研究提出一种多模态深度学习方法,利用视网膜成像和临床数据预测中风风险,实验结果显示在中风检测和风险预测方面优于传统方法。

详情

展开后加载摘要…

URL PDF HTML 收藏