arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-17 至 2025-12-17 共收录 169 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 18 篇

2509.18054 2025-12-17 cs.IR cs.AI cs.LG 79%

A Knowledge Graph-based Retrieval-Augmented Generation Framework for Algorithm Selection in the Facility Layout Problem

基于知识图谱的检索增强生成框架用于设施布局问题中的算法选择

Nikhil N S, Bilal Muhammed, Soban Babu Beemaraj, Amol Dilip Joshi

机构 * Indian Institute of Science(印度科学学院) TCS Research and Innovation(塔塔咨询公司研究与创新)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于知识图谱的检索增强生成框架,用于自动推荐设施布局问题中的算法选择,通过多维检索机制和大语言模型实现数据驱动的算法推荐。

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20047 2025-12-17 cs.CV eess.IV 78%

Med3DVLM: An Efficient Vision-Language Model for 3D Medical Image Analysis

Med3DVLM: 一种高效的视觉-语言模型用于3D医学图像分析

Yu Xin, Gorkem Can Ates, Kuang Gong, Wei Shao

机构 * University of Florida(佛罗里达大学)

专题命中 领域大模型 :language model(title,abstract)

AI总结 Med3DVLM通过三个创新提出,实现了高效的3D医学图像分析,显著提升了图像-文本检索、报告生成和视觉问答的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01115 2025-12-17 cs.AI cs.MA econ.TH physics.soc-ph 77%

Exploring Network-Knowledge Graph Duality: A Case Study in Agentic Supply Chain Risk Analysis

探索网络-知识图谱二元性:代理供应链风险分析的案例研究

Evan Heus, Rick Bookstaber, Dhruv Sharma

机构 * University of California, Berkeley(加州大学伯克利分校)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于LLM的代理框架,利用网络与知识图谱的二元性,通过图遍历和上下文壳技术实现供应链风险分析的实时可解释生成。

Comments Accepted to the 2nd Workshop on LLMs and Generative AI in Finance: International Conference on AI in Finance(ICAIF) 2025;7 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14429 2025-12-17 cs.AI cs.SE 70%

Seismology modeling agent: A smart assistant for geophysical researchers

地震建模代理:为地球物理研究者设计的智能助手

Yukun Ren, Siwei Yu, Kai Chen, Jianwei Ma

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出基于LLMs的SPECFEM智能交互工作流程,通过MCP协议降低地震模拟门槛,提升地球物理研究的自动化与可重复性。

Comments 26 pages, 15 figures. Code available at https://github.com/RenYukun1563/specfem-mcp

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13747 2025-12-17 cs.CV cs.AI 70%

Why Text Prevails: Vision May Undermine Multimodal Medical Decision Making

为何文本占上风:视觉可能损害多模态医疗决策制定

Siyuan Dai, Lunxiao Li, Kun Zhao, Eardi Lila, Paul K. Crane, Heng Huang, Dongkuan Xu, Haoteng Tang, Liang Zhan

机构 * University of Texas Rio Grande Valley(德克萨斯大学里奥格兰德谷大学) University of Pittsburgh(匹兹堡大学) NC State University(北卡罗来纳州立大学) University of Washington(华盛顿大学) University of Maryland(马里兰大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究发现文本推理在医疗多模态决策中优于多模态输入,提出三种策略以提升多模态医疗决策能力。

Comments Accepted by ICDM 2025 the Workshop on Synergy of AI and Multimodal Biomedical Data Mining

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14321 2025-12-17 cs.MA 67%

Multi-Agent Medical Decision Consensus Matrix System: An Intelligent Collaborative Framework for Oncology MDT Consultations

多智能体医疗决策共识矩阵系统:一种用于肿瘤多学科会诊的智能协作框架

Xudong Han, Xianglun Gao, Xiaoyi Qu, Zhenyu Yu

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 多智能体医疗决策共识矩阵系统通过智能体协作提升肿瘤多学科会诊的决策质量和效率。

Comments 14 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09605 2025-12-17 eess.IV cs.AI cs.LG q-bio.QM 62%

TomoGraphView: 3D Medical Image Classification with Omnidirectional Slice Representations and Graph Neural Networks

TomoGraphView: 基于全方位切片表示和图神经网络的3D医学图像分类

Johannes Kiechle, Stefan M. Fischer, Daniel M. Lang, Cosmin I. Bercea, Matthew J. Nyflot, Lina Felsner, Julia A. Schnabel, Jan C. Peeken

机构 * School of Computation, Information and Technology, Technical University of Munich(计算信息技术学院,慕尼黑技术大学) Department of Radiation Oncology, TUM School of Medicine, TUM University Hospital rechts der Isar, Technical University of Munich(放射肿瘤学系,TUM医学院,TUM大学医院rechts der Isar,慕尼黑技术大学) Institute of Machine Learning in Biomedical Imaging, Helmholtz Munich(生物医学影像机器学习研究所,海德堡慕尼黑) Institute of Radiation Medicine, Helmholtz Munich(放射医学研究所,海德堡慕尼黑) School of Biomedical Engineering and Imaging Sciences, King's College London(生物医学工程与成像科学学院,伦敦国王学院) Department of Radiation Oncology, University of Washington(放射肿瘤学系,华盛顿大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML))

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 TomoGraphView通过整合全方位体积切片与图神经网络,解决3D医学图像分类中切片方向限制和空间一致性问题。

Comments Preprint submitted to Medical Image Analysis (MedIA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18638 2025-12-17 cs.CV cs.AI 57%

Learning neuroimaging models from health system-scale data

从健康系统级数据中学习神经影像模型

Yiwei Lyu, Samir Harake, Asadur Chowdury, Soumyanil Banerjee, Rachel Gologorsky, Shixuan Liu, Anna-Katharina Meissner, Akshay Rao, Chenhui Zhao, Akhil Kondepudi, Cheng Jiang, Xinhai Hou, Rushikesh S. Joshi, Volker Neuschmelting, Ashok Srinivasan, Dawn Kleindorfer, Brian Athey, Vikas Gulani, Aditya Pandey, Honglak Lee, Todd Hollon

机构 * University of Michigan Computer Science and Engineering(密歇根大学计算机科学与工程系) University of Michigan Neurosugery(密歇根大学神经外科) University of Cologne Neurosugery(科隆大学神经外科) University of Michigan Radiology(密歇根大学放射学) University of Michigan Neurology(密歇根大学神经病学) University of Michigan Computational Medicine and Bioinformatics(密歇根大学计算医学与生物信息学)

专题命中 领域大模型 :language model(abstract);分类 cs.AI

AI总结 Prima是一款基于健康系统级数据训练的视觉语言模型,用于神经影像学,实现了高准确率的诊断和临床决策支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03189 2025-12-17 cs.CV 50%

Dynamic Prompt Generation for Interactive 3D Medical Image Segmentation Training

动态提示生成用于交互式3D医学图像分割训练

Tidiane Camaret Ndir, Alexander Pfefferle, Robin Tibor Schirrmeister

机构 * Medical Physics, Department of Diagnostic and Interventional Radiology, Medical Center—University of Freiburg, Faculty of Medicine, University of Freiburg(医学物理,诊断与介入放射科,弗赖堡大学医学中心,医学院,弗赖堡大学) University of Freiburg(弗赖堡大学) ELLIS Institute Tübingen(图宾根ELLIS研究所)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本文提出一种动态提示生成方法,结合内容感知自适应裁剪优化图像编码器,提升交互式3D医学图像分割的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14026 2025-12-17 cs.CV 50%

Unleashing the Power of Image-Tabular Self-Supervised Learning via Breaking Cross-Tabular Barriers

通过打破跨表格障碍释放图像-表格自监督学习的潜力

Yibing Fu, Yunpeng Zhao, Zhitao Zeng, Cheng Chen, Yueming Jin

专题命中 领域大模型 :pretraining(abstract)

AI总结 本文提出CITab框架,通过跨表格学习提升医学图像与表格数据的多模态自监督学习效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02677 2025-12-17 eess.IV cs.CV 50%

Multimodal Deep Learning for Stroke Prediction and Detection using Retinal Imaging and Clinical Data

基于视网膜成像和临床数据的多模态深度学习用于中风预测与检测

Saeed Shurrab, Aadim Nepal, Terrence J. Lee-St. John, Nicola G. Ghazi, Bartlomiej Piechowski-Jozwiak, Farah E. Shamout

机构 * Division of Engineering, New York University Abu Dhabi(纽约大学阿布扎赫尔分校工程系) Institute for Healthier Living Abu Dhabi(阿布扎赫尔健康生活研究所) Eye Institute at Cleveland Clinic Abu Dhabi(阿布扎赫尔克利夫兰医学中心眼科研究所) Canberra Hospital(堪培拉医院)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本研究提出一种多模态深度学习方法,利用视网膜成像和临床数据预测中风风险,实验结果显示在中风检测和风险预测方面优于传统方法。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 3 篇

2508.00969 2025-12-17 cs.LG cs.AI 62%

Masked Omics Modeling for Multimodal Representation Learning across Histopathology and Molecular Profiles

掩码组学建模用于病理学与分子特征的多模态表示学习

Lucas Robinet, Ahmad Berjaoui, Elizabeth Cohen-Jonathan Moyal

机构 * Oncopole(奥恩波尔) IRT Saint Exupéry(国际研究与技术圣埃克苏佩里) INSERM Cancer Research Center of Toulouse(里沃利癌症研究中心) Toulouse(图卢兹)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 MORPHEUS通过整合病理学图像和多组学数据,提出了一种多模态预训练策略,以提升癌症研究中的跨模态表示学习能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14092 2025-12-17 cs.CV cs.AI 57%

ProtoFlow: Interpretable and Robust Surgical Workflow Modeling with Learned Dynamic Scene Graph Prototypes

ProtoFlow: 基于学习动态场景图原型的可解释且鲁棒的手术流程建模

Felix Holm, Ghazal Ghazaei, Nassir Navab

机构 * Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML))

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI

AI总结 ProtoFlow通过学习动态场景图原型,实现了手术流程的可解释和鲁棒建模,提升了AI在手术培训和决策支持中的应用潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04960 2025-12-17 cs.CV cs.AI 57%

MIMIR: Masked Image Modeling for Mutual Information-based Adversarial Robustness

MIMIR:基于互信息的对抗鲁棒性图像建模

Xiaoyun Xu, Shujian Yu, Zhuoran Liu, Stjepan Picek

机构 * Radboud University Nijmegen(拉德博德大学尼姆维根分校) Vrije Universiteit Amsterdam(自由大学阿姆斯特丹) University of Zagreb Faculty of Electrical Engineering and Computing(Zagreb大学电子工程与计算学院)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI

AI总结 MIMIR通过互信息惩罚和自编码器的遮蔽图像建模,提升视觉变换器的对抗鲁棒性。

Comments Accepted by NDSS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 5 篇

2512.13441 2025-12-17 cs.CL q-bio.NC 88%

Large language models are not about natural language

大语言模型并非关于自然语言

Johan J. Bolhuis, Andrea Moro, Stephen Crain, Sandiway Fong

机构 * University of Cambridge, Department of Psychology(剑桥大学心理学系) University School for Advanced Studies(高级研究大学) Scuola Normale Superiore(规范大学) Macquarie University, Department of Linguistics(麦考瑞大学语言学系) University of Arizona, Department of Linguistics(亚利桑那大学语言学系)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 大语言模型并非基于自然语言构建,而是基于概率模型,而人类语言由内在计算系统生成层次化思维结构。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13914 2025-12-17 cs.SE cs.AI cs.HC 85%

Context Branching for LLM Conversations: A Version Control Approach to Exploratory Programming

上下文分支用于LLM对话:一种版本控制方法用于探索性编程

Bhargav Chickmagalur Nanjundappa, Spandan Maaheshwari

机构 * Northeastern University(东北大学)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 ContextBranch通过版本控制机制实现对话分支,解决多轮对话中上下文污染问题,提升探索性编程的响应质量和上下文意识。

Comments 11 pages, 4 figures, 2 tables, 1 code snippet, 4 algorithms

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14373 2025-12-17 cs.CV 71%

EcoScapes: LLM-Powered Advice for Crafting Sustainable Cities

EcoScapes: 由LLM驱动的城市可持续发展建议

Martin Röhn, Nora Gourmelon, Vincent Christlein

机构 * Technische Universität Nürnberg(图宾根技术大学) Friedrich-Alexander-Universität Erlangen-Nürnberg(埃尔兰根-纽伦堡弗里德里希-亚历山大大学)

专题命中 其他LLM :LLM(title)

AI总结 EcoScapes利用LLM、卫星影像和知识库,为小型城市提供气候适应策略建议,提升可持续发展能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15464 2025-12-17 cs.CV cs.LG 57%

SIGMMA: Hierarchical Graph-Based Multi-Scale Multi-modal Contrastive Alignment of Histopathology Image and Spatial Transcriptome

SIGMMA:基于层次图的多尺度多模态对比对齐:组织病理图像与空间转录组

Dabin Jeong, Amirhossein Vahidi, Ciro Ramírez-Suástegui, Marie Moullet, Kevin Ly, Mohammad Vali Sanian, Sebastian Birk, Yinshui Chang, Adam Boxall, Daniyal Jafree, Lloyd Steele, Vijaya Baskar MS, Muzlifah Haniffa, Mohammad Lotfollahi

机构 * Wellcome Sanger Institute(沃森桑格研究所) Cambridge Centre for AI in Medicine(剑桥人工智能医学中心) Institute of AI for Health(人工智能与健康研究所) Cambridge Stem Cell Institute(剑桥干细胞研究所)

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 SIGMMA通过多尺度多模态对比对齐,提升组织病理图像与空间转录组的跨模态对应表示,提高基因表达预测和跨模态检索性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04687 2025-12-17 cs.CV 50%

Guideline-Consistent Segmentation via Multi-Agent Refinement

通过多智能体细化实现指南一致的分割

Vanshika Vats, Ashwani Rathee, James Davis

专题命中 其他LLM :language model(abstract)

AI总结 本文提出一种多智能体无训练框架,通过Worker-Supervisor迭代细化架构实现指南一致的分割,有效应对复杂文本指南。

Comments To be published in The Fortieth AAAI Conference on Artificial Intelligence (AAAI 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏