arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12720 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 12720 篇

2510.08891 2026-07-12 cs.ET cs.AI cs.HC 70%

Designing and Evaluating an AI-enhanced Immersive Multidisciplinary Simulation (AIMS) for Interprofessional Education

设计和评估一种增强人工智能的沉浸式跨学科模拟(AIMS)用于多专业教育

Ruijie Wang, Jie Lu, Bo Pei, Evonne Jones, Jamey Brinson, Timothy Brown

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 AIMS通过整合AI和虚拟环境,为多专业教育提供沉浸式模拟,提升学生协作能力和临床推理能力。

Comments 15 pages. Immersive Learning Research-Academic (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.08009 2026-07-10 cs.CL cs.CY 新提交 70%

From Execution to Education: A Bloom-Aligned Framework for Measuring Educational Control in LLMs

从执行到教育:一个用于衡量大语言模型中教育控制的布鲁姆对齐框架

Yi Zhang, Julia Rayz

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究提出用于衡量大语言模型中教育控制的布鲁姆对齐框架,应用于编程任务。通过比较一对匹配模型在不同干预设置下的表现,揭示模型提高认知需求易、降低难的不对称性,并用语义 - δ聚类等方法刻画结果,表明强执行性能不意味着有教育控制能力。

Comments 24 pages, 20 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.07761 2026-07-10 cs.AI 新提交 70%

Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning

对齐临床需求与人工智能能力:医学推理大语言模型综述

Qi Peng, Jiatong Li, Sirui Huang, Yiyang Jiang, Kaisong Gong, Ronger Ding, Shijie Ye, Changmeng Zheng, Yi Cai, Xiaobo Yang, Jin Huang, Xiao-Yong Wei, Qing Li

机构 * The Hong Kong Polytechnic University(香港理工大学) Hong Kong University(香港大学) South China University of Technology(华南理工大学) University of Toronto(多伦多大学) Peking Union Medical College Hospital(北京地坛医院) West China Hospital, Sichuan University(四川大学华西医院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 综述医学大语言模型在医疗领域的进展,提出双视角方法,连接临床实践与计算方法,建立五级能力方案并关联推理模式与医学任务,引入基准数据集并报告模型结果,讨论进展与挑战及未来方向。

Comments Accepted by Machine Intelligence Research

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.05050 2026-07-10 cond-mat.mtrl-sci cs.AI physics.chem-ph 交叉投稿 70%

Autonomous heterogeneous catalyst discovery with a self-evolving multi-agent digital twin

自主异质催化剂发现:一种自进化多智能体数字孪生系统

Zhilong Song, Zongmin Zhang, Lixue Cheng

机构 * Department of Chemistry, Hong Kong University of Science and Technology(香港科技大学化学系) IAS Center for AI for Scientific Discoveries, Hong Kong University of Science and Technology(香港科技大学人工智能科学发现中心) Department of Computer Science and Engineering, Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系) Department of Chemical and Biological Engineering, Hong Kong University of Science and Technology(香港科技大学化学与生物工程系)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 提出CatDT(催化数字孪生),一种自进化多智能体系统,通过集成八种专业智能体和27种科学工具,在单个GPU上5-30分钟内自动构建工作催化剂数字孪生,实现从体相晶体和自然语言反应描述到稳定晶面预测、反应路径枚举、过渡态定位和动力学计算的全流程,在七个气固相基准上预测与实验偏差在0.5-2倍内,并独立发现丙烷脱氢非贵金属候选催化剂。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04920 2026-07-10 cs.AI 版本更新 70%

Conversational AI for Rapid Scientific Prototyping: A Case Study on ESA's ELOPE Competition

用于快速科学原型设计的对话式人工智能:以欧空局的ELOPE竞赛为例

Nils Einecke

机构 * Honda Research Institute Europe GmbH(本田欧洲研究机构)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 以欧空局ELOPE竞赛为例,探讨用ChatGPT进行快速科学原型设计。介绍其在竞赛中的表现,包括提供代码、推理等,虽有局限但凸显人机协作潜力,分析优缺点,提出将大语言模型集成到科学工作流程可增强快速原型设计。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24978 2026-07-09 cs.AI cond-mat.quant-gas quant-ph 70%

Agentic Exploration of Physics Models

物理模型的智能体探索

Maximilian Nägele, Florian Marquardt

机构 * Max Planck Institute for the Science of Light(马克斯·普朗克光科学研究所)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出 SciExplorer 智能体,利用大语言模型工具使用能力,无需领域特定蓝图即可探索未知物理系统,通过实验和观测恢复运动方程和哈密顿量。

Journal ref Maximilian Nägele and Florian Marquardt, 2026 Phys. Rev. X 16, 031002

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25605 2026-07-09 cs.IR cs.AI cs.DB 版本更新 70%

Health System Scale Semantic Search Across Unstructured Clinical Notes

跨无结构临床笔记的健康系统规模语义搜索

Faith Wavinya Mutinda, Spandana Makeneni, Anna Lin, Shivaji Dutta, Irit R. Rasooly, Patrick Dibussolo, Shivani Kamath Belman, Hessam Shahriari, Kevin Murphy, Alex B. Ruan, Barbara H. Chaiyachati, Sanjay Chainani, Robert W. Grundmeier, Scott M. Haag, Jeffrey M. Miller, Heather M. Griffis, Ian M. Campbell

机构 * Department of Biomedical and Health Informatics, Children’s Hospital of Philadelphia(儿童医院哲学学院生物医学与健康信息学系) Google Cloud(谷歌云) Department of Pediatrics, University of Pennsylvania(宾夕法尼亚大学儿科系) Division of Neonatology, Children’s Hospital of Philadelphia(儿童医院哲学学院新生儿科) Division of Human Genetics, Children’s Hospital of Philadelphia(儿童医院哲学学院人类遗传学部)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.AI

AI总结 本文提出了一种在大规模健康系统中实现语义搜索的解决方案,通过优化嵌入模型和分块策略,实现了亚秒级查询延迟和高准确率的临床问答任务,同时展示了在临床实用性评估中减少任务完成时间的效果。

Comments For associated code, see https://github.com/Ian-Campbell-Lab/clinical-semantic-search

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.05685 2026-07-08 cs.HC cs.AI 新提交 70%

Depression Symptoms and Relational Patterns in 187k ChatGPT Histories

18.7万条ChatGPT对话记录中的抑郁症状与关系模式

Neil K. R. Sehgal, Dunigan Folk, Lyle Ungar, Sharath Chandra Guntuku

机构 * University of Pennsylvania(宾夕法尼亚大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究有抑郁症状者如何使用ChatGPT,通过分析766名参与者的18.7万条对话,比较不同症状程度人群的使用模式、语言特点等,发现基于语言预测筛查效果不佳,认为这些记录可作大语言模型成非正式支持设施的证据。

Journal ref CSCW Companion '26: Companion Publication of the 2026 Conference on Computer-Supported Cooperative Work and Social Computing

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.13566 2026-07-08 cs.AI 新提交 70%

A Three-Layer Framework for AI in Scientific Discovery

人工智能在科学发现中的三层框架

Guojun Liao

机构 * Department of Mathematics, University of Texas at Arlington(德克萨斯大学阿灵顿分校数学系)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出AI在科学发现中的三层框架,核心创新是第二层:通过定性推理进行模型形成,识别框架结构不足并寻找缺失概念,通过三个案例说明其重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.17437 2026-07-08 cs.SE cs.LG 版本更新 70%

A semantic mutation metric for metamorphic relation adequacy in scientific computing programs

一种用于科学计算程序元突变关系充分性的语义突变度量

Meng Li, Xiaohua Yang, Jie Liu, Shiyu Yan

机构 * School of Computing, University of South China(南华大学计算机学院) Hunan Engineering Research Center of Software Evaluation and Testing for Intellectual Equipment(湖南软件测评与智能设备工程研究中心) CNNC Key Laboratory on High Trusted Computing(中核集团高可信计算重点实验室)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.LG

AI总结 本文提出了一种基于领域语义操作符的语义突变度量(SMS),旨在解决传统突变度量在科学计算中忽略领域语义的问题,通过引入五个领域语义操作符,提高了对元突变关系充分性的评估能力。

Comments 93 pages in elsarticle review mode (12pt double-spaced, ~28-35 pp typeset), 3 figures. Replication package: https://doi.org/10.5281/zenodo.20250664. Corresponding author: Meng Li (mlemon@usc.edu.cn)

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00738 2026-07-07 cs.DL cs.AI 新提交 70%

Phantom References: Hallucinated Citations That Survive Peer Review at Top-Tier Conferences

幻影引用:在顶级会议同行评审中存活的幻觉引用

Mark Russinovich, Ram Shankar Siva Kumar, Ahmed Salem

机构 * Microsoft(微软)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出RefChecker工具,通过多源验证和网络搜索复检,发现顶级会议论文中存在少量但可观的幻觉引用(如不存在的文献或作者列表严重不匹配),且同行评审未能完全阻止其进入档案记录。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29580 2026-07-07 cs.CL 版本更新 70%

MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibar

MAM-AI:面向桑给巴尔护士和助产士的设备端医疗检索增强生成系统

Yi Ren

机构 * École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院洛桑分校)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 针对撒哈拉以南非洲地区护士助产士难以获取权威指南的问题,提出完全运行于安卓设备上的医疗问答系统MAM-AI,采用300M嵌入模型检索87份指南文档,并用4B int4生成器离线生成带引用的回答,评估发现小生成器在安全性和有用性间存在权衡,通过优化提示词降低回避率。

Comments arXiv admin comment: This version has been removed by arXiv administrators as the submitter did not have the rights to agree to the license at the time of submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29467 2026-07-07 cs.CL cs.IR 版本更新 70%

mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Health

mamabench 和 mamaretrieval:评估孕产妇、新生儿和生殖健康领域医学检索增强生成的基准

Yi Ren

机构 * École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院洛桑校区)

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 针对助产士咨询的孕产妇、新生儿和生殖健康问题,构建了包含25,949个问题的QA基准mamabench和基于3,185个查询的段落级相关性基准mamaretrieval,采用分级相关标注和标签质量审计。

Comments arXiv admin comment: This version has been removed by arXiv administrators as the submitter did not have the rights to agree to the license at the time of submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11730 2026-07-07 cs.CV cs.HC cs.LG 版本更新 70%

Multimodal Ambivalence/Hesitancy Recognition in Videos for Personalized Digital Health Interventions

视频中矛盾/犹豫识别用于个性化数字健康干预

Manuela González-González, Soufiane Belharbi, Muhammad Osama Zeeshan, Masoumeh Sharafi, Muhammad Haseeb Aslam, Lorenzo Sia, Nicolas Richet, Marco Pedersoli, Alessandro Lameiras Koerich, Simon L Bacon, Eric Granger

机构 * LIVIA, Dept. of Systems Engineering, ETS Montreal, Canada(ETS蒙特利尔大学系统工程系LIVIA实验室) LIVIA, Dept. of Software and IT Engineering, ETS Montreal, Canada(ETS蒙特利尔大学软件与信息工程系LIVIA实验室) Dept. of Health, Kinesiology, & Applied Physiology, Concordia University, Montreal, Canada(康科迪亚大学健康、运动科学与应用生理学系) Montreal Behavioural Medicine Centre, CIUSSS Nord-de-l’Ile-de-Montréal, Canada(蒙特利尔行为医学中心,蒙特利尔北岛卫生与社会服务局)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文研究了通过深度学习模型在视频中进行矛盾/犹豫识别,以提升数字健康干预的个性化和成本效益,实验基于新的BAH视频数据集,发现需改进多模态模型以准确识别矛盾/犹豫。

Comments 11 pages, 4 figures, ACII 2026. arXiv admin note: substantial text overlap with arXiv:2505.19328

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07285 2026-07-07 cs.CL cs.CY 版本更新 70%

Why teaching resists automation in an AI-inundated era: Human judgment, non-modular work, and the limits of delegation

为何在人工智能泛滥的时代教学抵抗自动化:人类判断、非模块化工作与委托的局限

Songhee Han

机构 * Florida State University(佛罗里达州立大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文探讨了在人工智能普及背景下,教学工作难以自动化的原因,指出教学本质上具有解释性、关联性和专业判断,无法被完全自动化或委托给技术。

Comments Revised version; accepted for publication in TechTrends

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08029 2026-07-07 cs.LG cs.CV 版本更新 70%

CLARITY: Medical World Model for Guiding Treatment Decisions by Modeling Context-Aware Disease Trajectories in Latent Space

CLARITY:用于通过在潜在空间中建模情境感知疾病轨迹来指导治疗决策的医学世界模型

Tianxingjian Ding, Yuanhao Zou, Chen Chen, Mubarak Shah, Yu Tian

机构 * Institute of Artificial Intelligence, University of Central Florida(中佛罗里达大学人工智能研究所)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 CLARITY通过在潜在空间中建模时间间隔和患者特定数据,预测疾病演变轨迹,生成个性化治疗方案,并在医学领域实现最先进的治疗规划性能。

Comments Accepted to ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.22869 2026-07-07 cs.CL 70%

JBE-QA: Japanese Bar Exam QA Dataset for Assessing Legal Domain Knowledge

JBE-QA:日本司法考试问答数据集用于评估法律领域知识

Zhihan Cao, Fumihito Nishino, Hiroaki Yamada, Nguyen Ha Thanh, Yusuke Miyao, Ken Satoh

机构 * Institute of Science Tokyo, Japan(东京科学研究所,日本) ROIS-DS, Center for Juris-informatics, Japan(司法信息中心,日本) National Institute of Informatics, Japan(日本信息机构) University of Tokyo, Japan(东京大学,日本)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出JBE-QA数据集,用于评估大语言模型的法律知识,涵盖民法、刑法和宪法,通过分解问题为独立判断并添加上下文字段,共包含3464个平衡标注问题,评估26种LLM表现,显示具备推理能力的专有模型表现最佳。

Comments Three tables and one figure

Journal ref In Proc. of the 15th Language Resources and Evaluation Conference (LREC 2026) (pp. 5317-5327)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.13191 2026-07-03 physics.comp-ph cond-mat.mtrl-sci cs.AI 版本更新 70%

From Experiments to Expertise: Scientific Knowledge Consolidation for AI-Driven Computational Physics

从实验到专长:面向AI驱动计算物理的科学知识巩固

Haonan Huang

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出QMatSuite平台,通过记录、检索和反思机制巩固计算材料科学中的知识,减少推理开销并提高模拟准确性。

Comments v2: camera-ready version, accepted at the ICML 2026 Workshop on AI for Physics (AI4Physics@ICML 2026). 20 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00856 2026-07-02 q-fin.CP cs.LG 新提交 70%

Shapley in Context: Explaining Financial Language with Domain Expertise

语境中的Shapley:利用领域专业知识解释金融语言

Dangxing Chen, Pengzhan Guo

机构 * Zu Chongzhi Center, Duke Kunshan University(杜克昆山大学祖冲之中心) Digital Innovation Research Center, Duke Kunshan University(杜克昆山大学数字创新研究中心)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究Shapley值在大型语言模型金融文本解释中的适用性,通过理论分析和实验验证其与金融领域知识的一致性。

Comments European Journal of Finance

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00828 2026-07-02 cs.DB cs.AI 新提交 70%

Exploring the Semantic Gap in Agentic Data Systems: A Formative Study of Operationalization Failures in Analytical Workflows

探索智能体数据系统中的语义鸿沟:分析工作流中操作化失败的形成性研究

Jalal Mahmud, Eser Kandogan

机构 * Megagon Labs(Megagon实验室)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究通过跨领域分析236个分析意图,识别出153种反复出现的操作化失败,归纳为五类语义鸿沟,揭示用户分析概念与系统可用信息之间的差距,并提出未来智能体数据系统需要更丰富的语义表示。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00479 2026-07-02 cs.LG stat.ML 新提交 70%

Ghost in the Kernel: In-Context Learning with Efficient Transformers via Domain Generalization

内核中的幽灵:通过领域泛化实现高效Transformer的上下文学习

Peilin Liu, Ding-Xuan Zhou

机构 * School of Mathematics and Statistics, University of Sydney(悉尼大学数学与统计学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文从领域泛化角度研究线性Transformer的近似与泛化能力,证明其通过上下文分布到响应函数的映射实现上下文学习,并基于理论提出激活函数与损失函数设计新视角。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00147 2026-07-02 cs.AI 新提交 70%

RareDxR1: Autonomous Medical Reasoning for Rare Disease Diagnosis Beyond Human Annotation

RareDxR1:超越人类标注的罕见病诊断自主医学推理

Deyang Jiang, Haoran Wu, Ziyi Wang, Yiming Rong, Yunlong Zhao, Ye Jin, Bo Xu

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出端到端推理大模型RareDxR1,通过知识内化与自主进化学习框架,结合反思增强推理采样和双级课程强化学习,直接从非结构化临床笔记诊断罕见病,无需依赖结构化表型或人类标注,在多个基准上达到最先进准确率。

Comments 7 pages, 3 figures. Accepted to IEEE International Conference on Multimedia and Expo (ICME) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30296 2026-07-02 cs.AI 版本更新 70%

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

ManimAgent: 用于视觉教育的自进化多模态智能体

Wenjia Jiang, Zongyuan Cai, Yuanhang Shao, Chenru Wang, Boyan Han, Zhixue Song, Keyu Chen, Shengwei An, Xu Yang, Zhou Yang

机构 * University of Alberta(阿尔伯塔大学) Southeast University(东南大学) Virginia Tech(弗吉尼亚理工学院) Xidian University(西安电子科技大学) Vivavia Inc(Vivavia公司)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出ManimAgent,通过双通道情节记忆库跨任务传递反思经验,无需权重更新或人工种子,在代码生成任务中提升通过率并减少反思轮次。

Comments Project page: https://manimagent.github.io/. Code: https://github.com/jwj1342/Paper2Manim

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.17406 2026-07-02 cs.AI 版本更新 70%

EvoMaster: A Foundational Evolving Agent Framework for Agentic Science at Scale

EvoMaster:一种用于大规模代理科学的基础进化代理框架

Xinyu Zhu, Yuzhu Cai, Zexi Liu, Cheng Wang, Fengyang Li, Wenkai Jin, Wanxu Liu, Zehao Bing, Bingyang Zheng, Jingyi Chai, Shuo Tang, Rui Ye, Yuwen Du, Xianghe Pang, Yaxin Du, Tingjia Miao, Yuzhi Zhang, Ruoxue Liao, Zhaohan Ding, Linfeng Zhang, Yanfeng Wang, Weinan E, Siheng Chen

机构 * School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) SciLand DP Technology(DP技术)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 EvoMaster通过持续自我进化机制,使代理能迭代优化假设并积累知识,实现跨学科的高效科学发现,其易用性与性能在多个基准测试中均表现优异。

Comments 44 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30906 2026-07-01 cs.AI 新提交 70%

Investigating Multi-Agent Deliberation in Law

法律中的多智能体审议研究

Cor Steging, Ludi van Leeuwen, Tadeusz Zbiegień

机构 * Bernoulli Institute of Mathematics, Computer Science and Artificial Intelligence, University of Groningen(格罗宁根大学伯努利数学、计算机科学与人工智能研究所) Department of Legal Theory, Jagiellonian University(雅盖隆大学法律理论系)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究基于大语言模型的多智能体审议方法在法律推理任务中的应用,提出两种受法庭程序和论证启发的框架,实验表明其与基线模型性能相当但答案差异显著,尤其擅长需要多角度批判性思维的问题。

Comments This manuscript has been accepted for presentation at the AIDA2J Workshop during the 21st International Conference of AI & Law in Singapore, June 8 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.14579 2026-07-01 cs.CV cs.LG 版本更新 70%

Medical Image Spatial Grounding with Semantic Sampling

基于语义采样的医学图像空间定位

Andrew Seohwan Yu, Mohsen Hariri, Kunio Nakamura, Mingrui Yang, Xiaojuan Li, Vipin Chaudhary

机构 * Case Western Reserve University(凯斯西储大学) Cleveland Clinic(克利夫兰诊所)

专题命中 领域大模型 :language model(abstract);prompting(abstract);分类 cs.LG

AI总结 研究视觉语言模型在医学图像三维空间定位中的挑战,提出MIS-Ground基准和MIS-SemSam优化方法,通过语义采样提升定位精度13.06%。

Comments 10 pages, 2 figures. Accepted at MICCAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30442 2026-06-30 cs.AI 70%

The FIL Hypothesis: Inductive Biases Help with Kernel Engineering

FIL假说:归纳偏置有助于核工程

Nikolai Rozanov, Subhabrata Dutta, Preslav Nakov, Iryna Gurevych

机构 * NLP Department, MBZUAI(自然语言处理部门,MBZUAI) Ubiquitous Knowledge Processing, TU Darmstadt(普遍知识处理,图腾斯达特大学) Department of Computing, Imperial College London(计算部门,伦敦帝国理工学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出反馈信息循环(FIL)假说,认为未来AI应用将面临长反馈循环的缩放极限,并引入基于归纳偏置的方法,在GPU编程任务中验证其优于纯数据驱动方法。

Comments 10 pages main, 17 pages abstract, pre-print

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.30096 2026-06-30 cs.CL cs.IT math.IT 70%

Information Dynamics of Language Communication

语言交际的信息动力学

Leonardo S. Goodall, Andrea I. Luppi, Pedro A. M. Mediano

机构 * Calleva Research Centre, University of Oxford, UK(牛津大学卡勒瓦研究中心) St John’s College, University of Cambridge, UK(剑桥大学圣约翰学院) Montréal Neurological Institute, McGill University, Canada(蒙特利尔神经科学研究所,麦吉尔大学,加拿大) Centre for Eudaimonia and Human Flourishing, University of Oxford, UK(幸福与人类繁荣中心,牛津大学,英国) Department of Computing, Imperial College London, UK(伦敦帝国理工学院计算机系)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 提出信息论框架量化语义信息在对话中的定向流动,通过语义转移熵和语义部分信息分解测量信息传递,在四个实验中验证其能检测认知僵化对话、说服者主导作用、心理治疗质量及议论文协同贡献。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29375 2026-06-30 cs.CL 70%

TriageRA-CCF: Source-Side Clinical Confidence and Coverage Signals for Adaptive Rank Budgeting in Medical LLMs

TriageRA-CCF:面向医疗大语言模型的源端临床置信度与覆盖度信号自适应秩预算分配

Shucan Ji, Yining Huang, Hongliang Guo

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 针对医疗问题难度差异大导致固定低秩适配效果不佳的问题,提出TriageRA-CCF方法,利用源端训练数据的置信度、临床覆盖度和反事实接近缺失信号指导自适应秩预算分配,在Qwen3-8B和Llama3.1-8B上平均准确率优于LoRA、DoRA和MoELoRA基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29024 2026-06-30 cs.CL 70%

Conversational Domain Adaptation of IndicTrans2 across 21 Indic Languages via Experience Replay and Model Soups

基于经验回放和模型汤的IndicTrans2跨21种印度语言对话域适应

Aditya Pratap Singh

专题命中 领域大模型 :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 通过经验回放和模型汤方法,将IndicTrans2-1B适应至21种印度语言的对话风格,在保持通用领域性能的同时,平均对话chrF提升6.2。

Comments 8 pages, 3 figures, 3 tables. Code: https://github.com/Aditya-PS-05/indictrans2-conversational Model: https://huggingface.co/adipras1407/indictrans2-en-indic-1B-conversational

详情

展开后加载摘要…

URL PDF HTML 收藏