Comments15 pages. Accepted at AIED 2026 (27th International Conference on Artificial Intelligence in Education). Published version: Artificial Intelligence in Education, LNCS vol. 16582, Springer, Cham, first online 25 June 2026 (cite as 2027)
Journal refIn: Blanchard, E.G., Chen, G., Chi, M., Isotani, S. (eds) Artificial Intelligence in Education. AIED 2026. Lecture Notes in Computer Science, vol 16582. Springer, Cham (2027)
MedScore: Generalizable Factuality Evaluation of Free-Form Medical Answers by Domain-adapted Claim Decomposition and Verification
Heyuan Huang, Alexandra DeLucia, Vijay Murari Tiyyala, Mark Dredze
机构
*
Center for Language and Speech Processing(语言与语音处理中心)
;
Johns Hopkins University(约翰霍普金斯大学)
专题命中
领域大模型
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL
CommentsAdded generalizability experiment and examples on non-medical free-form answer. Added ablation study for MedCorp verification corpus and MedScore decomposition prompt
Journal refFindings of the Association for Computational Linguistics: ACL 2026, pages 14149-14180, San Diego, California, United States. Association for Computational Linguistics
The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure
能力悖论:更聪明的审计员如何使多智能体系统更不安全
Qiqi Liu, Runhan Song, Shilin Ye
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
Max Planck Institute for Security and Privacy(马克斯·普朗克安全与隐私研究所)
;
Henan Yinzhu Safety Technology Co., Ltd.(河南亿众安全技术有限公司)
;
Harbin Institute of Technology, Faculty of Computing(哈尔滨工业大学计算机学院)
专题命中
领域大模型
:large language model(abstract);language model(abstract);prompting(abstract);分类 cs.AI
Deep Expert Injection for Anchoring Retinal VLMs with Domain-Specific Knowledge
深度专家注入用于带有领域特定知识的视网膜视觉语言模型的锚定
Shuai Lu, Meng Wang, Jia Guo, Jiawei Du, Bo Liu, Shengzhu Yang, Weihang Zhang, Huazhu Fu, Huiqi Li
机构
*
Beijing Institute of Technology, Beijing, China(北京理工大学)
;
National University of Singapore, Singapore(新加坡国立大学)
;
Tsinghua University, Beijing, China(清华大学)
;
The Hong Kong Polytechnic University, Hong Kong(香港理工大学)
Comments8 pages, 6 figures, published in the proceedings of EDULEARN26
Journal refK. Rajaratnam, W. Gan, Y. Sun (2026) TOWARDS REDUCING FOREIGN LANGUAGE ANXIETY USING LEVEL-APPROPRIATE EMBODIED CONVERSATIONAL AGENTS, EDULEARN26 Proceedings, Article 1459
Benchmarking Resource-Efficient LLMs for Research Topic Ontology Generation in the Biomedical Field
用于生物医学领域研究主题本体生成的资源高效语言模型基准测试
Tanay Aggarwal, Angelo Salatino, Francesco Osborne, Enrico Motta
机构
*
Knowledge Media Institute, The Open University, Milton Keynes, UK(开放大学知识媒体研究所)
;
Department of Business and Law, University of Milano-Bicocca, Milan, IT(米兰-比科卡大学商业与法律系)
专题命中
领域大模型
:large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL
Empowering Medical Equipment Sustainability in Low-Resource Settings: An AI-Powered Diagnostic and Support Platform for Biomedical Technicians
赋能低资源环境下的医疗设备可持续性:一个由AI驱动的诊断与支持平台,用于生物医学技术人员
Bernes Lorier Atabonfack, Ahmed Tahiru Issah, Mohammed Hardi Abdul Baaki, Clemence Ingabire, Tolulope Olusuyi, Maruf Adewole, Udunna C. Anazodo, Timothy X Brown
机构
*
Carnegie Mellon University Africa(卡内基梅隆大学非洲分校)
;
Medical Artificial Intelligence Laboratory(医学人工智能实验室)
;
University of Pennsylvania(宾夕法尼亚大学)
;
McGill University(麦吉尔大学)
专题命中
领域大模型
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
VASP Agent: An Agentic Framework for Autonomous First-principles Calculations
VASP智能体:一种用于自主第一性原理计算的智能框架
Zeyu Xia, Jinzhe Ma, Congjie Zheng, Zhongyao Wang, Shufei Zhang, Yuqiang Li, Hang Su, P. Hu, Changshui Zhang, Xingao Gong, Wanli Ouyang, Lei Bai, Dongzhan Zhou, Mao Su
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Department of Computer Science and Technology(计算机科学与技术系)
;
School of Physical Science and Technology(物理科学与技术学院)
;
Department of Automation(自动化系)
;
Beijing National Research Center for Information Science and Technology (BNRist)(北京信息科学与技术国家研究中心)
;
School of Chemistry and Chemical Engineering(化学与化工学院)
;
Key Laboratory of Computational Physical Sciences (Ministry of Education)(计算物理科学重点实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究所)
专题命中
领域大模型
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews
TAMA:一种使用多智能体大语言模型进行临床访谈的人机协作主题分析框架
Huimin Xu, Seungjun Yi, Terence Lim, Jiawei Xu, Andrew Well, Carlos Mery, Aidong Zhang, Yuji Zhang, Heng Ji, Keshav Pingali, Yan Leng, Ying Ding
机构
*
School of Information, University of Texas at Austin(信息学院,德克萨斯大学奥斯汀分校)
;
Department of Biomedical Engineering, University of Texas at Austin(生物医学工程系,德克萨斯大学奥斯汀分校)
;
College of Natural Sciences, University of Texas at Austin(自然科学院,德克萨斯大学奥斯汀分校)
;
Graphen, Inc.(Graphen公司)
;
Department of Cardiac Surgery, Division of Pediatric Cardiac Surgery, Vanderbilt University School of Medicine(心脏外科系,范德比尔特大学医学中心)
;
Pediatric Heart Institute, Monroe Carell Jr. Children’s Hospital at Vanderbilt(儿童心脏研究所,范德比尔特儿童医院)
;
Department of Computer Science, University of Virginia(计算机科学系,弗吉尼亚大学)
;
Department of Computer Science, University of Illinois at Urbana-Champaign(计算机科学系,伊利诺伊大学厄巴纳-香槟分校)
;
Department of Computer Science, University of Texas at Austin(计算机科学系,德克萨斯大学奥斯汀分校)
;
McCombs School of Business, University of Texas at Austin(麦克阿瑟商学院,德克萨斯大学奥斯汀分校)
;
Dell Medical School, University of Texas at Austin(德克萨斯大学奥斯汀分校德莱尔医学院)
专题命中
领域大模型
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL
Simulating Validity: Modal Decoupling in MLLM Generated Feedback on Science Drawings
模拟有效性:在MLLM生成的科学图表反馈中的模态解耦
Arne Bewersdorff, Nejla Yuruk, Xiaoming Zhai
机构
*
University of Georgia, AI4STEM Education Center(佐治亚大学AI4STEM教育中心)
;
Gazi University, Department of Mathematics and Science Education(加齐大学数学与科学教育系)
专题命中
领域大模型
:large language model(abstract);language model(abstract);prompting(abstract);分类 cs.AI