Do LLM Self-Explanations Help Users Predict Model Behavior? Evaluating Counterfactual Simulatability with Pragmatic Perturbations
LLM自我解释是否有助于用户预测模型行为?通过语用扰动评估反事实可模拟性
Pingjun Hong, Benjamin Roth
机构
*
Faculty of Computer Science, University of Vienna(计算机科学系,维也纳大学)
;
UniVie Doctoral School Computer Science, University of Vienna(UniVie 计算机科学博士学院,维也纳大学)
;
Faculty of Philological and Cultural Studies, University of Vienna(文学与文化研究系,维也纳大学)
CPGPrompt: Translating Clinical Guidelines into LLM-Executable Decision Support
CPGPrompt:将临床指南转化为LLM可执行的决策支持
Ruiqi Deng, Geoffrey Martin, Tony Wang, Gongbo Zhang, Yi Liu, Chunhua Weng, Yanshan Wang, Justin F Rousseau, Yifan Peng
机构
*
Information Science (Health Tech), Cornell Tech, New York, NY, USA(信息科学(健康科技),康奈尔科技,纽约,纽约州)
;
Systems Engineering, Cornell University, Ithaca, NY, USA(系统工程,康奈尔大学,伊萨卡,纽约州)
;
Population Health Sciences, Weill Cornell Medicine, New York, NY, USA(人口健康科学,韦尔医学院,纽约,纽约州)
;
Computer and Information Science, Cornell University, Ithaca, NY, USA(计算机与信息科学,康奈尔大学,伊萨卡,纽约州)
;
Department of Biomedical Informatics, Columbia University, New York, NY, USA(生物医学信息学系,哥伦比亚大学,纽约,纽约州)
;
Department of Medicine, Weill Cornell Medicine, New York, NY, USA(医学系,韦尔医学院,纽约,纽约州)
;
Department of Health Information Management, University of Pittsburgh, Pittsburgh, PA, USA(健康信息管理系,匹兹堡大学,匹兹堡,宾夕法尼亚州)
;
Clinical and Translational Science Institute, University of Pittsburgh, Pittsburgh, PA, USA(临床与转化科学研究所,匹兹堡大学,匹兹堡,宾夕法尼亚州)
;
Department of Neurology, UT Southwestern Medical Center, Dallas, TX, USA(神经病学系,德克萨斯西南医学中心,达拉斯,德克萨斯州)
;
Peter O’Donnell Jr. Brain Institute, UT Southwestern Medical Center, Dallas, TX, USA(彼得·奥·donnell Jr.脑研究所,德克萨斯西南医学中心,达拉斯,德克萨斯州)
;
Clinical Informatics Center, University of Texas Southwestern Medical Center, Dallas, USA(临床信息学中心,德克萨斯西南医学中心,达拉斯,美国)
;
Institute of Artificial Intelligence for Digital Health, Weill Cornell Medicine, New York, NY USA(数字健康人工智能研究所,韦尔医学院,纽约,纽约州)
MARVEL: A Multi Agent-based Research Validator and Enabler using Large Language Models
MARVEL:基于大语言模型的多智能体研究验证器与促进器
Nikhil Mukund, Yifang Luo, Fan Zhang, Lisa Barsotti, Erik Katsavounidis
机构
*
MIT Kavli Institute for Astrophysics and Space Research and LIGO Laboratory(麻省理工学院凯斯利天文与空间研究所及LIGO实验室)
;
Massachusetts Institute of Technology(麻省理工学院)
;
State Key Laboratory of Ocean Sensing & Ocean College(海洋传感国家重点实验室)
;
Zhejiang University(浙江大学)
;
NSF AI Institute for Artificial Intelligence and Fundamental Interactions (IAIFI)(国家科学基金会人工智能与基本相互作用研究所)
;
Cambridge, MA, USA(美国马萨诸塞州剑桥市)
Attractive Metadata Attack: Inducing LLM Agents to Invoke Malicious Tools
吸引性元数据攻击:诱导LLM代理调用恶意工具
Kanghua Mo, Li Hu, Yucheng Long, Zhihao Li
机构
*
Cyberspace Institute of Advanced Technology, Guangzhou University(广州大学网络空间研究院)
;
Department of Electrical and Electronic Engineering, The Hong Kong Polytechnic University(香港理工大学电子与电气工程系)