This Treatment Works, Right? Evaluating LLM Sensitivity to Patient Question Framing in Medical QA
这种治疗有效吗?评估LLM在医疗问答中对患者问题表述的敏感性
Hye Sun Yun, Geetika Kapoor, Michael Mackert, Ramez Kouzy, Wei Xu, Junyi Jessy Li, Byron C. Wallace
机构
*
Northeastern University(东北大学)
;
UC Berkeley(加州大学伯克利分校)
;
UT Austin(德克萨斯大学奥斯汀分校)
;
UT MD Anderson Cancer Center(德克萨斯大学MD安德森癌症中心)
;
Georgia Institute of Technology(佐治亚理工学院)
专题命中
领域大模型
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
Australian Artificial Intelligence Institute, University of Technology Sydney(澳大利亚人工智能研究所,悉尼科技大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Shenzhen University(深圳大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Sichuan University(四川大学)
专题命中
领域大模型
:foundation model(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
LLM-Meta-SR: In-Context Learning for Evolving Selection Operators in Symbolic Regression
LLM-Meta-SR: 上下文学习用于符号回归中演化的选择算子
Hengzhe Zhang, Qi Chen, Bing Xue, Wolfgang Banzhaf, Mengjie Zhang
机构
*
Centre for Data Science and Artificial Intelligence & School of Engineering and Computer Science, Victoria University of Wellington(惠灵顿维多利亚大学数据科学与人工智能中心及工程与计算机科学学院)
;
Department of Computer Science and Engineering, Michigan State University(密歇根州立大学计算机科学与工程系)
专题命中
领域大模型
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Silicon Bureaucracy and AI Test-Oriented Education: Contamination Sensitivity and Score Confidence in LLM Benchmarks
硅 bureaucracy 与 AI 考试导向教育:LLM 测试基准中的污染敏感性与分数可信度
Yiliang Song, Hongjun An, Jiangan Chen, Xuanchen Yan, Huan Song, Jiawei Shao, Xuelong Li
机构
*
Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究院(TeleAI))
;
Guangxi Normal University(广西师范大学)
;
Northwestern Polytechnical University(西北工业大学)
专题命中
领域大模型
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
the School of Chinese Medicine, the Beijing University of Chinese Medicine, Beijing, China(中国中医药大学中医学院,北京中医药大学,北京,中国)
;
the School of Pharmacy, Nanjing University of Chinese Medicine, Nanjing, China(南京中医药大学药学院,南京,中国)
;
the Infectious disease department, Dongfang Hospital, Beijing University of Chinese Medicine, Beijing, China(北京中医药大学东四医院感染科,北京,中国)
;
the Gulou Hospital of Traditional Chinese Medicine of Beijing,Beijing, China(北京中医药大学附属鼓楼中医医院,北京,中国)
;
the Department of Computer Science, Johns Hopkins University, Baltimore, USA(约翰霍普金斯大学计算机科学系,美国马里兰州巴尔的摩市)
;
the Department of Education, Dongzhimen Hospital, Beijing University of Chinese Medicine, Beijing, China(北京中医药大学东直门医院教育部,北京,中国)
;
the Department of Pediatrics, Wangjing Hospital, China Academy of Chinese Medical Sciences, Beijing, China(中国医学科学院北京协和医院儿科部,北京,中国)
;
the School of Information Engineering, Huzhou University, Huzhou, China(湖州大学信息工程学院,湖州,中国)
;
Research Center for Scientific Data Hub, Zhejiang Lab, Hangzhou, China(浙江省实验室科学数据中心研究中心,杭州,中国)
;
the Frontier Basic Research Center, Zhejiang Lab, Hangzhou, China(浙江省实验室前沿基础研究中心,杭州,中国)
;
the Research Center for High Efficiency Computing Infrastructure, Zhejiang Lab, Hangzhou, China(浙江省实验室高效计算基础设施研究中心,杭州,中国)
专题命中
领域大模型
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI