LAAF: A Layered Accountability Architecture Framework for LLM Applications
LAAF:面向大语言模型应用的分层问责架构框架
Prachi Chaturvedi, Shahnawaz Ahmad, Ehsan Nowroozi, Muhammad Waqas, George Loukas, Alireza Jolfaei, Lucas Cordeiro, Pierre Dantas
机构
*
School of Computer Science Engineering and Technology, Bennett University(班尼特大学计算机科学与工程技术学院)
;
Centre for Sustainable Cyber Security (CS2), University of Greenwich(格林威治大学可持续网络安全中心)
;
School of Electrical Engineering, Computing & Mathematical Sciences, Curtin University(科廷大学电气工程、计算与数学科学学院)
;
Systems and Software Security (S3), University of Manchester(曼彻斯特大学系统与软件安全中心)
专题命中
领域大模型
:LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI
Can LLMs Accurately Score Medical Diagnoses and Clinical Reasoning?
LLM能否准确评分医学诊断和临床推理?
Amy Rouillard, Sitwala Mundia, Linda Camara, Ziyaad Dangor, Michael Cameron Gramanie, Ismail Kalla, Shabir A. Madhi, Kajal Morar, Marlvin T. Ncube, Haroon Saloojee, Bruce A. Bassett
机构
*
Wits MIND Institute, University of the Witwatersrand, Johannesburg, South Africa(维特士心理研究所,沃斯兰德大学,约翰内斯堡,南非)
;
Grai Labs, Cape Town, South Africa(格雷实验室,开普敦,南非)
;
South African Medical Research Council Vaccines and Infectious Diseases Analytics Research Unit, Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(南非医学研究理事会疫苗和传染病分析研究组,健康科学学院,沃斯兰德大学,约翰内斯堡,南非)
;
Department of Internal Medicine, Charlotte Maxeke Johannesburg Academic Hospital, and Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(内科学系,查理·马克斯凯约翰内斯堡学术医院,以及健康科学学院,沃斯兰德大学,约翰内斯堡,南非)
;
Department of Paediatrics and Child Health, Faculty of Health Sciences, University of the Witwatersrand, Johannesburg, South Africa(儿科学与儿童健康系,健康科学学院,沃斯兰德大学,约翰内斯堡,南非)
;
Wits MIND Institute, University of the Witwatersrand, Johannesbu(维特士心理研究所,沃斯兰德大学,约翰内斯堡)
专题命中
领域大模型
:LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
MedFabric: Gold Evidence Hides the Difficulty of Word-Level Medical Fabrication Detection
MedFabric:黄金证据掩盖了词级医学编造检测的难度
Tung Sum Thomas Kwok, Qian Qian, Xiaofeng Lin, Dongxu Zhang, Jun Han, Zhichao Yang, Davin Hill, Tamer Soliman, Sanjit Singh Batra, Robert Tillman, Guang Cheng
专题命中
领域大模型
:LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
在域转移下对用于乳腺钼靶成像的基础模型的稳健性进行基准测试
Giang Nguyen, Raghav Mehta, Emma A.M. Stanley, Tian Xia, Thi Hao Nguyen, Hieu Pham, Ben Glocker
机构
*
College of Engineering and Computer Science, VinUniversity(工程与计算机科学学院,文大大学)
;
Imperial College London(伦敦帝国理工学院)
;
Radiology Department, Vietnam National Cancer Hospital(越南国家癌症医院放射科)
;
VinUni-Illinois Smart Health Center, VinUniversity(文大大学 - 伊利诺伊智能健康中心,文大大学)
;
The Computer Vision and Medical AI Lab, VinUniversity(计算机视觉与医学人工智能实验室,文大大学)
MVC-Bench: Benchmarking Calibration of Medical Vision-Language Models
MVC-Bench:医学视觉-语言模型的校准基准测试
Ashshak Sharifdeen, Shihab Aaqil Ahamed, Ufaq Khan, Muhammad Akhtar Munir Sujair Ibrahim, Mohamed Rafeek Mareer Ahamed, Yutong Xie, Imran Razzak, Muhammad Haris Khan
机构
*
Mohamed bin Zayed University of AI(穆罕默德·本·扎耶德人工智能大学)
;
Sabaragamuwa University of Sri Lanka(斯里兰卡萨巴拉加穆瓦大学)
;
Digital Platform Development, SLT PLC(SLT公共有限公司数字平台开发部)
MOMO: A framework for seamless physical, verbal, and graphical robot skill learning and adaptation
MOMO:一种无缝物理、语言和图形机器人技能学习与适应框架
Markus Knauer, Edoardo Fiorini, Maximilian Mühlbauer, Stefan Schneyer, Promwat Angsuratanawech, Florian Samuel Lay, Timo Bachmann, Samuel Bustamante, Korbinian Nottensteiner, Freek Stulp, Alin Albu-Schäffer, João Silvério, Thomas Eiband
机构
*
German Aerospace Center (DLR), Institute of Robotics and Mechatronics (RMC)(德国航空航天中心(DLR)机器人与机电研究所)
;
School of Computation, Information and Technology (CIT), Technical University of Munich (TUM)(计算、信息与技术学院(CIT),慕尼黑技术大学)
Comments59 pages, 4 figures, 8 tables. Includes Appendix A (variable definitions), Appendix B (supplementary tables), and Online Appendix C (meta prompts)
机构
*
Department of Computer Science and Engineering, HKUST, Hong Kong SAR, China(香港科技大学计算机科学与工程系)
;
School of Law, Tsinghua University, Beijing, China(清华大学法学院)
专题命中
领域大模型
:large language model(abstract);language model(abstract);分类 cs.CL
机构
*
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
LinkedIn Corporation(领英公司)
;
Northeastern University(东北大学)
;
University of California, Davis(加州大学戴维斯分校)