机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
University of Pennsylvania(宾夕法尼亚大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Hong Kong Polytechnic University(香港理工大学)
;
Nanjing University of Posts and Telecommunications(南京邮电大学)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL
X-MuTeST: A Multilingual Benchmark for Explainable Hate Speech Detection and A Novel LLM-consulted Explanation Framework
X-MuTeST:一个用于可解释仇恨言论检测的多语言基准及一种新颖的LLM咨询解释框架
Mohammad Zia Ur Rehman, Sai Kartheek Reddy Kasu, Shashivardhan Reddy Koppula, Sai Rithwik Reddy Chirra, Shwetank Shekhar Singh, Nagendra Kumar
机构
*
Indian Institute of Technology Indore(印度理工学院Indore)
;
Indian Institute of Information Technology Dharwad(印度信息科技学院Dharwad)
;
Arizona State University(亚利桑那州立大学)
;
Indian Institute of Technology Mandi(印度理工学院Mandi)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL
Context-aware LLM-based AI Agents for Human-centered Energy Management Systems in Smart Buildings
面向人类中心的基于大语言模型的AI代理用于智能建筑中的能源管理系统
Tianzhi He, Farrokh Jazizadeh
机构
*
organization= School of Civil \& Environmental Engineering
;
Construction Management, The University of Texas at San Antonio , addressline= BSE 1.310, One UTSA Circle , city= San Antonio , postcode= 78249 , state= TX , country= U.S.
;
organization= Department of Civil
;
Environmental Engineering, Virginia Polytechnic Institute
;
State University , addressline= 200 Patton Hall, 750 Drillfield , city= Blacksburg , postcode= 24060 , state= VA , country= U.S.
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI
Heaven-Sent or Hell-Bent? Benchmarking the Intelligence and Defectiveness of LLM Hallucinations
天降还是地狱?评估大语言模型幻觉的智能与缺陷
Chengxu Yang, Jingling Yuan, Siqi Cai, Jiawei Jiang, Chuang Hu
机构
*
Wuhan University of Technology(武汉理工大学)
;
Hubei Key Laboratory of Transportation Internet of Things(湖北省交通运输物联网重点实验室)
;
BreathingCORE
;
Wuhan University(武汉大学)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL
DarkPatterns-LLM: A Multi-Layer Benchmark for Detecting Manipulative and Harmful AI Behavior
DarkPatterns-LLM: 一种多层基准用于检测操纵性和有害的AI行为
Sadia Asif, Israel Antonio Rosales Laguan, Haris Khan, Shumaila Asif, Muneeb Asif
机构
*
Rensselaer Polytechnic Institute(拉特格斯理工学院)
;
National University of Sciences and Technology(国立科学与技术大学)
;
School of Electrical Engineering & Computer Science(电子与计算机科学学院)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI
A Real-World Evaluation of LLM Medication Safety Reviews in NHS Primary Care
对NHS初级医疗中基于LLM的药物安全审查系统的真实世界评估
Oliver Normand, Esther Borsi, Mitch Fruin, Lauren E Walker, Jamie Heagerty, Chris C. Holmes, Anthony J Avery, Iain E Buchan, Harry Coppock
机构
*
i.AI, Department for Science, Innovation, and Technology(i.AI,科学、创新与技术部门)
;
Centre for Experimental Therapeutics, University of Liverpool(实验治疗中心,利物浦大学)
;
Civic Health Innovation Labs, University of Liverpool(公民健康创新实验室,利物浦大学)
;
Downing Street(唐宁街10号)
;
Department of Statistics, University of Oxford(统计系,牛津大学)
;
Ellison Institute of Technology(埃利森技术研究所)
;
Centre for Academic Primary Care, University of Nottingham(学术初级护理中心,诺丁汉大学)
;
The UK AI Security Institute(英国人工智能安全研究所)
;
Department of Computing, Imperial College London(计算系,伦敦帝国学院)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI
FASTRIC: Prompt Specification Language for Verifiable LLM Interactions
FASTRIC:用于可验证大语言模型交互的提示规范语言
Wen-Long Jin
机构
*
Department of Civil and Environmental Engineering(土木与环境工程系)
;
California Institute for Telecommunications and Information Technology(电信与信息科技学院)
;
Institute of Transportation Studies(交通研究学院)
;
University of California, Irvine, CA 92697-3600(加州大学伊维德分校)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL
A Multi-Agent LLM Defense Pipeline Against Prompt Injection Attacks
针对提示注入攻击的多智能体LLM防御管道
S M Asif Hossain, Ruksat Khan Shayoni, Mohd Ruhul Ameen, Akif Islam, M. F. Mridha, Jungpil Shin
机构
*
School of Computing, Wichita State University, Kansas, USA(威斯康星州立大学计算机学院)
;
College of Engineering and Computer Sciences, Marshall University, Huntington, WV, USA(马歇尔大学工程与计算机科学学院)
;
Department of Computer Science and Engineering, University of Rajshahi, Bangladesh(拉贾沙希大学计算机科学与工程系)
;
Department of Computer Science and Engineering, American International University-Bangladesh, Dhaka, Bangladesh(美国国际大学-孟加拉国计算机科学与工程系)
;
School of Computer Science and Engineering, The University of Aizu, Aizuwakamatsu, Japan(立命馆大学计算机科学与工程学院)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG
机构
*
Department of Computer Science, Shantou University(汕头大学计算机科学系)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
BNRIST, Tsinghua University(清华大学脑科学与类脑智能研究院)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL
机构
*
Department of Computer and Network Engineering, College of Information Technology, United Arab Emirates University(阿联酋大学计算机与网络工程系)
;
Technology Innovation Institute(技术创新研究所)
;
Eötvös Loránd University(埃奥瓦大学)
;
Department of Computer Science, Guelma University(古尔马大学计算机科学系)
;
De Montfort University(德蒙福特大学)
;
Khalifa University of Science and Technology(卡里玛科学技术大学)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI
RadOnc-GPT: An Autonomous LLM Agent for Real-Time Patient Outcomes Labeling at Scale
RadOnc-GPT:一种用于大规模实时患者结果标注的自主大语言模型代理
Jason Holmes, Yuexing Hao, Mariana Borras-Osorio, Federico Mastroleo, Santiago Romero Brufau, Valentina Carducci, Katie M Van Abel, David M Routman, Andrew Y. K. Foong, Liv M Muller, Satomi Shiraishi, Daniel K Ebner, Daniel J Ma, Sameer R Keole, Samir H Patel, Mirek Fatyga, Martin Bues, Brad J Stish, Yolanda I Garces, Michelle A Neben Wittich, Robert L Foote, Sujay A Vora, Nadia N Laack, Mark R Waddle, Wei Liu
机构
*
Mayo Clinic, Phoenix, AZ, USA(梅奥诊所,凤凰城,亚利桑那州,美国)
;
Mayo Clinic, Rochester, MN, USA(梅奥诊所,罗切斯特,明尼苏达州,美国)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI