Crisis-Bench: Benchmarking Strategic Ambiguity and Reputation Management in Large Language Models
Crisis-Bench: 大型语言模型中战略模糊与声誉管理的基准测试
Cooper Lin, Maohao Ran, Yanting Zhang, Zhenglin Wan, Hongwei Fan, Yibo Xu, Yike Guo, Wei Xue, Jun Song
机构
*
Hong Kong University of Science and Technology(香港科技大学)
;
Hong Kong Baptist University(香港 Baptist 大学)
;
National University of Singapore(新加坡国立大学)
;
Imperial College London(伦敦帝国理工学院)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI
Jailbreaking Large Language Models through Iterative Tool-Disguised Attacks via Reinforcement Learning
通过强化学习的迭代工具伪装攻击对大型语言模型进行劫持
Zhaoqi Wang, Zijian Zhang, Daqing He, Pengtao Kou, Xin Li, Jiamou Liu, Jincheng An, Yong Liu
机构
*
School of Cyberspace Science and Technology, Beijing Institute of Technology(信息科学与技术学院,北京理工大学)
;
School of Computer Science and Technology, Beijing Institute of Technology(计算机科学与技术学院,北京理工大学)
;
School of Computer Science, University of Auckland(计算机科学学院,奥克兰大学)
;
QAX Security Center, Qi-AnXin Technology Group Inc.(QAX安全中心,启安新科技集团)
;
Qi-AnXin Technology Group Inc.(启安新科技集团)
;
Zhongguancun Laboratory(中关村实验室)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI
Reinforcement Learning of Large Language Models for Interpretable Credit Card Fraud Detection
基于大语言模型的强化学习在可解释性信用卡欺诈检测中的应用
Cooper Lin, Yanting Zhang, Maohao Ran, Wei Xue, Hongwei Fan, Yibo Xu, Zhenglin Wan, Sirui Han, Yike Guo, Jun Song
机构
*
Hong Kong University of Science and Technology(香港科学与技术大学)
;
Hong Kong Baptist University(香港 Baptist 大学)
;
Imperial College London(伦敦帝国理工学院)
;
National University of Singapore(新加坡国立大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
Efficient Inference for Noisy LLM-as-a-Judge Evaluation
高效噪声LLM-as-a-Judge评估
Yiqun T Chen, Sizhu Lu, Sijia Li, Moran Guo, Shengyi Li
机构
*
Departments of Biostatistics and Computer Science, Johns Hopkins University(生物统计学与计算机科学系,约翰霍普金斯大学)
;
Department of Statistics, University of California, Berkeley(统计学系,加州大学伯克利分校)
;
Department of Biostatistics, University of California, Los Angeles(生物统计学系,加州大学洛杉矶分校)
;
Department of Biostatistics, Johns Hopkins University(生物统计学系,约翰霍普金斯大学)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.LG
MedRiskEval: Medical Risk Evaluation Benchmark of Language Models, On the Importance of User Perspectives in Healthcare Settings
MedRiskEval: 医疗风险评估基准,语言模型在医疗领域中的重要性:用户视角的重要性
Jean-Philippe Corbeil, Minseon Kim, Maxime Griot, Sheela Agarwal, Alessandro Sordoni, Francois Beaulieu, Paul Vozila
机构
*
Microsoft Healthcare & Life Sciences(微软医疗与生命科学)
;
Microsoft Research Montréal, Canada(微软研究蒙特利尔分校)
;
Université catholique de Louvain, Belgium(列日大学)
;
Mila, Université de Montréal, Canada(蒙特利尔大学Mila)
专题命中
评测与基准
:language model(title,abstract);large language model(abstract);分类 cs.CL
机构
*
Harbin Institute of Technology(哈尔滨工业大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Xidian University(西安电子科技大学)
;
East China Normal University(东华大学)
;
Wuhan University(武汉大学)
;
Southeast University(东南大学)
;
National University of Defense Technology(国防科技大学)
;
Chinese University of Hong Kong(香港中文大学)
Securing the AI Supply Chain: What Can We Learn From Developer-Reported Security Issues and Solutions of AI Projects?
保障人工智能供应链:我们能从AI项目开发人员报告的安全问题和解决方案中获得什么启示?
The Anh Nguyen, Triet Huynh Minh Le, M. Ali Babar
机构
*
School of Computer Science
;
Information Technology,\ University Adelaide Australia
;
Information Technology,\ University \&\ Systems Adelaide Australia
;
Information Technology,\ University
;
Information Technology,\ University \&\ Systems
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI
MAGneT: Coordinated Multi-Agent Generation of Synthetic Multi-Turn Mental Health Counseling Sessions
MAGneT: 协调多智能体生成合成多轮心理健康咨询会话
Aishik Mandal, Tanmoy Chakraborty, Iryna Gurevych
机构
*
Ubiquitous Knowledge Processing Lab (UKP Lab), Department of Computer Science and Hessian Center for AI (hessian.AI), Technische Universität Darmstadt(德累斯顿技术大学计算机科学系、普遍知识处理实验室(UKP Lab)、黑森人工智能中心(hessian.AI))
;
National Research Center for Applied Cybersecurity ATHENE, Germany(应用网络安全国家研究中心ATHENE,德国)
;
Department of Electrical Engineering, Indian Institute of Technology Delhi, India(印度德里印度理工学院电气工程系)
;
Yardi School of Artificial Intelligence, Indian Institute of Technology Delhi, India(印度德里印度理工学院人工智能学院)