Automotive-ENV: Benchmarking Multimodal Agents in Vehicle Interface Systems
机构 * Australian Artificial Intelligence Institute(澳大利亚人工智能研究所) ; University of Liverpool(利物浦大学)
专题命中 安全评测 :safety(abstract);分类 cs.CL
Comments 10 pages, 5 figures,
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
机构 * Australian Artificial Intelligence Institute(澳大利亚人工智能研究所) ; University of Liverpool(利物浦大学)
专题命中 安全评测 :safety(abstract);分类 cs.CL
Comments 10 pages, 5 figures,
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL
机构 * Luxembourg Institute of Science and Technology (LIST)(卢森堡科学与技术研究院) ; RMT Labs(RMT实验室)
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL
Comments Accepted at the IEEE GLOBECOM Workshops 2025: "Large AI Model over Future Wireless Networks"
机构 * University of Southern California(南加州大学) ; University of Queensland(昆士兰大学) ; University of California, San Diego(加州大学圣地亚哥分校) ; University of Buffalo(布法罗大学) ; University of California, Merced(加州大学默塞德分校)
专题命中 安全评测 :alignment(abstract);分类 cs.CL
机构 * Amazon(亚马逊公司) ; Alexa, Amazon(亚马逊Alexa部门) ; Virginia Tech(弗吉尼亚理工大学) ; University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
专题命中 安全评测 :alignment(abstract);分类 cs.LG
Comments 17 pages, 9 figures, 14 tables, Findings of the Association for Computational Linguistics: EMNLP 2025
机构 * Youngstown State University(扬斯敦州立大学)
专题命中 安全评测 :alignment(abstract);分类 cs.CL
机构 * Ixent Games
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI
Comments Version 2: This article consolidates and replaces a previous version to present the complete research in a single, comprehensive manuscript
机构 * Stability AI ; SketchX, University of Surrey(SketchX,大学)
专题命中 安全评测 :alignment(abstract);分类 cs.AI
Comments Project Page: https://hmrishavbandy.github.io/sd35flash/
机构 * University of California, Berkeley(加州大学伯克利分校)
专题命中 安全评测 :alignment(abstract);分类 cs.AI
Comments 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: Evaluating the Evolving LLM Lifecycle: Benchmarks, Emergent Abilities, and Scaling
机构 * Peking University(北京大学) ; LLM-Core
专题命中 安全评测 :trustworthy(abstract);分类 cs.CL
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI
Comments 11 pages, 2 figures, 2 tables
机构 * Department of Computer Science University of Quebec in Outaouais (UQO)(计算机科学系魁北克大学 Outaouais 分校)
专题命中 安全评测 :safety(abstract);分类 cs.AI
机构 * University of Pennsylvania(宾夕法尼亚大学) ; BigHat Biosciences(BigHat生物技术公司)
专题命中 安全评测 :safety(abstract);分类 cs.LG
Comments NeurIPS 2025 AI4Science Workshop and NeurIPS 2025 Multi-modal Foundation Models and Large Language Models for Life Sciences Workshop
专题命中 安全评测 :alignment(abstract);分类 cs.CL
机构 * Arizona State University(亚利桑那州立大学) ; University of California, Berkeley(加州大学伯克利分校)
专题命中 安全评测 :safety(abstract);分类 cs.AI
Comments 17 pages, 9 figures
专题命中 安全评测 :alignment(abstract);分类 cs.CL
Comments Accepted by EMNLP 2025 Main (Oral), 25 pages, 7 figures
机构 * University of Southern Mississippi(密苏里州南方大学)
专题命中 安全评测 :alignment(abstract);分类 cs.CL
机构 * Department of Computer Science University of Pittsburgh(计算机科学系宾夕法尼亚大学)
专题命中 安全评测 :alignment(abstract);分类 cs.AI
Comments To Appear in EMNLP2025
机构 * School of Computing Technologies RMIT University Melbourne VIC Australia ; Department of Biological Sciences ; Department of Computer Science \& Information Systems Birla Institute of Technology ; Institute of Innovation, Science ; Sustainability Federation University Australia Ballarat VIC Australia ; RMIT University ; Birla Institute of Technology ; Federation University Australia
专题命中 安全评测 :trustworthy(abstract);分类 cs.LG
机构 * Worcester Polytechnic Institute(沃斯特理工大学)
专题命中 安全评测 :alignment(abstract);分类 cs.CL
机构 * National Yang Ming Chiao Tung University(国家阳明交通大学) ; NVIDIA
专题命中 安全评测 :alignment(abstract);分类 cs.CL
Comments Accepted for EMNLP 2025 main conference
机构 * NUS(国立新加坡大学) ; SonarSource SA
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI
Comments 4 pages
机构 * University of Rostock(罗斯托克大学) ; University of Rostock, University of Marburg(罗斯托克大学、马堡大学)
专题命中 安全评测 :alignment(abstract);分类 cs.LG
Comments 10 pages, 7 figures, added repo link
机构 * EPFL(苏黎世联邦理工学院) ; MIT(麻省理工学院) ; Georgia Institute of Technology(佐治亚理工学院)
专题命中 安全评测 :alignment(abstract);分类 cs.CL
Comments EMNLP 2025. Project Page at https://language-to-cognition.epfl.ch
机构 * Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(莫扎伊德大学人工智能学院) ; University of Notre Dame(诺特丹大学) ; IBM Research(IBM研究院)
专题命中 安全评测 :DPO(abstract);分类 cs.CL
机构 * Department of Computer Science and Engineering, University of Notre Dame(诺丁汉大学计算机科学与工程系) ; MIT(麻省理工学院) ; CalTech(加州理工学院) ; MBZUAI(穆斯林人工智能研究所) ; CMU(卡内基梅隆大学) ; Department of Chemistry & Biochemistry, University of Notre Dame(诺丁汉大学化学与生物化学系)
专题命中 安全评测 :alignment(abstract);分类 cs.CL
机构 * Computer Science Department(计算机科学系) ; Applied Economics Department(应用经济学系)
专题命中 安全评测 :trustworthy(abstract);分类 cs.AI
Comments 6 pages, 3 figures, 4 tables, 1 algorithm, accepted in the Robustness and Security of Large Language Models (ROSE-LLM) special session at ICMLA 2025
专题命中 安全评测 :safety(abstract);分类 cs.AI
机构 * Laboratory for Emerging Intelligence(新兴智能实验室) ; University of California, San Diego(加州大学圣地亚哥分校)
专题命中 安全评测 :alignment(abstract);分类 cs.CL
Comments Published at EMNLP 2025 (Main)
机构 * AI Singapore(AI新加坡) ; VISTEC ; MBZUAI
专题命中 安全评测 :alignment(abstract);分类 cs.CL
Comments Accepted to EMNLP 2025 (Main). Model and Dataset: https://huggingface.co/collections/airesearch/wangchan-thai-instruction-6835722a30b98e01598984fd