SpecDiff-2: Scaling Diffusion Drafter Alignment For Faster Speculative Decoding
专题命中 安全评测 :alignment(title);分类 cs.CL
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 安全评测 :alignment(title);分类 cs.CL
机构 * Carnegie Mellon University(卡内基梅隆大学)
专题命中 安全评测 :alignment(title);分类 cs.LG
机构 * NII LLMC(日本信息处理学会大语言模型中心) ; The University of Tokyo(东京大学) ; NAIST(日本科学技术大学) ; Nagoya Institute of Technology(名古屋技术大学)
专题命中 安全评测 :alignment(title);分类 cs.CL
Comments Accepted to EMNLP 2025 (Main Conference). Models and evaluation results available at: https://github.com/llm-jp/massive-sft
机构 * Computer Vision Lab, CAIDAS & IFI, University of Würzburg, Germany(计算机视觉实验室,CAIDAS与IFI,乌尔姆大学,德国) ; University of Bucharest, Romania(布加勒斯特大学,罗马尼亚)
专题命中 安全评测 :safety(title);分类 cs.AI
Comments Accepted at CVPR 2025
专题命中 安全评测 :trustworthy(title);分类 cs.LG
Comments 1tables,6 figs,11pages
机构 * Department of Statistics, Eskisehir Technical University, Turkiye(埃斯基谢普大学统计系) ; Leiden Institute of Advanced Computer Science, Leiden University, the Netherlands(莱顿大学高级计算机科学研究所) ; Faculty of Mathematics and Information Science, Warsaw University of Technology, Poland(华沙理工大学数学与信息科学学院) ; Informatics and Mechanics, University of Warsaw, Faculty of Mathematics, Poland(华沙大学信息技术与力学系)
专题命中 安全评测 :trustworthy(title);分类 cs.LG
Comments Accepted at 28th International Conference on Discovery Science 2025
Journal ref In: Džeroski, S., Levatić, J., Pio, G., Simidjievski, N. (eds) Discovery Science. DS 2025. Lecture Notes in Computer Science, vol 16090. Springer, Cham
机构 * The Graduate University for Advanced Studies (SOKENDAI)(高级研究大学(SOKENDAI)) ; National Institute of Informatics(信息研究所)
专题命中 安全评测 :alignment(title);分类 cs.CL
专题命中 安全评测 :trustworthy(title);分类 cs.AI
Comments 9 pages, 1 figure, 4 tables
专题命中 安全评测 :trustworthy(title);分类 cs.CY
机构 * Advanced Institute of So-Go-Chi (Convergence Knowledge) Informatics(融合知识研究院) ; Tohoku University(东北大学) ; Japan Advanced Institute of Science and Technology(日本先进科学研究院) ; Faculty of Information and Communication Technology(信息与通信技术学院) ; Mahidol University(玛希敦大学)
专题命中 安全评测 :trustworthy(title);分类 cs.AI
Journal ref Natural Language Processing and Information Systems (NLDB 2025)
机构 * Amirkabir University of Technology(阿姆irkabir技术大学) ; Part AI Research Center(Part人工智能研究中心) ; University of Mazandaran(马赞德兰大学) ; King’s College London(伦敦国王学院)
专题命中 安全评测 :alignment(title);分类 cs.CL
Comments Preprint. Under review
机构 * Department of Electrical and Electronic Engineering, The University of Hong Kong, Hong Kong SAR, China(香港大学电子与电气工程系) ; College of Information Science and Electronic Engineering, Zhejiang University, Hangzhou, China(浙江大学信息科学与电子工程学院) ; Department of Computer Science and Engineering, University of Notre Dame, Notre Dame, IN, USA(Notre Dame 大学计算机科学与工程系) ; Center for Advanced Semiconductor and Integrated Circuit, The University of Hong Kong, Hong Kong SAR, China(香港大学先进半导体与集成电路中心)
专题命中 安全评测 :trustworthy(title);分类 cs.LG
机构 * AITech Lab(AITech实验室) ; Computer Science and Engineering Faculty(计算机科学与工程学院) ; Ho Chi Minh City University of Technology(胡志明市技术大学) ; VNUHCM
专题命中 安全评测 :alignment(title);分类 cs.AI
专题命中 安全评测 :alignment(title);分类 cs.AI
Comments The result is no longer believeable. Teaching force issue exists in the infer time of LLM
机构 * University of Notre Dame(圣约翰大学) ; Northeastern University(东北大学) ; Virginia Tech(弗吉尼亚理工大学) ; University of Washington(华盛顿大学) ; Johns Hopkins University(约翰霍普金斯大学)
专题命中 安全评测 :trustworthy(title);分类 cs.CL
机构 * RTFL Project Contributor(RTFL项目贡献者)
专题命中 安全评测 :trustworthy(title);分类 cs.AI
Comments 6 pages (IEEE conference format), 10 figures. Source code available at https://github.com/abhitall/federated-credit-risk-rtfl.git
专题命中 安全评测 :trustworthy(title);分类 cs.CL
机构 * University of Twente(代尔夫特理工大学) ; Inria(法国国家信息与自动化技术研究所) ; Institute of Science and Technology Austria(奥地利科学与技术研究所)
专题命中 安全评测 :trustworthy(title);分类 cs.LG
机构 * Department of Computer Science, University of Manchester(曼彻斯特大学计算机科学系) ; Idiap Research Institute(Idiap研究机构) ; National Biomarker Centre, CRUK-MI, Univ. of Manchester(国家生物标记中心,CRUK-MI,曼彻斯特大学)
专题命中 安全评测 :alignment(title);分类 cs.CL
Comments 19 pages, 8 figures, 8 tables
机构 * Reserve Bank of Australia(澳大利亚储备银行)
专题命中 安全评测 :trustworthy(title);分类 cs.CL
Comments 16 pages, 6 figures
机构 * IFCA.unican.es(IFCA大学)
专题命中 安全评测 :trustworthy(title);分类 cs.AI
Comments 17 pages, 2 figures
机构 * Sandia National Laboratories(桑迪亚国家实验室) ; The George Washington University(乔治·华盛顿大学) ; University of Michigan(密歇根大学) ; The University of Texas at Austin(德克萨斯大学奥斯汀分校)
专题命中 安全评测 :trustworthy(title);分类 cs.LG
机构 * Machine Intelligence, Interaction, and Imagination (Mi 3 ) Laboratory(机器智能、交互与想象(Mi 3)实验室) ; University of California, Merced(加州大学默塞德分校) ; Aalborg Universitet(奥胡斯大学)
专题命中 安全评测 :safety(title);分类 cs.AI
专题命中 安全评测 :trustworthy(title);分类 cs.AI
专题命中 安全评测 :alignment(title);分类 cs.AI
专题命中 安全评测 :alignment(title);分类 cs.LG
Comments Accepted at the 2025 IEEE/ACM 22nd International Conference on Mining Software Repositories (MSR) - Data and Tool Showcase Track
专题命中 安全评测 :trustworthy(title);分类 cs.AI
专题命中 安全评测 :safety(abstract,comments);AI safety(abstract,comments);分类 cs.CY
Comments To be published in the proceedings of the 2024 Conference on Frontier AI Safety Commitments. 6 pages
专题命中 安全评测 :trustworthy(title);分类 cs.AI
专题命中 安全评测 :trustworthy(title);分类 cs.LG