Communication Bias in Large Language Models: A Regulatory Perspective
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CL、cs.AI、cs.CY
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CL、cs.AI、cs.CY
机构 * Intelligent Maintenance and Operations Systems Lab, EPFL, Lausanne, Switzerland(智能维护与运营系统实验室,EPFL,拉沃斯纳,瑞士)
专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG
机构 * Department of Civil, Construction, and Environmental Engineering, San Diego State University, San Diego, CA, United States(土木、建设与环境工程系,圣地亚哥州立大学) ; Department of Electrical and Computer Engineering, University of California, San Diego, San Diego, CA, United States(电气与计算机工程系,加州大学圣地亚哥分校) ; Department of Mechanical and Aerospace Engineering, University of California, San Diego, San Diego, CA, United States(机械与航空航天工程系,加州大学圣地亚哥分校) ; Department of Computer Science, San Diego State University, San Diego, CA, United States(计算机科学系,圣地亚哥州立大学)
专题命中 其他安全 :safety(title,abstract);AI safety(title);alignment(abstract)
机构 * Department of Cognitive Science University of California, San Diego(认知科学系,加州大学圣地亚哥分校)
专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.AI
Comments Accepted at EMNLP 2025 (camera-ready)
机构 * Center on Frontiers of Computing Studies, School of Compter Science, Peking University(前沿计算研究中心,计算机科学学院,北京大学) ; Eastern Institute of Technology, Ningbo(宁波技术研究所) ; Qualcomm AI Research(高通人工智能研究) ; Inst. for Artificial Intelligence, Peking University(人工智能研究所,北京大学)
专题命中 其他安全 :alignment(title,abstract);分类 cs.AI
Comments ICCV 2025
机构 * Department of Artificial Intelligence, Yonsei University(人工智能系,延世大学) ; MAAP LAB, MODULABS(MODULABS 音频实验室) ; KRAFTON ; AI Matics ; Department of Media Software, Sungkyul University(媒体软件系,松谷大学)
专题命中 其他安全 :alignment(title,abstract)
Comments NeurIPS 2025 AI for Music Workshop
机构 * Brown University(布朗大学) ; Rochester Institute of Technology(罗切斯特理工大学)
专题命中 其他安全 :alignment(title);分类 cs.AI
Comments 7 pages, neurips workshop
专题命中 其他安全 :safety(abstract);分类 cs.AI
Comments 17 pages, 12 figures
专题命中 其他安全 :alignment(abstract)
机构 * Mortimer B. Zuckerman Mind Brain Behavior Institute(莫蒂默·B·齐克曼脑行为研究所) ; Columbia University(哥伦比亚大学) ; Microsoft Research(微软研究院)
专题命中 其他安全 :alignment(abstract)
Comments 19 pages, 5 figures