arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-08-15 至 2025-08-15 共收录 3 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 幻觉与事实性 3 篇

2508.10010 2025-08-15 cs.CL 57%

An Audit and Analysis of LLM-Assisted Health Misinformation Jailbreaks Against LLMs

Ayana Hussain, Patrick Zhao, Nicholas Vincent

专题命中 幻觉与事实性 :jailbreak(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09458 2025-08-15 cs.HC cs.AI cs.ET 57%

Hallucination vs interpretation: rethinking accuracy and precision in AI-assisted data extraction for knowledge synthesis

Xi Long, Christy Boscardin, Lauren A. Maggio, Joseph A. Costello, Ralph Gonzales, Rasmyah Hammoudeh, Ki Lai, Yoon Soo Park, Brian C. Gin

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06776 2025-08-15 cs.RO cs.SY eess.SY 50%

Chance-constrained Linear Quadratic Gaussian Games for Multi-robot Interaction under Uncertainty

Kai Ren, Giulio Salizzoni, Mustafa Emre Gürsoy, Maryam Kamgarpour

机构 * SYCAMORE Lab, École Polytechnique Fédérale de Lausanne (EPFL)(SYCAMORE实验室,瑞士联邦理工学院(EPFL))

专题命中 幻觉与事实性 :safety(abstract)

Comments Published in IEEE Control Systems Letters

Journal ref IEEE Control Systems Letters, vol. 9, pp. 2061-2066, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏