arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-07-24 至 2025-07-24 共收录 32 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 10 篇

2507.17165 2025-07-24 cs.SE 50%

Can LLMs Write CI? A Study on Automatic Generation of GitHub Actions Configurations

Taher A. Ghaleb, Dulina Rathnayake

专题命中 其他安全 :alignment(abstract)

Comments Accepted at the 41st IEEE International Conference on Software Maintenance and Evolution 2025 (ICSME'25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17050 2025-07-24 cs.CV 50%

Toward Scalable Video Narration: A Training-free Approach Using Multimodal Large Language Models

Tz-Ying Wu, Tahani Trigui, Sharath Nittur Sridhar, Anand Bodas, Subarna Tripathi

机构 * Intel(英特尔)

专题命中 其他安全 :alignment(abstract)

Comments Accepted to CVAM Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏