arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-12-11 至 2025-12-11 共收录 35 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8 篇

2107.05664 2025-12-11 cs.RO cs.AI 57%

Altruistic Maneuver Planning for Cooperative Autonomous Vehicles Using Multi-agent Advantage Actor-Critic

为合作自主车辆的利他性动作规划使用多智能体优势Actor-Critic

Behrad Toghi, Rodolfo Valiente, Dorsa Sadigh, Ramtin Pedarsani, Yaser P. Fallah

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文提出一种多智能体优势Actor-Critic算法,用于自动驾驶车辆在混合交通环境中的利他性动作规划,以提升交通效率与安全。

Comments Accepted to 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2021) - Workshop on Autonomous Driving: Perception, Prediction and Planning

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09627 2025-12-11 cs.SE 50%

LogICL: Distilling LLM Reasoning to Bridge the Semantic Gap in Cross-Domain Log Anomaly Detection

LogICL: 通过蒸馏大语言模型推理来弥合跨域日志异常检测中的语义鸿沟

Jingwei Ye, Zhi Wang, Chenbin Su, Jieshuai Yang, Jiayi Ding, Chunbo Liu, Ge Chu

专题命中 其他安全 :alignment(abstract)

AI总结 LogICL通过蒸馏大语言模型推理,提升跨域日志异常检测的语义理解与泛化能力,实现更准确的检测与解释。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09573 2025-12-11 cs.CV 50%

Investigate the Low-level Visual Perception in Vision-Language based Image Quality Assessment

研究基于视觉-语言的图像质量评估中的低级视觉感知

Yuan Li, Zitang Sun, Yen-Ju Chen, Shin'ya Nishida

机构 * Graduate School of Informatics, Kyoto University(京都大学信息科学研究生院)

专题命中 其他安全 :alignment(abstract)

AI总结 本文研究了基于视觉-语言模型的图像质量评估中低级视觉感知的问题,发现现有模型在检测基本失真时存在偏差,并通过改进视觉编码器对齐度提升失真识别准确率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09473 2025-12-11 cs.HC 50%

An Efficient Interaction Human-AI Synergy System Bridging Visual Awareness and Large Language Model for Intensive Care Units

一种高效的交互人机协同系统:连接视觉意识与大语言模型用于重症监护室

Yibowen Zhao, Yiming Cao, Zhiqi Shen, Juan Du, Yonghui Xu, Lizhen Cui, Cyril Leung

专题命中 其他安全 :safety(abstract)

AI总结 本文提出一种基于云-边-端架构的人机协同系统,通过视觉感知数据提取和大语言模型驱动的语义交互,提升ICU中患者数据处理的效率和安全性。

Comments This paper has been accepted by the Late Breaking Papers of the 2025 International Conference on Human Computer Interaction (HCII 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11436 2025-12-11 cs.CV 50%

TAViS: Text-bridged Audio-Visual Segmentation with Foundation Models

TAViS: 基于文本的音频视觉分割与基础模型

Ziyang Luo, Nian Liu, Xuguang Yang, Salman Khan, Rao Muhammad Anwer, Hisham Cholakkal, Fahad Shahbaz Khan, Junwei Han

机构 * Northwestern Polytechnical University(西北工业大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)

专题命中 其他安全 :alignment(abstract)

AI总结 TAViS通过文本桥接机制结合多模态基础模型与分割模型,实现高效的音频视觉分割与跨模态对齐。

Comments ICCV2025,code:https://github.com/Sssssuperior/TAViS

详情

展开后加载摘要…

URL PDF HTML 收藏