arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-09-18 至 2025-09-18 共收录 38 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 安全评测 17 篇

2509.13731 2025-09-18 cs.RO 50%

Reinforcement Learning for Robotic Insertion of Flexible Cables in Industrial Settings

Jeongwoo Park, Seabin Lee, Changmin Park, Wonjong Lee, Changjoo Nam

机构 * Dept. of Electronic Engineering, Sogang University(电子工程系,成均馆大学) Dept. of Artificial Intelligence, Sogang University(人工智能系,成均馆大学)

专题命中 安全评测 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13629 2025-09-18 cs.CV 50%

SAMIR, an efficient registration framework via robust feature learning from SAM

Yue He, Min Liu, Qinghao Liu, Jiazheng Wang, Yaonan Wang, Hang Zhang, Xiang Chen

专题命中 安全评测 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10426 2025-09-18 cs.RO cs.MA 50%

DECAMP: Towards Scene-Consistent Multi-Agent Motion Prediction with Disentangled Context-Aware Pre-Training

Jianxin Shi, Zengqi Peng, Xiaolong Chen, Tianyu Wo, Jun Ma

机构 * School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院) Robotics and Autonomous Systems Thrust, The Hong Kong University of Science and Technology (GZ)(香港科技大学(广州)机器人与自主系统方向) Division of Emerging Interdisciplinary Areas, The Hong Kong University of Science and Technology(香港科技大学新兴交叉领域研究所)

专题命中 安全评测 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 其他安全 5 篇

2509.13706 2025-09-18 cs.CL cs.AI 81%

Automated Triaging and Transfer Learning of Incident Learning Safety Reports Using Large Language Representational Models

Peter Beidler, Mark Nguyen, Kevin Lybarger, Ola Holmberg, Eric Ford, John Kang

专题命中 其他安全 :safety(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13846 2025-09-18 cs.CV cs.LG 79%

Consistent View Alignment Improves Foundation Models for 3D Medical Image Segmentation

Puru Vaish, Felix Meister, Tobias Heimann, Christoph Brune, Jelmer M. Wolterink

机构 * Department of Applied Mathematics, Technical Medical Centre, University of Twente(代尔夫特理工大学应用数学系) Digital Technology and Innovation, Siemens Healthineers, Erlangen, Germany(西门子医疗创新部,埃尔朗根,德国)

专题命中 其他安全 :alignment(title,abstract);分类 cs.LG

Comments MICCAI 2025: 1st Place in Transformer track and 2nd Place in Convolution track of SSL3D-OpenMind challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13839 2025-09-18 cs.RO 78%

Pre-Manipulation Alignment Prediction with Parallel Deep State-Space and Transformer Models

Motonari Kambara, Komei Sugiura

机构 * Keio University(keio大学)

专题命中 其他安全 :alignment(title,abstract)

Comments Published in Advanced Robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20030 2025-09-18 cs.GT cs.DS 78%

Polynomial-Time Approximation Schemes via Utility Alignment: Unit-Demand Pricing and More

Robin Bowers, Marius Garbea, Emmanouil Pountourakis, Samuel Taggart

专题命中 其他安全 :alignment(title,abstract)

Comments To appear in FOCS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13704 2025-09-18 cs.AI cs.SE 57%

InfraMind: A Novel Exploration-based GUI Agentic Framework for Mission-critical Industrial Management

Liangtao Lin, Zhaomeng Zhu, Tianwei Zhang, Yonggang Wen

机构 * Nanyang Technological University(南洋理工大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏