arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 3317 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 安全训练 3317 篇

2406.01722 2024-06-12 econ.GN q-fin.EC 50%

On Labs and Fabs: Mapping How Alliances, Acquisitions, and Antitrust are Shaping the Frontier AI Industry

Tomás Aguirre

专题命中 安全训练 :safety(abstract)

Comments This is a preliminary version of the paper. Comments are welcome. E-mail: t6aguirre@gmail.com

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04159 2024-06-07 cs.RO cs.MA 50%

MARLander: A Local Path Planning for Drone Swarms using Multiagent Deep Reinforcement Learning

Demetros Aschu, Robinroy Peter, Sausar Karaf, Aleksey Fedoseev, Dzmitry Tsetserukou

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03223 2024-06-06 cs.RO 50%

Object Manipulation in Marine Environments using Reinforcement Learning

Ahmed Nader, Muhayy Ud Din, Mughni Irfan, Irfan Hussain

专题命中 安全训练 :safety(abstract)

Comments 8 pages

Journal ref 15th IFAC Conference on Control Applications in Marine Systems, Robotics and Vehicles (CAMS 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.14854 2024-06-06 cs.MA 50%

MatrixWorld: A pursuit-evasion platform for safe multi-agent coordination and autocurricula

Lijun Sun, Yu-Cheng Chang, Chao Lyu, Chin-Teng Lin, Yuhui Shi

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01449 2024-06-04 cs.CV 50%

SLANT: Spurious Logo ANalysis Toolkit

Maan Qraitem, Piotr Teterwak, Kate Saenko, Bryan A. Plummer

专题命中 安全训练 :harmlessness(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17225 2024-05-28 econ.EM 50%

Quantifying the Reliance of Black-Box Decision-Makers on Variables of Interest

Daniel Vebman

专题命中 安全训练 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.18852 2024-05-28 cs.PL cs.SE 50%

VERT: Verified Equivalent Rust Transpilation with Large Language Models as Few-Shot Learners

Aidan Z. H. Yang, Yoshiki Takashima, Brandon Paulsen, Josiah Dodds, Daniel Kroening

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11391 2024-05-21 eess.SY cs.SY 50%

Optimal control barrier functions for RL based safe powertrain control

Habtamu Hailemichael, Beshah Ayalew, Andrej Ivanco

专题命中 安全训练 :safety(abstract)

Journal ref IFAC-PapersOnLine Volume 56, Issue 3, 2023, Pages 385-390

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.15696 2024-05-14 cs.RO 50%

Delay-Aware Multi-Agent Reinforcement Learning for Cooperative Adaptive Cruise Control with Model-based Stability Enhancement

Jiaqi Liu, Ziran Wang, Peng Hang, Jian Sun

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.06195 2024-05-03 eess.SY cs.SY 50%

Distributed Safe Navigation of Multi-Agent Systems using Control Barrier Function-Based Optimal Controllers

Pol Mestres, Carlos Nieto-Granda, Jorge Cortés

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10690 2024-05-01 cs.RO 50%

Bridging Intelligence and Instinct: A New Control Paradigm for Autonomous Robots

Shimian Zhang, Qiuhong Lu

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.14218 2024-04-23 cs.HC 50%

Designing Safe and Engaging AI Experiences for Children: Towards the Definition of Best Practices in UI/UX Design

Grazia Ragone, Paolo Buono, Rosa Lanzilotti

专题命中 安全训练 :safety(abstract)

Comments 4 pages, The paper has been peer-reviewed and presented at the "CHI 2024 Workshop on Child-centred AI Design", May 11, 2024, Honolulu, HI, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.13127 2024-04-23 cs.SI physics.soc-ph 50%

Uncovering large inconsistencies between machine learning derived gridded settlement datasets

Vedran Sekara, Andrea Martini, Manuel Garcia-Herranz, Do-Hyung Kim

专题命中 安全训练 :alignment(abstract)

Comments 14 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09818 2024-04-16 cs.AR 50%

Error Detection and Correction Codes for Safe In-Memory Computations

Luca Parrini, Taha Soliman, Benjamin Hettwer, Jan Micha Borrmann, Simranjeet Singh, Ankit Bende, Vikas Rana, Farhad Merchant, Norbert Wehn

专题命中 安全训练 :safety(abstract)

Comments This paper will be presented at 29th IEEE European Test Symposium 2024 (ETS) 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.03079 2024-04-05 cs.DC 50%

vPALs: Towards Verified Performance-aware Learning System For Resource Management

Guoliang He, Gingfung Yeung, Sheriffo Ceesay, Adam Barker

专题命中 安全训练 :safety(abstract)

Comments presented at Deployable AI Workshop at AAAI-2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.14348 2024-04-05 cs.MA stat.ML 50%

DePAint: A Decentralized Safe Multi-Agent Reinforcement Learning Algorithm considering Peak and Average Constraints

Raheeb Hassan, K. M. Shadman Wadith, Md. Mamun or Rashid, Md. Mosaddek Khan

专题命中 安全训练 :safety(abstract)

Comments accepted for publication in Springer Applied Intelligence Journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.11156 2024-04-03 math.OC cs.MA cs.RO cs.SY eess.SY math.DS 50%

Collaborative Safe Formation Control for Coupled Multi-Agent Systems

Brooks A. Butler, Chi Ho Leung, Philip E. Paré

专题命中 安全训练 :safety(abstract)

Comments This work has been accepted to be presented at the 2024 European Control Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.20218 2024-04-01 cs.CR 50%

Decentralized Multimedia Data Sharing in IoV: A Learning-based Equilibrium of Supply and Demand

Jiani Fan, Minrui Xu, Jiale Guo, Lwin Khin Shar, Jiawen Kang, Dusit Niyato, Kwok-Yan Lam

专题命中 安全训练 :safety(abstract)

Journal ref IEEE Transactions on Vehicular Technology (Volume: 73, Issue: 3, March 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.19492 2024-03-29 cs.CV 50%

Segmentation tool for images of cracks

Andrii Kompanets, Remco Duits, Davide Leonetti, Nicky van den Berg, H. H., Snijder

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18996 2024-03-29 cs.CV 50%

Envisioning MedCLIP: A Deep Dive into Explainability for Medical Vision-Language Models

Anees Ur Rehman Hashmi, Dwarikanath Mahapatra, Mohammad Yaqub

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14369 2024-03-22 eess.SY cs.SY 50%

A Control Barrier Function Composition Approach for Multi-Agent Systems in Marine Applications

Yujia Yang, Chris Manzie, Ye Pu

专题命中 安全训练 :safety(abstract)

Comments 11 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10083 2024-03-18 cs.RO 50%

HeR-DRL:Heterogeneous Relational Deep Reinforcement Learning for Decentralized Multi-Robot Crowd Navigation

Xinyu Zhou, Songhao Piao, Wenzheng Chi, Liguo Chen, Wei Li

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.02345 2024-02-20 eess.SY cs.SY 50%

Communication-Efficient Decentralized Multi-Agent Reinforcement Learning for Cooperative Adaptive Cruise Control

Dong Chen, Kaixiang Zhang, Yongqiang Wang, Xunyuan Yin, Zhaojian Li, Dimitar Filev

专题命中 安全训练 :safety(abstract)

Comments 14 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.11035 2024-01-23 cs.CV 50%

Image Safeguarding: Reasoning with Conditional Vision Language Model and Obfuscating Unsafe Content Counterfactually

Mazal Bethany, Brandon Wherry, Nishant Vishwamitra, Peyman Najafirad

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.06226 2024-01-15 cs.RO 50%

Learning Crowd Behaviors in Navigation with Attention-based Spatial-Temporal Graphs

Yanying Zhou, Jochen Garcke

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.02001 2024-01-05 cs.SI 50%

Close to Human-Level Agreement: Tracing Journeys of Violent Speech in Incel Posts with GPT-4-Enhanced Annotations

Daniel Matter, Miriam Schirmer, Nir Grinberg, Jürgen Pfeffer

专题命中 安全训练 :alignment(abstract)

Comments 10 pages, 2 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.12966 2023-12-25 math.OC cs.RO 50%

Rate-Tunable Control Barrier Functions: Methods and Algorithms for Online Adaptation

Hardik Parwana, Dimitra Panagou

专题命中 安全训练 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01059 2023-12-05 cs.RO 50%

Swarm-GPT: Combining Large Language Models with Safe Motion Planning for Robot Choreography Design

Aoran Jiao, Tanmay P. Patel, Sanjmi Khurana, Anna-Mariya Korol, Lukas Brunke, Vivek K. Adajania, Utku Culha, Siqi Zhou, Angela P. Schoellig

专题命中 安全训练 :safety(abstract)

Comments 10 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13714 2023-11-27 cs.RO cs.MA cs.SY eess.SY math.OC 50%

Learning Safe Control for Multi-Robot Systems: Methods, Verification, and Open Challenges

Kunal Garg, Songyuan Zhang, Oswin So, Charles Dawson, Chuchu Fan

专题命中 安全训练 :safety(abstract)

Comments Submitted to Annual Reviews in Control

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.10041 2023-11-17 cs.RO 50%

Interpretable Reinforcement Learning for Robotics and Continuous Control

Rohan Paleja, Letian Chen, Yaru Niu, Andrew Silva, Zhaoxin Li, Songan Zhang, Chace Ritchie, Sugju Choi, Kimberlee Chestnut Chang, Hongtei Eric Tseng, Yan Wang, Subramanya Nageshrao, Matthew Gombolay

专题命中 安全训练 :safety(abstract)

Comments arXiv admin note: text overlap with arXiv:2202.02352

详情

展开后加载摘要…

URL PDF HTML 收藏