arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2404.10237 2024-09-04 cs.CV cs.CL 57%

Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models

Songtao Jiang, Tuo Zheng, Yan Zhang, Yeying Jin, Li Yuan, Zuozhu Liu

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16750 2024-09-04 cs.AI cs.AR 57%

All Artificial, Less Intelligence: GenAI through the Lens of Formal Verification

Deepak Narayan Gadde, Aman Kumar, Thomas Nalapat, Evgenii Rezunov, Fabio Cappellini

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments Published in DVCon U.S. 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.00005 2024-09-04 cs.IT cs.AI math.IT 57%

Csi-LLM: A Novel Downlink Channel Prediction Method Aligned with LLM Pre-Training

Shilong Fan, Zhenyu Liu, Xinyu Gu, Haozhen Li

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.17471 2024-08-31 cs.CV cs.CY 57%

Real-Time Automated donning and doffing detection of PPE based on Yolov4-tiny

Anusha Verma, Ghazal Ghajari, K M Tawsik Jawad, Hugh P. Salehi, Fathi Amsaad

专题命中 其他安全 :safety(abstract);分类 cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11911 2024-08-30 cs.CL 57%

InstructERC: Reforming Emotion Recognition in Conversation with Multi-task Retrieval-Augmented Large Language Models

Shanglin Lei, Guanting Dong, Xiaoping Wang, Keheng Wang, Runqi Qiao, Sirui Wang

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.15245 2024-08-29 cs.CV cs.AI 57%

An Edge AI System Based on FPGA Platform for Railway Fault Detection

Jiale Li, Yulin Fu, Dongwei Yan, Sean Longyu Ma, Chiu-Wing Sham

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments Accepted at the 2024 IEEE 13th Global Conference on Consumer Electronics (GCCE 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.14972 2024-08-28 cs.CL 57%

AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems

Chi-Min Chan, Jianxuan Yu, Weize Chen, Chunyang Jiang, Xinyu Liu, Weijie Shi, Zhiyuan Liu, Wei Xue, Yike Guo

专题命中 其他安全 :safety(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.14023 2024-08-27 cs.CV cs.AI 57%

Video-CCAM: Enhancing Video-Language Understanding with Causal Cross-Attention Masks for Short and Long Videos

Jiajun Fei, Dian Li, Zhidong Deng, Zekun Wang, Gang Liu, Hui Wang

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.14008 2024-08-27 cs.CV cs.AI 57%

LMM-VQA: Advancing Video Quality Assessment with Large Multimodal Models

Qihang Ge, Wei Sun, Yu Zhang, Yunhao Li, Zhongpeng Ji, Fengyu Sun, Shangling Jui, Xiongkuo Min, Guangtao Zhai

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.07729 2024-08-27 cs.CV cs.AI cs.RO 57%

SSL-Interactions: Pretext Tasks for Interactive Trajectory Prediction

Prarthana Bhattacharyya, Chengjie Huang, Krzysztof Czarnecki

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments Accepted at IV-2024. 13 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.12609 2024-08-26 cs.RO cs.AI 57%

Enhanced Prediction of Multi-Agent Trajectories via Control Inference and State-Space Dynamics

Yu Zhang, Yongxiang Zou, Haoyu Zhang, Zeyu Liu, Houcheng Li, Long Cheng

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.12194 2024-08-26 cs.CL 57%

Large Language Models as Foundations for Next-Gen Dense Retrieval: A Comprehensive Empirical Assessment

Kun Luo, Minghao Qin, Zheng Liu, Shitao Xiao, Jun Zhao, Kang Liu

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments Submitted to EMNLP24

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09717 2024-08-22 cs.CL 57%

UniBridge: A Unified Approach to Cross-Lingual Transfer Learning for Low-Resource Languages

Trinh Pham, Khoi M. Le, Luu Anh Tuan

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments First two authors contribute equally. Accepted at ACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.16769 2024-08-22 eess.SY cs.AI cs.SY 57%

2-Level Reinforcement Learning for Ships on Inland Waterways: Path Planning and Following

Martin Waltz, Niklas Paulig, Ostap Okhrin

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.10086 2024-08-20 cs.AI 57%

ARMADA: Attribute-Based Multimodal Data Augmentation

Xiaomeng Jin, Jeonghwan Kim, Yu Zhou, Kuan-Hao Huang, Te-Lin Wu, Nanyun Peng, Heng Ji

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.09682 2024-08-20 cs.AI 57%

Simulating Field Experiments with Large Language Models

Yaoyu Chen, Yuheng Hu, Yingda Lu

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 17 pages, 5 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.09530 2024-08-20 cs.AI 57%

PA-LLaVA: A Large Language-Vision Assistant for Human Pathology Image Understanding

Dawei Dai, Yuanhui Zhang, Long Xu, Qianlan Yang, Xiaojing Shen, Shuyin Xia, Guoyin Wang

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 8 pages, 4 figs

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05709 2024-08-19 cs.RO cs.FL cs.LG 57%

TR2MTL: LLM based framework for Metric Temporal Logic Formalization of Traffic Rules

Kumar Manas, Stefan Zwicklbauer, Adrian Paschke

专题命中 其他安全 :safety(abstract);分类 cs.LG

Comments Accepted for publication in Proceedings of the IEEE Intelligent Vehicles Symposium (IV), Jeju Island - Korea, 2-5 June 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.07990 2024-08-16 cs.CL 57%

FuseChat: Knowledge Fusion of Chat Models

Fanqi Wan, Longguang Zhong, Ziyi Yang, Ruijun Chen, Xiaojun Quan

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.07526 2024-08-15 cs.SE cs.CR cs.LG 57%

Learning-based Models for Vulnerability Detection: An Extensive Study

Chao Ni, Liyu Shen, Xiaodan Xu, Xin Yin, Shaohua Wang

专题命中 其他安全 :safety(abstract);分类 cs.LG

Comments 13 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09442 2024-08-15 physics.med-ph cs.LG physics.app-ph physics.bio-ph 57%

An insertable glucose sensor using a compact and cost-effective phosphorescence lifetime imager and machine learning

Artem Goncharov, Zoltan Gorocs, Ridhi Pradhan, Brian Ko, Ajmal Ajmal, Andres Rodriguez, David Baum, Marcell Veszpremi, Xilin Yang, Maxime Pindrys, Tianle Zheng, Oliver Wang, Jessica C. Ramella-Roman, Michael J. McShane, Aydogan Ozcan

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments 24 Pages, 4 Figures

Journal ref ACS Nano (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06929 2024-08-14 cs.CL 57%

Evaluating Cultural Adaptability of a Large Language Model via Simulation of Synthetic Personas

Louis Kwok, Michal Bravansky, Lewis D. Griffin

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments 18 pages, 8 figures, Published as a conference paper at COLM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.07821 2024-08-14 cs.LG stat.ML 57%

When to Accept Automated Predictions and When to Defer to Human Judgment?

Daniel Sikar, Artur Garcez, Tillman Weyde, Robin Bloomfield, Kaleem Peeroo

专题命中 其他安全 :safety(abstract);分类 cs.LG

Comments 9 pages, 10 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06332 2024-08-13 cs.CL 57%

Animate, or Inanimate, That is the Question for Large Language Models

Leonardo Ranaldi, Giulia Pucci, Fabio Massimo Zanzotto

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.17150 2024-08-12 cs.CL cs.SE 57%

SimCT: A Simple Consistency Test Protocol in LLMs Development Lifecycle

Fufangchen Zhao, Guoqiang Jin, Rui Zhao, Jiangheng Huang, Fei Tan

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00575 2024-08-12 cs.CL 57%

Instruction-tuning Aligns LLMs to the Human Brain

Khai Loong Aw, Syrielle Montariol, Badr AlKhamissi, Martin Schrimpf, Antoine Bosselut

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments COLM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.03528 2024-08-09 cs.SE cs.AI 57%

Exploring the extent of similarities in software failures across industries using LLMs

Martin Detloff

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.03892 2024-08-08 cs.SE cs.AI 57%

MORTAR: A Model-based Runtime Action Repair Framework for AI-enabled Cyber-Physical Systems

Renzhi Wang, Zhehua Zhou, Jiayang Song, Xuan Xie, Xiaofei Xie, Lei Ma

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.15821 2024-08-07 cs.GT cs.AI cs.MA 57%

Cooperation and Control in Delegation Games

Oliver Sourbut, Lewis Hammond, Harriet Wood

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments Published at IJCAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.02751 2024-08-07 cs.LG 57%

A Novel Hybrid Approach for Tornado Prediction in the United States: Kalman-Convolutional BiLSTM with Multi-Head Attention

Jiawei Zhou

专题命中 其他安全 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏