arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2306.04707 2023-06-09 cs.CL cs.AI 62%

Improving Open Language Models by Learning from Organic Interactions

Jing Xu, Da Ju, Joshua Lane, Mojtaba Komeili, Eric Michael Smith, Megan Ung, Morteza Behrooz, William Ngan, Rashel Moritz, Sainbayar Sukhbaatar, Y-Lan Boureau, Jason Weston, Kurt Shuster

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.00301 2023-06-07 cs.LG cs.CL 62%

CapText: Large Language Model-based Caption Generation From Image Context and Description

Shinjini Ghosh, Sagnik Anupam

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments Update 6/6/23: Fixed typographic error in abstract

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16806 2023-06-07 cs.CL cs.AI 62%

Do GPTs Produce Less Literal Translations?

Vikas Raunak, Arul Menezes, Matt Post, Hany Hassan Awadalla

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments ACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01282 2023-06-05 cs.LG cs.AI 62%

Recent Advances in Graph-based Machine Learning for Applications in Smart Urban Transportation Systems

Hongde Wu, Sen Yan, Mingming Liu

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15614 2023-06-01 cs.LG cs.AI 62%

Reverse Engineering Self-Supervised Learning

Ido Ben-Shaul, Ravid Shwartz-Ziv, Tomer Galanti, Shai Dekel, Yann LeCun

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11846 2023-05-22 cs.CV cs.CL cs.LG cs.SD eess.AS 62%

Any-to-Any Generation via Composable Diffusion

Zineng Tang, Ziyi Yang, Chenguang Zhu, Michael Zeng, Mohit Bansal

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments Project Page: https://codi-gen.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.06358 2023-05-12 cs.AI cs.CL 62%

Accessible Instruction-Following Agent

Kairui Zhou

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.02469 2023-05-05 cs.HC cs.AI cs.LG 62%

The System Model and the User Model: Exploring AI Dashboard Design

Fernanda Viégas, Martin Wattenberg

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments 10 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.12328 2023-04-26 q-bio.GN cs.AI cs.LG 62%

Virus2Vec: Viral Sequence Classification Using Machine Learning

Sarwan Ali, Babatunde Bello, Prakash Chourasia, Ria Thazhe Punathil, Pin-Yu Chen, Imdad Ullah Khan, Murray Patterson

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments 11 Pages 6 Figures Accepted in conference Conference on Health, Inference, and Learning (CHIL) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.11507 2023-04-25 cs.LG cs.AI 62%

Machine learning framework for end-to-end implementation of Incident duration prediction

Smrithi Ajit, Varsha R Mouli, Skylar Knickerbocker, Jonathan S. Wood

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.08991 2023-04-19 cs.CL cs.AI 62%

D2CSE: Difference-aware Deep continuous prompts for Contrastive Sentence Embeddings

Hyunjae Lee

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.05839 2023-04-13 cs.LG cs.AI 62%

Optimal Interpretability-Performance Trade-off of Classification Trees with Black-Box Reinforcement Learning

Hector Kohler, Riad Akrour, Philippe Preux

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.05711 2023-04-06 cs.CV cs.CL cs.LG 62%

Synopses of Movie Narratives: a Video-Language Dataset for Story Understanding

Yidan Sun, Qin Chao, Yangfeng Ji, Boyang Li

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments 25 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.02131 2023-03-16 cs.CV cs.CL cs.LG 62%

Masked Vision and Language Modeling for Multi-modal Representation Learning

Gukyeong Kwon, Zhaowei Cai, Avinash Ravichandran, Erhan Bas, Rahul Bhotika, Stefano Soatto

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments International Conference on Learning Representations (ICLR) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.07586 2023-03-15 cs.AI cs.LG 62%

Teacher-Student Knowledge Distillation for Radar Perception on Embedded Accelerators

Steven Shaw, Kanishka Tyagi, Shan Zhang

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments submitted at ASILOMAR,2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.06151 2023-03-14 cs.LG cs.AI 62%

NoiseCAM: Explainable AI for the Boundary Between Noise and Adversarial Attacks

Wenkai Tan, Justus Renkhoff, Alvaro Velasquez, Ziyu Wang, Lusi Li, Jian Wang, Shuteng Niu, Fan Yang, Yongxin Liu, Houbing Song

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments Submitted to IEEE Fuzzy 2023. arXiv admin note: text overlap with arXiv:2303.06032

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.12023 2023-03-14 cs.LG cs.AI cs.CV stat.ML 62%

Generative Modeling Helps Weak Supervision (and Vice Versa)

Benedikt Boecking, Nicholas Roberts, Willie Neiswanger, Stefano Ermon, Frederic Sala, Artur Dubrawski

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments Published as a conference paper at ICLR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.03140 2023-03-07 cs.CR cs.AI cs.CY 62%

Cybersecurity of AI medical devices: risks, legislation, and challenges

Elisabetta Biasin, Erik Kamenjasevic, Kaspar Rosager Ludvigsen

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.02995 2023-03-07 cs.CV cs.CL cs.LG 62%

HiCLIP: Contrastive Language-Image Pretraining with Hierarchy-aware Attention

Shijie Geng, Jianbo Yuan, Yu Tian, Yuxiao Chen, Yongfeng Zhang

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments Accepted at ICLR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.10983 2023-02-23 cs.SD cs.CL cs.LG eess.AS 62%

Do Orcas Have Semantic Language? Machine Learning to Predict Orca Behaviors Using Partially Labeled Vocalization Data

Sophia Sandholm

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.02574 2023-02-21 cs.CL cs.LG 62%

Contextual Semantic Parsing for Multilingual Task-Oriented Dialogues

Mehrad Moradshahi, Victoria Tsai, Giovanni Campagna, Monica S. Lam

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments Published in EACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.13143 2023-02-15 cs.AI cs.LG cs.SY eess.SY 62%

CoTV: Cooperative Control for Traffic Light Signals and Connected Autonomous Vehicles using Deep Reinforcement Learning

Jiaying Guo, Long Cheng, Shen Wang

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.04215 2023-02-09 eess.AS cs.AI cs.LG cs.SD eess.SP 62%

A Vector Quantized Approach for Text to Speech Synthesis on Real-World Spontaneous Speech

Li-Wei Chen, Shinji Watanabe, Alexander Rudnicky

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments Accepted to AAAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.11407 2023-01-18 cs.LG cs.AI stat.ML 62%

Toward Explainable AI for Regression Models

Simon Letzgus, Patrick Wagner, Jonas Lederer, Wojciech Samek, Klaus-Robert Müller, Gregoire Montavon

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments 17 pages, 10 figures, published; changes: 1. references to code and xai-regression.org added (p. 1/2, end of introduction), 2. adjustment of sign-error in restructuring section (p. 8, just above Fig. 4)

Journal ref IEEE Signal Processing Magazine (Volume: 39, Issue: 4, July 2022) 40-58

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.03134 2023-01-12 cs.LG cs.AI eess.SP 62%

A Semi-supervised Approach for Activity Recognition from Indoor Trajectory Data

Mashud Rana, Ashfaqur Rahman, Daniel Smith

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.08567 2022-12-19 cs.LG cs.AI cs.CR cs.LO 62%

Optimized Symbolic Interval Propagation for Neural Network Verification

Philipp Kern, Marko Kleine Büning, Carsten Sinz

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments Published at the 1st Workshop on Formal Verification of Machine Learning (WFVML 2022) (https://www.ml-verification.com/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.01478 2022-12-16 cs.CY cs.LG 62%

A machine learning model to identify corruption in México's public procurement contracts

Andrés Aldana, Andrea Falcón-Cortés, Hernán Larralde

专题命中 其他安全 :safety(abstract);分类 cs.CY、cs.LG

Comments 17 pages, 8 figures. On revision in Government Information Quarterly

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.06576 2022-12-14 cs.LG cs.AI cs.CR cs.CV 62%

AI Model Utilization Measurements For Finding Class Encoding Patterns

Peter Bajcsy, Antonio Cardone, Chenyi Ling, Philippe Dessauw, Michael Majurski, Tim Blattner, Derek Juba, Walid Keyrouz

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments 45 pages, 29 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.05885 2022-12-13 cs.CV cs.AI cs.LG 62%

Image-based Artificial Intelligence empowered surrogate model and shape morpher for real-time blank shape optimisation in the hot stamping process

Haosu Zhou, Nan Li

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments 32 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.03084 2022-12-07 cs.CV cs.AI cs.LG 62%

Land Use Prediction using Electro-Optical to SAR Few-Shot Transfer Learning

Marcel Hussing, Karen Li, Eric Eaton

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments Published at Tackling Climate Change with Machine Learning workshop at NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏