arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8034 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8034 篇

2311.00047 2023-11-02 cs.AI cs.CL cs.CV cs.LG 67%

Grounding Visual Illusions in Language: Do Vision-Language Models Perceive Illusions Like Humans?

Yichi Zhang, Jiayi Pan, Yuchen Zhou, Rui Pan, Joyce Chai

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at EMNLP 2023 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18366 2023-10-31 cs.CL cs.AI cs.LG 67%

A Multilingual Virtual Guide for Self-Attachment Technique

Alicia Jiayun Law, Ruoyu Hu, Lisa Alazraki, Anandha Gopalan, Neophytos Polydorou, Abbas Edalat

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

Journal ref 2022 IEEE 4th International Conference on Cognitive Machine Intelligence (CogMI)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.17218 2023-10-27 cs.CY cs.AI cs.HC cs.LG cs.SY eess.SY 67%

Artificial intelligence in government: Concepts, standards, and a unified framework

Vincent J. Straub, Deborah Morgan, Jonathan Bright, Helen Margetts

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.CY、cs.LG

Comments 35 pages with references and appendix, 3 tables, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.01423 2023-10-24 cs.CL cs.AI cs.LG 67%

An Empirical Study of AI Generated Text Detection Tools

Arslan Akram

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 15 Pages, 4 Figures, 2 Tables, 42 References

Journal ref 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.03135 2023-10-13 cs.CV cs.AI cs.CL cs.LG 67%

Distilling Large Vision-Language Model with Out-of-Distribution Generalizability

Xuanlin Li, Yunhao Fang, Minghua Liu, Zhan Ling, Zhuowen Tu, Hao Su

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Published at International Conference on Computer Vision (ICCV) 2023. Poster at https://xuanlinli17.github.io/pdfs/iccv23_large_vlm_distillation_poster.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03840 2023-10-09 cs.LG cs.AI cs.CL 67%

Contextualized Structural Self-supervised Learning for Ontology Matching

Zhu Wang

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.10891 2023-09-21 cs.CL cs.AI cs.LG 67%

Self-Augmentation Improves Zero-Shot Cross-Lingual Transfer

Fei Wang, Kuan-Hao Huang, Kai-Wei Chang, Muhao Chen

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

Comments AACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14683 2023-08-29 cs.CL cs.AI cs.LG 67%

Fine-Tuning Llama 2 Large Language Models for Detecting Online Sexual Predatory Chats and Abusive Texts

Thanh Thi Nguyen, Campbell Wilson, Janis Dalins

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13563 2023-08-29 cs.CL cs.AI cs.IR cs.LG 67%

Large Language Models in Analyzing Crash Narratives -- A Comparative Study of ChatGPT, BARD and GPT-4

Maroa Mumtarin, Md Samiullah Chowdhury, Jonathan Wood

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09970 2023-08-22 cs.CL cs.AI cs.LG 67%

Tackling Vision Language Tasks Through Learning Inner Monologues

Diji Yang, Kezhen Chen, Jinmeng Rao, Xiaoyuan Guo, Yawen Zhang, Jie Yang, Yi Zhang

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09804 2023-08-22 cs.CV cs.AI cs.CL cs.LG 67%

VL-PET: Vision-and-Language Parameter-Efficient Tuning via Granularity Control

Zi-Yuan Hu, Yanyang Li, Michael R. Lyu, Liwei Wang

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

Comments ICCV 2023 (17 pages, 6 figures, 22 tables)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.07326 2023-08-16 cs.AI cs.CL cs.LG 67%

AI Text-to-Behavior: A Study In Steerability

David Noever, Sam Hyams

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.12128 2023-07-25 cs.CV cs.AI cs.CY cs.LG 67%

AI on the Road: A Comprehensive Analysis of Traffic Accidents and Accident Detection System in Smart Cities

Victor Adewopo, Nelly Elsayed, Zag Elsayed, Murat Ozer, Victoria Wangia-Anderson, Ahmed Abdelgawad

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY、cs.LG

Comments 8,8

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.09782 2023-07-24 cs.LG cs.AI cs.CL 67%

ZeroQuant-FP: A Leap Forward in LLMs Post-Training W4A8 Quantization Using Floating-Point Formats

Xiaoxia Wu, Zhewei Yao, Yuxiong He

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.03419 2023-07-12 cs.CY cs.AI cs.DS cs.LG 67%

QI2 -- an Interactive Tool for Data Quality Assurance

Simon Geerkens, Christian Sieberichs, Alexander Braun, Thomas Waschulzik

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.04114 2023-07-11 cs.LG cs.AI cs.CL cs.CV cs.MM 67%

FILM: How can Few-Shot Image Classification Benefit from Pre-Trained Language Models?

Zihao Jiang, Yunkai Dang, Dong Pang, Huishuai Zhang, Weiran Huang

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.00633 2023-07-04 cs.RO cs.AI cs.CY cs.HC cs.LG 67%

Effects of Explanation Specificity on Passengers in Autonomous Driving

Daniel Omeiza, Raunak Bhattacharyya, Nick Hawes, Marina Jirotka, Lars Kunze

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.04636 2023-06-30 cs.AI cs.CL cs.LG 67%

"That Is a Suspicious Reaction!": Interpreting Logits Variation to Detect NLP Adversarial Attacks

Edoardo Mosca, Shreyash Agarwal, Javier Rando, Georg Groh

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

Comments ACL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.00020 2023-06-02 cs.CL cs.AI cs.LG 67%

GPT4GEO: How a Language Model Sees the World's Geography

Jonathan Roberts, Timo Lüddecke, Sowmen Das, Kai Han, Samuel Albanie

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.12219 2023-05-26 cs.LG cs.AI cs.CL 67%

Collaborative Development of NLP models

Fereshte Khani, Marco Tulio Ribeiro

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.04013 2023-04-04 cs.CL cs.AI cs.LG 67%

There is No Big Brother or Small Brother: Knowledge Infusion in Language Models for Link Prediction and Question Answering

Ankush Agarwal, Sakharam Gawade, Sachin Channabasavarajendra, Pushpak Bhattacharyya

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.01817 2023-03-20 cs.AI cs.CY cs.LG 67%

Liability regimes in the age of AI: a use-case driven analysis of the burden of proof

David Fernández Llorca, Vicky Charisi, Ronan Hamon, Ignacio Sánchez, Emilia Gómez

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY、cs.LG

Comments Paper published at the Journal of Artificial Intelligence Research

Journal ref Journal of Artificial Intelligence Research, Vol. 76 (2023), pp. 613-644

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.03430 2023-02-21 cs.LG cs.AI cs.CL cs.CV cs.MM 67%

Foundations and Trends in Multimodal Machine Learning: Principles, Challenges, and Open Questions

Paul Pu Liang, Amir Zadeh, Louis-Philippe Morency

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.12737 2022-11-24 cs.CV cs.AI cs.CL cs.LG 67%

RoentGen: Vision-Language Foundation Model for Chest X-ray Generation

Pierre Chambon, Christian Bluethgen, Jean-Benoit Delbrouck, Rogier Van der Sluijs, Małgorzata Połacin, Juan Manuel Zambrano Chaves, Tanishq Mathew Abraham, Shivanshu Purohit, Curtis P. Langlotz, Akshay Chaudhari

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 19 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.06318 2022-11-14 cs.CY cs.AI cs.LG 67%

Artificial Intelligence and Life in 2030: The One Hundred Year Study on Artificial Intelligence

Peter Stone, Rodney Brooks, Erik Brynjolfsson, Ryan Calo, Oren Etzioni, Greg Hager, Julia Hirschberg, Shivaram Kalyanakrishnan, Ece Kamar, Sarit Kraus, Kevin Leyton-Brown, David Parkes, William Press, AnnaLee Saxenian, Julie Shah, Milind Tambe, Astro Teller

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY、cs.LG

Comments 52 pages, https://ai100.stanford.edu/2016-report

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.14504 2022-10-11 cs.CL cs.AI cs.LG 67%

GERNERMED++: Transfer Learning in German Medical NLP

Johann Frei, Ludwig Frei-Stuber, Frank Kramer

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.09900 2022-09-21 cs.CL cs.AI cs.LG 67%

LINGUIST: Language Model Instruction Tuning to Generate Annotated Utterances for Intent Classification and Slot Tagging

Andy Rosenbaum, Saleh Soltan, Wael Hamza, Yannick Versley, Markus Boese

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted to The 29th International Conference on Computational Linguistics (COLING 2022) October 12-17, 2022, Gyeongju, Republic of Korea https://coling2022.org/

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10726 2022-09-15 cs.CL cs.AI cs.LG 67%

TWEET-FID: An Annotated Dataset for Multiple Foodborne Illness Detection Tasks

Ruofan Hu, Dongyu Zhang, Dandan Tao, Thomas Hartvigsen, Hao Feng, Elke Rundensteiner

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI、cs.LG

Comments LREC 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.02222 2022-08-04 cs.LG cs.AI cs.CY eess.SP 67%

Blockchain associated machine learning and IoT based hypoglycemia detection system with auto-injection feature

Rahnuma Mahzabin, Fahim Hossain Sifat, Sadia Anjum, Al-Akhir Nayan, Muhammad Golam Kibria

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY、cs.LG

Journal ref Indonesian Journal of Electrical Engineering and Computer Science, Vol. 27, No. 1, pp. 447-455, July 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.10086 2022-06-17 cs.AI cs.CY cs.LG 67%

Learning Models of Individual Behavior in Chess

Reid McIlroy-Young, Russell Wang, Siddhartha Sen, Jon Kleinberg, Ashton Anderson

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.CY、cs.LG

Comments 12 pages, 11 figures, 5 tables, Published in the Proceedings of the 28th ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD 2022), Code https://github.com/CSSLab/maia-individual

详情

展开后加载摘要…

URL PDF HTML 收藏