arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2402.01091 2024-02-05 cs.CL cs.CY cs.SI 62%

Reading Between the Tweets: Deciphering Ideological Stances of Interconnected Mixed-Ideology Communities

Zihao He, Ashwin Rao, Siyi Guo, Negar Mokhberian, Kristina Lerman

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.12241 2024-02-05 cs.CY cs.AI 62%

Positive AI: Key Challenges in Designing Artificial Intelligence for Wellbeing

Willem van der Maden, Derek Lomas, Malak Sadek, Paul Hekkert

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.03674 2024-02-05 cs.LG cs.AI cs.SE 62%

Machine Learning with Requirements: a Manifesto

Eleonora Giunchiglia, Fergus Imrie, Mihaela van der Schaar, Thomas Lukasiewicz

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01053 2024-02-05 cs.CL cs.AI 62%

Plan-Grounded Large Language Models for Dual Goal Conversational Settings

Diogo Glória-Silva, Rafael Ferreira, Diogo Tavares, David Semedo, João Magalhães

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.04645 2024-02-01 q-bio.NC cs.AI cs.CL eess.AS 62%

Do self-supervised speech and language models extract similar representations as human brain?

Peili Chen, Linyang He, Li Fu, Lu Fan, Edward F. Chang, Yuanning Li

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments To appear in 2024 IEEE International Conference on Acoustics, Speech and Signal Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.01295 2024-01-30 cs.CL cs.AI 62%

Efficiently Aligned Cross-Lingual Transfer Learning for Conversational Tasks using Prompt-Tuning

Lifu Tu, Jin Qu, Semih Yavuz, Shafiq Joty, Wenhao Liu, Caiming Xiong, Yingbo Zhou

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments Accepted to the Finding of the ACL: EACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.10213 2024-01-19 cs.CV cs.CY cs.LG 62%

Improving automatic detection of driver fatigue and distraction using machine learning

Dongjiang Wu

专题命中 其他安全 :alignment(abstract);分类 cs.CY、cs.LG

Comments Master's thesis, 55 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04118 2024-01-18 cs.CV cs.AI cs.LG 62%

Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play

Timothy Schaumlöffel, Arthur Aubret, Gemma Roig, Jochen Triesch

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments Proceedings of the 2023 IEEE International Conference on Development and Learning (ICDL)

Journal ref "Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play," 2023 IEEE International Conference on Development and Learning (ICDL), Macau, China, 2023, pp. 67-72

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.08581 2024-01-18 cs.CV cs.AI cs.LG 62%

Temporal Embeddings: Scalable Self-Supervised Temporal Representation Learning from Spatiotemporal Data for Multimodal Computer Vision

Yi Cao, Swetava Ganguli, Vipul Pandey

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments Extended abstract accepted for presentation at BayLearn 2023. 3 pages, 7 figures. Abstract based on IEEE IGARSS 2023 research track paper: arXiv:2304.13143

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.07263 2024-01-17 cs.LG cs.AI 62%

BET: Explaining Deep Reinforcement Learning through The Error-Prone Decisions

Xiao Liu, Jie Zhao, Wubing Chen, Mao Tan, Yongxing Su

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments This is an early version of a paper that submitted to IJCAI 2024 8 pages, 4 figures and 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.07037 2024-01-17 cs.CL cs.AI 62%

xCoT: Cross-lingual Instruction Tuning for Cross-lingual Chain-of-Thought Reasoning

Linzheng Chai, Jian Yang, Tao Sun, Hongcheng Guo, Jiaheng Liu, Bing Wang, Xiannian Liang, Jiaqi Bai, Tongliang Li, Qiyao Peng, Zhoujun Li

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments 11 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.05605 2024-01-12 cs.CL cs.LG 62%

Scaling Laws for Forgetting When Fine-Tuning Large Language Models

Damjan Kalajdzievski

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.12803 2024-01-10 cs.LG cs.CL 62%

Data Augmentations for Improved (Large) Language Model Generalization

Amir Feder, Yoav Wald, Claudia Shi, Suchi Saria, David Blei

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.LG

Comments Published at NeurIPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.12458 2023-12-21 cs.CL cs.AI 62%

When Parameter-efficient Tuning Meets General-purpose Vision-language Models

Yihang Zhai, Haixin Wang, Jianlong Chang, Xinlong Yang, Jinan Sun, Shikun Zhang, Qi Tian

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.06453 2023-12-20 cs.CL cs.LG 62%

Narrowing the Gap between Supervised and Unsupervised Sentence Representation Learning with Large Language Model

Mingxin Li, Richong Zhang, Zhijie Nie, Yongyi Mao

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments Accepted at AAAI24

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10332 2023-12-19 cs.IR cs.AI cs.LG 62%

ProTIP: Progressive Tool Retrieval Improves Planning

Raviteja Anantha, Bortik Bandyopadhyay, Anirudh Kashi, Sayantan Mahinder, Andrew W Hill, Srinivas Chappidi

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments preprint version

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07705 2023-12-14 q-bio.NC cs.AI cs.CV cs.LG 62%

Brain-optimized inference improves reconstructions of fMRI brain activity

Reese Kneeland, Jordyn Ojeda, Ghislain St-Yves, Thomas Naselaris

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments 7 pages, 8 figures, submitted to the 2023 AAAI Workshop on Brain Encoding and Decoding. arXiv admin note: text overlap with arXiv:2306.00927

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.06861 2023-12-13 cs.CY cs.CL 62%

Disentangling Perceptions of Offensiveness: Cultural and Moral Correlates

Aida Davani, Mark Díaz, Dylan Baker, Vinodkumar Prabhakaran

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.05461 2023-12-12 cs.LG cs.AI 62%

STREAMLINE: An Automated Machine Learning Pipeline for Biomedicine Applied to Examine the Utility of Photography-Based Phenotypes for OSA Prediction Across International Sleep Centers

Ryan J. Urbanowicz, Harsh Bandhey, Brendan T. Keenan, Greg Maislin, Sy Hwang, Danielle L. Mowery, Shannon M. Lynch, Diego R. Mazzotti, Fang Han, Qing Yun Li, Thomas Penzel, Sergio Tufik, Lia Bittencourt, Thorarinn Gislason, Philip de Chazal, Bhajan Singh, Nigel McArdle, Ning-Hung Chen, Allan Pack, Richard J. Schwab, Peter A. Cistulli, Ulysses J. Magalang

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments 23 pages, 7 figures, 1 table, 1 supplemental information document (77 pages), and 7 ancillary files

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04917 2023-12-11 cs.SE cs.AI cs.LG 62%

Operationalizing Assurance Cases for Data Scientists: A Showcase of Concepts and Tooling in the Context of Test Data Quality for Machine Learning

Lisa Jöckel, Michael Kläs, Janek Groß, Pascal Gerber, Markus Scholz, Jonathan Eberle, Marc Teschner, Daniel Seifert, Richard Hawkins, John Molloy, Jens Ottnad

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments Accepted for publication at International Conference on Product-Focused Software Process Improvement (Profes 2023), https://conf.researchr.org/home/profes-2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04416 2023-12-08 cs.LG cs.CY 62%

Monitoring Sustainable Global Development Along Shared Socioeconomic Pathways

Michelle W. L. Wan, Jeffrey N. Clark, Edward A. Small, Elena Fillola Mayoral, Raúl Santos-Rodríguez

专题命中 其他安全 :alignment(abstract);分类 cs.CY、cs.LG

Comments 5 pages, 1 figure. Presented at NeurIPS 2023 Workshop: Tackling Climate Change with Machine Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03728 2023-12-08 cs.CL cs.AI 62%

Real Customization or Just Marketing: Are Customized Versions of Chat GPT Useful?

Eduardo C. Garrido-Merchán, Jose L. Arroyo-Barrigüete, Francisco Borrás-Pala, Leandro Escobar-Torres, Carlos Martínez de Ibarreta, Jose María Ortiz-Lozano, Antonio Rua-Vieites

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments 9 pages, 1 figure, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05736 2023-12-07 cs.CL cs.LG 62%

LLMLingua: Compressing Prompts for Accelerated Inference of Large Language Models

Huiqiang Jiang, Qianhui Wu, Chin-Yew Lin, Yuqing Yang, Lili Qiu

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments Accepted at EMNLP 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00194 2023-12-04 cs.LG cs.CL 62%

Robust Concept Erasure via Kernelized Rate-Distortion Maximization

Somnath Basu Roy Chowdhury, Nicholas Monath, Avinava Dubey, Amr Ahmed, Snigdha Chaturvedi

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments NeurIPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17686 2023-11-30 cs.CL cs.AI 62%

AviationGPT: A Large Language Model for the Aviation Domain

Liya Wang, Jason Chou, Xin Zhou, Alex Tien, Diane M Baumgartner

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16143 2023-11-29 cs.CR cs.AI cs.CV cs.LG 62%

Ransomware Detection and Classification using Machine Learning

Kavitha Kunku, ANK Zaman, Kaushik Roy

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Journal ref IEEE Symposium on Computational Intelligence in Cyber Security (IEEE CICS) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16030 2023-11-28 cs.AI cs.LG math.OC 62%

Machine Learning-Enhanced Aircraft Landing Scheduling under Uncertainties

Yutian Pang, Peng Zhao, Jueming Hu, Yongming Liu

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.11136 2023-11-28 cs.CY cs.AI cs.SI 62%

Is There Any Social Principle for LLM-Based Agents?

Jitao Bai, Simiao Zhang, Zhonghao Chen

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.CY

Comments 4 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.11797 2023-11-21 cs.CL cs.AI cs.CV cs.HC cs.MA 62%

Igniting Language Intelligence: The Hitchhiker's Guide From Chain-of-Thought Reasoning to Language Agents

Zhuosheng Zhang, Yao Yao, Aston Zhang, Xiangru Tang, Xinbei Ma, Zhiwei He, Yiming Wang, Mark Gerstein, Rui Wang, Gongshen Liu, Hai Zhao

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.10933 2023-11-21 cs.AI cs.CL cs.CV cs.HC 62%

Representing visual classification as a linear combination of words

Shobhit Agarwal, Yevgeniy R. Semenov, William Lotter

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments To be published in the Proceedings of the 3rd Machine Learning for Health symposium, Proceedings of Machine Learning Research (PMLR)

详情

展开后加载摘要…

URL PDF HTML 收藏