arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1852 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1852 篇

2111.02038 2022-03-23 cs.SE cs.LG 57%

Fair-SSL: Building fair ML Software with less data

Joymallya Chakraborty, Suvodeep Majumder, Huy Tu

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.LG

Journal ref International Workshop on Equitable Data and Technology (FairWare 2022 ), May 9, 2022, Pittsburgh, PA, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.09616 2022-03-21 astro-ph.CO cs.LG 57%

DeepLSS: breaking parameter degeneracies in large scale structure with deep learning analysis of combined probes

Tomasz Kacprzak, Janis Fluri

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG

Comments 18 pages, 10 figures, 2 tables, submitted to Physical Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.01252 2022-03-11 cs.CY 57%

Australia's Approach to AI Governance in Security and Defence

Susannah Kate Devitt, Damian Copeland

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CY

Comments 60 pages, 7 boxes, 2 figures, 2 annexes, submitted for Eds M. Raska, Z. Stanley-Lockman, & R. Bitzinger. AI Governance for National Security and Defence: Assessing Military AI Strategic Perspectives. Routledge

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.11629 2022-02-24 cs.AI stat.ML 57%

A Complete Criterion for Value of Information in Soluble Influence Diagrams

Chris van Merwijk, Ryan Carey, Tom Everitt

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI

Comments In Proceedings of the AAAI 2022 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.15208 2022-02-04 cs.CV cs.AI 57%

HRNET: AI on Edge for mask detection and social distancing

Kinshuk Sengupta, Praveen Ranjan Srivastava

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI

Journal ref SN Computer Science, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.01659 2022-01-06 cs.CY 57%

From the Ground Truth Up: Doing AI Ethics from Practice to Principles

James Brusseau

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CY

Journal ref AI & Soc (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.15234 2022-01-05 cs.AI cs.GT 57%

Artificial Intelligence Development Races in Heterogeneous Settings

Theodor Cimpeanu, Francisco C. Santos, Luis Moniz Pereira, Tom Lenaerts, The Anh Han

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI

Comments 42 pages, 13 figures, accepted for publication in Nature Scientific Reports

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.06883 2021-12-14 cs.SE cs.AI cs.DC 57%

A Methodology for a Scalable, Collaborative, and Resource-Efficient Platform to Facilitate Healthcare AI Research

Raphael Y. Cohen, Vesela P. Kovacheva

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.16122 2021-11-03 cs.HC cs.CY 57%

Zombies in the Loop? Humans Trust Untrustworthy AI-Advisors for Ethical Decisions

Sebastian Krügel, Andreas Ostermaier, Matthias Uhl

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.05164 2021-10-12 cs.CY 57%

Ethical Assurance: A practical approach to the responsible design, development, and deployment of data-driven technologies

Christopher Burr, David Leslie

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.00421 2021-09-22 cs.CL 57%

The Highs and Lows of Simple Lexical Domain Adaptation Approaches for Neural Machine Translation

Nikolay Bogoychev, Pinzhen Chen

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL

Comments Accepted at Workshop on Insights from Negative Results in NLP 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.05388 2021-09-14 cs.CL 57%

The Impact of Positional Encodings on Multilingual Compression

Vinit Ravishankar, Anders Søgaard

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.14099 2021-07-30 cs.CY 57%

The ghost of AI governance past, present and future: AI governance in the European Union

Charlotte Stix

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.13076 2021-07-29 cs.HC cs.LG 57%

Interactive Storytelling for Children: A Case-study of Design and Development Considerations for Ethical Conversational AI

ennifer Chubba, Sondess Missaouib, Shauna Concannonc, Liam Maloneyb, James Alfred Walker

专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.00403 2021-06-09 cs.AI cs.MA math.DS nlin.AO q-bio.PE 57%

Mediating Artificial Intelligence Developments through Negative and Positive Incentives

The Anh Han, Luis Moniz Pereira, Tom Lenaerts, Francisco C. Santos

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.15133 2021-06-01 cs.CY 57%

An Assessment of the AI Regulation Proposed by the European Commission

Patrick Glauner

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY

Comments To appear in the 2022 Springer book "The Future Circle of Healthcare: AI, 3D Printing, Longevity, Ethics, and Uncertainty Mitigation" edited by Sepehr Ehsani, Patrick Glauner, Philipp Plugmann and Florian M. Thieringer

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.01056 2021-05-17 cs.AI cs.MA 57%

Improving Confidence in the Estimation of Values and Norms

Luciano Cavalcante Siebert, Rijk Mercuur, Virginia Dignum, Jeroen van den Hoven, Catholijn Jonker

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI

Comments 16 pages, 3 figures, pre-print for the International Workshop on Coordination, Organizations, Institutions, Norms and Ethics for Governance of Multi-Agent Systems (COINE), co-located with AAMAS 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.02851 2021-05-07 cs.AI 57%

Algorithmic Ethics: Formalization and Verification of Autonomous Vehicle Obligations

Colin Shea-Blymyer, Houssam Abbas

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI

Comments To be published in ACT Transactions on Cyber-Physical Systems Special Issue on Artificial Intelligence and Cyber-Physical Systems. arXiv admin note: text overlap with arXiv:2009.00738

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.03103 2021-03-05 cs.LG 57%

Interpretable Artificial Intelligence through the Lens of Feature Interaction

Michael Tsang, James Enouen, Yan Liu

专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.02647 2021-03-04 cs.RO cs.CV cs.LG 57%

From Learning to Relearning: A Framework for Diminishing Bias in Social Robot Navigation

Juana Valeria Hurtado, Laura Londoño, Abhinav Valada

专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG

Journal ref Frontiers in Robotics and AI, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.12406 2021-02-25 cs.CY 57%

Actionable Principles for Artificial Intelligence Policy: Three Pathways

Charlotte Stix

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CY

Journal ref Sci Eng Ethics 27, 15 (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.08812 2021-01-25 cs.CY 57%

The Internet of Things in Ports: Six Key Security and Governance Challenges for the UK (Policy Brief)

Feja Lesniewska, Uchenna D Ani, Jeremy M Watson, Madeline Carr

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY

Comments 4 pages, 3 Figures, Policy Briefing, Based on research funded by EPSR and carried out by UCL STEaPP NIPC-ALIoTT collaboration project under the PETRAS Cybersecurity Hub

Journal ref The Internet of Things in Ports: Six Key Security and Governance Challenges for the UK (A Policy Brief). London: PETRAS National Centre of Excellence for IoT System Cybersecurity (2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.02048 2020-12-04 cs.CY 57%

Ethical Testing in the Real World: Evaluating Physical Testing of Adversarial Machine Learning

Kendra Albert, Maggie Delano, Jonathon Penney, Afsaneh Rigot, Ram Shankar Siva Kumar

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY

Comments Accepted to NeurIPS 2020 Workshop on Dataset Curation and Security; Also accepted at Navigating the Broader Impacts of AI Research Workshop. All authors contributed equally. The list of authors is arranged alphabetically

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.12530 2020-08-31 cs.CY cs.SY eess.SY physics.soc-ph 57%

Investigating Taxi and Uber competition in New York City: Multi-agent modeling by reinforcement-learning

Saeed Vasebi, Yeganeh M. Hayeri

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY

Comments 12 pages, 10 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.04254 2020-08-11 cs.LG cs.CV stat.ML 57%

Informative Dropout for Robust Representation Learning: A Shape-bias Perspective

Baifeng Shi, Dinghuai Zhang, Qi Dai, Zhanxing Zhu, Yadong Mu, Jingdong Wang

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.LG

Comments Accepted to ICML2020. Code is available at https://github.com/bfshi/InfoDrop

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.01522 2020-07-06 cs.LG stat.ML 57%

Dueling Deep Q-Network for Unsupervised Inter-frame Eye Movement Correction in Optical Coherence Tomography Volumes

Yasmeen M. George, Suman Sedai, Bhavna J. Antony, Hiroshi Ishikawa, Gadi Wollstein, Joel S. Schuman, Rahil Garnavi

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.01978 2020-05-18 cs.LG stat.ML 57%

FANNet: Formal Analysis of Noise Tolerance, Training Bias and Input Sensitivity in Neural Networks

Mahum Naseer, Mishal Fatima Minhas, Faiq Khalid, Muhammad Abdullah Hanif, Osman Hasan, Muhammad Shafique

专题命中 AI治理与伦理 :safety(abstract);分类 cs.LG

Comments To appear at the 23rd Design, Automation and Test in Europe (DATE 2020). Grenoble, France

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.04644 2020-04-10 cs.LG 57%

On the Ethics of Building AI in a Responsible Manner

Shai Shalev-Shwartz, Shaked Shammah, Amnon Shashua

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.00844 2020-04-08 cs.DC cs.IT cs.LG math.IT 57%

Machine Learning at the Wireless Edge: Distributed Stochastic Gradient Descent Over-the-Air

Mohammad Mohammadi Amiri, Deniz Gunduz

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.LG

Comments IEEE Transactions on Signal Processing, Early Access, Mar. 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.11157 2020-03-26 cs.CY 57%

AI loyalty: A New Paradigm for Aligning Stakeholder Interests

Anthony Aguirre, Gaia Dempsey, Harry Surden, Peter B. Reiner

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏