arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1852 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1852 篇

2510.06306 2025-10-09 cs.HC 50%

"Grillz on a hijabi": Intersectional Identities in Fostering Critical AI Literacy

Jaemarie Solyst, Chloe Fong, Faisal Nurdin, Rotem Landesman, R. Benjamin Shapiro

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06222 2025-10-09 cs.HC econ.GN q-fin.EC 50%

Inducing State Anxiety in LLM Agents Reproduces Human-Like Biases in Consumer Decision-Making

Ziv Ben-Zion, Zohar Elyoseph, Tobias Spiller, Teddy Lazebnik

专题命中 AI治理与伦理 :safety(abstract)

Comments Manuscript Main Text - 20 pages, including 3 Figures and 1 Table. Supplementary Materials - 10 pages, including 4 Supplemental Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04038 2025-10-07 eess.SY cs.SY 50%

Distributed MPC-based Coordination of Traffic Perimeter and Signal Control: A Lexicographic Optimization Approach

Viet Hoang Pham, Hyo-Sung Ahn

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25963 2025-10-01 cs.CV 50%

Self-Supervised Anatomical Consistency Learning for Vision-Grounded Medical Report Generation

Longzhen Yang, Zhangkai Ni, Ying Wen, Yihang Liu, Lianghua He, Heng Tao Shen

机构 * Tongji University(同济大学) East China Normal University(华东师范大学) Shanghai Eye Disease Prevention and Treatment Center(上海眼病防治中心)

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22424 2025-09-29 q-bio.OT 50%

Desiderata for a biomedical knowledge network: opportunities, challenges and future Directions

Chunlei Wu, Hongfang Liu, Jason Flannick, Mark A. Musen, Andrew I. Su, Lawrence Hunter, Thomas M. Powers, Cathy H. Wu

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments 6 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06700 2025-09-09 cs.NI 50%

Sovereign AI for 6G: Towards the Future of AI-Native Networks

Swarna Bindu Chetty, David Grace, Simon Saunders, Paul Harris, Eirini Eleni Tsiropoulou, Tony Quek, Hamed Ahmadi

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18600 2025-08-27 cs.GT cs.MA econ.GN q-fin.EC 50%

Bias-Adjusted LLM Agents for Human-Like Decision-Making via Behavioral Economics

Ayato Kitadai, Yusuke Fukasawa, Nariaki Nishino

专题命中 AI治理与伦理 :alignment(abstract)

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10286 2025-08-19 cs.HC 50%

Artificial Emotion: A Survey of Theories and Debates on Realising Emotion in Artificial Intelligence

Yupei Li, Qiyang Sun, Michelle Schlicher, Yee Wen Lim, Björn W. Schuller

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18328 2025-07-25 cs.NI 50%

Enhanced Velocity-Adaptive Scheme: Joint Fair Access and Age of Information Optimization in Vehicular Networks

Xiao Xu, Qiong Wu, Pingyi Fan, Kezhi Wang, Nan Cheng, Wen Chen, Khaled B. Letaief

专题命中 AI治理与伦理 :safety(abstract)

Comments This paper has been submitted to IEEE TMC

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21898 2025-07-09 cs.HC 50%

Bias, Accuracy, and Trust: Gender-Diverse Perspectives on Large Language Models

Aimen Gaba, Emily Wall, Tejas Ramkumar Babu, Yuriy Brun, Kyle Hall, Cindy Xiong Bearfield

专题命中 AI治理与伦理 :trustworthy(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05046 2025-07-08 cs.HC 50%

What Shapes User Trust in ChatGPT? A Mixed-Methods Study of User Attributes, Trust Dimensions, Task Context, and Societal Perceptions among University Students

Kadija Bouyzourn, Alexandra Birch

专题命中 AI治理与伦理 :alignment(abstract)

Comments 25 pages, 11 tables, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04534 2025-07-08 cs.SI 50%

Simulating User Watch-Time to Investigate Bias in YouTube Shorts Recommendations

Selimhan Dagtas, Mert Can Cakmak, Nitin Agarwal

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01069 2025-07-03 cs.CE cs.SE 50%

Agentic AI in Product Management: A Co-Evolutionary Model

Nishant A. Parikh

专题命中 AI治理与伦理 :alignment(abstract)

Comments 41 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16889 2025-07-02 cs.CV 50%

Beyond Diagnostic Performance: Revealing and Quantifying Ethical Risks in Pathology Foundation Models

Weiping Lin, Shen Liu, Runchen Zhu, Yixuan Lin, Baoshun Wang, Liansheng Wang

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments 33 pages,5 figure,23 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19081 2025-06-25 cond-mat.soft 50%

Emergent collective dynamics from motile photokinetic organisms

J. Morales, P. Munoz, D. Noto, H. N Ulloa, F. Guzman-Lastra

专题命中 AI治理与伦理 :alignment(abstract)

Comments 11 pages 5 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03409 2025-06-19 cs.CR 50%

Technical Options for Flexible Hardware-Enabled Guarantees

James Petrie, Onni Aarne

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14166 2025-06-18 cs.HC 50%

Affective-CARA: A Knowledge Graph Driven Framework for Culturally Adaptive Emotional Intelligence in HCI

Nirodya Pussadeniya, Bahareh Nakisa, Mohmmad Naim Rastgoo

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07766 2025-06-11 cs.CR cs.SE econ.TH 50%

Realigning Incentives to Build Better Software: a Holistic Approach to Vendor Accountability

Gergely Biczók, Sasha Romanosky, Mingyan Liu

专题命中 AI治理与伦理 :safety(abstract)

Comments accepted to WEIS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17937 2025-05-29 cs.HC 50%

Survival Games: Human-LLM Strategic Showdowns under Severe Resource Scarcity

Zhihong Chen, Yiqian Yang, Jinzhao Zhou, Qiang Zhang, Chin-Teng Lin, Yiqun Duan

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19895 2025-05-27 cs.CV 50%

Underwater Diffusion Attention Network with Contrastive Language-Image Joint Learning for Underwater Image Enhancement

Afrah Shaahid, Muzammil Behzad

机构 * King Fahd University of Petroleum and Minerals(国王法赫德石油与矿物大学)

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19045 2025-05-27 econ.GN q-fin.EC 50%

A General Theory of Growth, Employment, and Technological Change: Experiential Matrix Theory and the Transition from GDP to Humanist Experiential Growth in the Age of Artificial Intelligence

Christian Callaghan

专题命中 AI治理与伦理 :alignment(abstract)

Comments 57 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04752 2025-05-14 cs.IR 50%

Investigating Popularity Bias Amplification in Recommender Systems Employed in the Entertainment Domain

Dominik Kowald

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments Accepted at EWAF'25, summarizes fairness and popularity bias research presented in Dr. Kowald's habilitation: https://domkowald.github.io/documents/others/2024habilitation_recsys.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07377 2025-05-13 cs.IR 50%

Process-Supervised LLM Recommenders via Flow-guided Tuning

Chongming Gao, Mengyao Gao, Chenxiao Fan, Shuai Yuan, Wentao Shi, Xiangnan He

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted by SIGIR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05596 2025-05-12 econ.GN q-fin.EC 50%

Multi-level Governance, Smart Meter Adoption, and Utilities' Energy Efficiency Savings in the U.S

Yue Gao, Jing Zhang

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07787 2025-04-11 cs.SE 50%

Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language Models

Yisong Xiao, Aishan Liu, Siyuan Liang, Xianglong Liu, Dacheng Tao

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted by ISSTA 2025.20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07610 2025-04-11 cs.MA 50%

What Contributes to Affective Polarization in Networked Online Environments? Evidence from an Agent-Based Model

Narayani Vedam, Subhayan Mukerjee, Prasanta Bhattacharya

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17671 2025-04-04 cs.CV 50%

A Bias-Free Training Paradigm for More General AI-generated Image Detection

Fabrizio Guillaro, Giada Zingarini, Ben Usman, Avneesh Sud, Davide Cozzolino, Luisa Verdoliva

机构 * University Federico II of Naples(那不勒斯费德里科二世大学) Google DeepMind(谷歌DeepMind)

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21165 2025-03-28 eess.SY cs.AR cs.SY 50%

Extending Silicon Lifetime: A Review of Design Techniques for Reliable Integrated Circuits

Shaik Jani Babu, Fan Hu, Linyu Zhu, Sonal Singhal, Xinfei Guo

专题命中 AI治理与伦理 :safety(abstract)

Comments This work is under review by ACM

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09230 2025-03-24 cs.CV 50%

Surgical Text-to-Image Generation

Chinedu Innocent Nwoye, Rupak Bose, Kareem Elgohary, Lorenzo Arboit, Giorgio Carlino, Joël L. Lavanchy, Pietro Mascagni, Nicolas Padoy

机构 * University of Strasbourg(斯特拉斯堡大学) Fondazione Policlinico Universitario Agostino Gemelli IRCCS(阿戈斯蒂诺·杰梅里基金会天主教大学综合医院) IHU Strasbourg(斯特拉斯堡大学医院研究所) University of Basel(巴塞尔大学)

专题命中 AI治理与伦理 :alignment(abstract)

Comments 13 pages, 13 figures, 3 tables, published in Pattern Recognition Letters 2025, project page at https://camma-public.github.io/endogen/

Journal ref Pattern Recognition Letters, Volume 190, April 2025, Pages 73-80

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15949 2025-03-21 cs.CV 50%

CausalCLIPSeg: Unlocking CLIP's Potential in Referring Medical Image Segmentation with Causal Intervention

Yaxiong Chen, Minghong Wei, Zixuan Zheng, Jingliang Hu, Yilei Shi, Shengwu Xiong, Xiao Xiang Zhu, Lichao Mou

机构 * Wuhan University of Technology(武汉理工大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) MedAI Technology (Wuxi) Co. Ltd.(MedAI科技(无锡)有限公司) Technical University of Munich(慕尼黑工业大学)

专题命中 AI治理与伦理 :alignment(abstract)

Comments MICCAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏