arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-08-11 至 2025-08-11 共收录 38 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 12 篇

2508.06103 2025-08-11 cs.CL cs.IR 57%

Few-Shot Prompting for Extractive Quranic QA with Instruction-Tuned LLMs

Mohamed Basem, Islam Oshallah, Ali Hamdi, Ammar Mohammed

机构 * Faculty of Computer Science MSA University Giza, Egypt(计算机科学学院 MSA大学 埃及吉扎) Faculty of Computer Science MSA University Egypt(计算机科学学院 MSA大学 埃及)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments 6 pages , 2 figures , Accepted in IMSA 2025,Egypt , https://imsa.msa.edu.eg/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23673 2025-08-11 cs.AI 57%

HASD: Hierarchical Adaption for pathology Slide-level Domain-shift

Jingsong Liu, Han Li, Chen Yang, Michael Deutges, Ario Sadafi, Xin You, Katharina Breininger, Nassir Navab, Peter J. Schüffler

机构 * Institute of Pathology, Technical University of Munich, TUM School of Medicine and Health(病理研究所,慕尼黑技术大学,TUM医学院和健康学院) Computer Aided Medical Procedures (CAMP), TU Munich(计算机辅助医疗程序(CAMP),慕尼黑技术大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心(MCML)) Institute of AI for Health, Computational Health Center, Helmholtz Munich(健康人工智能研究所,计算健康中心,海德堡慕尼黑) Center for AI and Data Science (CAIDAS), Julius-Maximilians-Universität Würzburg(人工智能与数据科学中心(CAIDAS),维尔茨堡约纳斯-马克斯-迈克尔-大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments Accepted by MICCAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08501 2025-08-11 cs.LG 57%

Learning to Match Unpaired Data with Minimum Entropy Coupling

Mustapha Bounoua, Giulio Franzese, Pietro Michiardi

机构 * Department of Data Science, EURECOM, France(数据科学系,EURECOM,法国)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01979 2025-08-11 cs.CL 57%

Gradient-Regularized Latent Space Modulation in Large Language Models for Structured Contextual Synthesis

Derek Yotheringhay, Beatrix Nightingale, Maximilian Featherstone, Edmund Worthington, Hugo Ashdown

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00301 2025-08-11 cs.CL 57%

Contextual Morphogenesis in Large Language Models: A Novel Approach to Self-Organizing Token Representations

Alistair Dombrowski, Beatrix Engelhardt, Dimitri Fairbrother, Henry Evidail

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18826 2025-08-11 cs.CL 57%

Structural Embedding Projection for Contextual Large Language Model Inference

Vincent Enoasmo, Cedric Featherstonehaugh, Xavier Konstantinopoulos, Zacharias Huntington

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06168 2025-08-11 cs.IR 50%

Improving Table Retrieval with Question Generation from Partial Tables

Hsing-Ping Liang, Che-Wei Chang, Yao-Chung Fan

专题命中 其他安全 :alignment(abstract)

Comments TRL@ACL2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09505 2025-08-11 cs.RO 50%

TruckV2X: A Truck-Centered Perception Dataset

Tenghui Xie, Zhiying Song, Fuxi Wen, Jun Li, Guangzhao Liu, Zijian Zhao

专题命中 其他安全 :safety(abstract)

Journal ref IEEE Robotics and Automation Letters, vol. 10, no. 9, pp. 9312-9319, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏