arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

2025-12-12 至 2025-12-12 共收录 34 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 6 篇

2502.01113 2025-12-12 cs.IR cs.AI cs.CL 62%

GFM-RAG: Graph Foundation Model for Retrieval Augmented Generation

基于图的检索增强生成模型:图基础模型

Linhao Luo, Zicheng Zhao, Gholamreza Haffari, Dinh Phung, Chen Gong, Shirui Pan

机构 * Monash University(墨尔本大学) Nanjing University of Science and Technology(南京理工大学) Shanghai Jiao Tong University(上海交通大学) Griffith University(格里菲斯大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

AI总结 GFM-RAG是一种基于图的检索增强生成模型,通过创新的图神经网络捕捉复杂查询-知识关系,实现了在未见数据集上的高效性能和泛化能力。

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10435 2025-12-12 cs.CL 57%

Semantic Reconstruction of Adversarial Plagiarism: A Context-Aware Framework for Detecting and Restoring "Tortured Phrases" in Scientific Literature

对抗性剽窃的语义重建:一种基于上下文的框架,用于检测和恢复科学文献中的‘扭曲短语’

Agniva Maiti, Prajwal Panth, Suresh Chandra Satapathy

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出SRAP框架,通过双阶段架构检测并恢复科学文献中的对抗性剽窃,显著提升恢复准确率。

Comments 10 pages, 5 figures; unpublished manuscript; submitted to arXiv for dissemination

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10316 2025-12-12 cs.CV 50%

ConStruct: Structural Distillation of Foundation Models for Prototype-Based Weakly Supervised Histopathology Segmentation

ConStruct:用于基于原型的弱监督病理分割的基模态结构蒸馏

Khang Le, Ha Thach, Anh M. Vu, Trang T. K. Vo, Han H. Huynh, David Yang, Minh H. N. Le, Thanh-Huy Nguyen, Akash Awasthi, Chandra Mohan, Zhu Han, Hien Van Nguyen

机构 * Ho Chi Minh City University of Technology(胡志明市技术大学) University of Technology Sydney(悉尼技术大学) University of Houston(休斯顿大学) University of Information Technology(信息科技大学) College of Medical Science and Technology, Taipei Medical University(台北医学院医学科技学院) Department of Computer Science, Emory University(埃默里大学计算机科学系) Montefiore Medical Center, Albert Einstein College of Medicine(蒙特福伊医疗中心,阿尔伯特·爱因斯坦医学院) School of Computer Science, Carnegie Mellon University(卡内基梅隆大学计算机科学学院)

专题命中 其他安全 :alignment(abstract)

AI总结 ConStruct通过结合CONCH的形态感知表示、SegFormer的多尺度结构线索和文本引导的语义对齐,提出一种高效的弱监督病理分割方法,提升分割精度和计算效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13901 2025-12-12 cs.CV 50%

Dressing the Imagination: A Dataset for AI-Powered Translation of Text into Fashion Outfits and A Novel NeRA Adapter for Enhanced Feature Adaptation

为想象力着装:一个用于人工智能驱动的文本到时尚装扮的数据库及一种新的NeRA适配器用于增强特征适应

Gayatri Deshmukh, Somsubhra De, Chirag Sehgal, Jishu Sen Gupta, Sparsh Mittal

机构 * IIT Madras(印度理工学院Madras分校) Delhi Technological University(德里技术大学) IIT BHU(印度理工学院BHU分校) IIT Roorkee(印度理工学院Roorkee分校)

专题命中 其他安全 :alignment(abstract)

AI总结 FLORA数据集和NeRA适配器旨在提升人工智能生成时尚设计的精度与风格丰富度。

Comments Accepted as a Conference Paper at WACV 2026 (USA)

详情

展开后加载摘要…

URL PDF HTML 收藏