arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-02-24 至 2026-02-24 共收录 347 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 32 篇

2602.19881 2026-02-24 cs.CV cs.AI 57%

Make Some Noise: Unsupervised Remote Sensing Change Detection Using Latent Space Perturbations

在潜空间中制造噪声:利用潜空间扰动的无监督遥感变化检测

Blaž Rolih, Matic Fučka, Filip Wolf, Luka Čehovin Zajc

机构 * University of Ljubljana, Faculty of Computer and Information Science(卢布尔雅那大学计算机与信息科学学院)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 MaSoN通过在潜在特征空间中生成多样化变化,实现了无监督遥感变化检测的端到端框架,提升了多种变化类型的泛化能力和性能表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19158 2026-02-24 cs.AI 57%

DoAtlas-1: A Causal Compilation Paradigm for Clinical AI

DoAtlas-1:一种用于临床AI的因果编译范式

Yulong Li, Jianxu Chen, Xiwei Liu, Chuanyue Suo, Rong Xia, Zhixiang Lu, Yichen Li, Xinlin Zhuang, Niranjana Arun Menon, Yutong Xie, Eran Segal, Imran Razzak

机构 * Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) Xi'an Jiaotong-Liverpool University(西安交通大学利物浦大学)

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI

AI总结 DoAtlas-1通过因果编译范式将医学证据转化为可执行代码,提升临床AI的可审计性和可验证性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18607 2026-02-24 cs.AI 57%

Feedback-based Automated Verification in Vibe Coding of CAS Adaptation Built on Constraint Logic

基于反馈的自动验证在CAS适应中的Vibe编码实现

Michal Töpfer, František Plášil, Tomáš Bureš, Petr Hnětynka

专题命中 领域大模型 :LLM(abstract);分类 cs.AI

AI总结 本文提出通过Vibe编码反馈循环结合精确约束逻辑,实现CAS适应中AM的自动验证与生成,有效提升代码正确性与适应性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19823 2026-02-24 cs.CV 50%

Open-vocabulary 3D scene perception in industrial environments

工业环境中的开放词汇3D场景感知

Keno Moenck, Adrian Philip Florea, Julian Koch, Thorsten Schüppstuhl

机构 * Hamburg University of Technology, Institute of Aircraft Production Technology(汉堡工业大学航空生产技术研究所)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本研究提出了一种无需训练的开放词汇3D感知流程,通过合并预计算超点生成掩码,提升工业环境中的3D场景感知能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14512 2026-02-24 cs.CV 50%

MedVAR: Towards Scalable and Efficient Medical Image Generation via Next-scale Autoregressive Prediction

MedVAR:通过下一步尺度自回归预测实现可扩展和高效的医学图像生成

Zhicheng He, Yunpeng Zhao, Junde Wu, Ziwei Niu, Zijun Li, Bohan Li, Lanfen Lin, Yueming Jin

机构 * National University of Singapore(新加坡国立大学) University of Oxford(牛津大学) Zhejiang University(浙江大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 领域大模型 :foundation model(abstract)

AI总结 MedVAR通过下一步尺度自回归预测方法,实现了高效且可扩展的医学图像生成,为医学生成基础模型提供了新的方向。

Comments 23 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18811 2026-02-24 cs.CV 50%

Learning Multi-Modal Prototypes for Cross-Domain Few-Shot Object Detection

跨域少样本目标检测中的多模态原型学习

Wanqi Wang, Jingcai Guo, Yuxiang Cai, Zhi Chen

机构 * University of Chinese Academy of Sciences(中国科学院大学) The Hong Kong Polytechnic University(香港理工大学) Zhejiang University(浙江大学) The University of Southern Queensland(昆士兰大学)

专题命中 领域大模型 :language model(abstract)

AI总结 本文提出LMP方法,通过结合文本和视觉信息,提升跨域少样本目标检测的精度和性能。

Comments Accepted to CVPR 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18614 2026-02-24 cs.CV 50%

Effect of Patch Size on Fine-Tuning Vision Transformers in Two-Dimensional and Three-Dimensional Medical Image Classification

二维和三维医学图像分类中视觉变换器微调受补丁大小的影响

Massoud Dehghan, Ramona Woitek, Amirreza Mahbod

机构 * Research Center for Medical Image Analysis and Artificial Intelligence(医学图像分析与人工智能研究中心) Department of Medicine(医学系) Faculty of Medicine and Dentistry(医学与牙科学院) Danube Private University(多瑙私立大学)

专题命中 领域大模型 :foundation model(abstract)

AI总结 本文研究了补丁大小对ViT在医学图像分类中的影响,发现较小补丁尺寸能提升分类性能,且通过模型集成策略进一步提高效果。

Comments 29 pages

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 16 篇

2510.24856 2026-02-24 cs.CL 88%

Do Large Language Models Grasp The Grammar? Evidence from Grammar-Book-Guided Probing in Luxembourgish

大型语言模型是否理解语法规则?来自卢森堡语语法书引导探测的证据

Lujun Li, Yewei Song, Lama Sleem, Yiqun Wang, Yangjie Xu, Cedric Lothritz, Niccolo Gentile, Radu State, Tegawende F. Bissyande, Jacques Klein

机构 * University of Luxembourg(卢森堡大学) Luxembourg Institute of Science and Technology(卢森堡科学与技术研究院)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 本研究通过卢森堡语语法书引导探测,探讨大型语言模型对语法规则的理解,发现翻译性能与语法规则理解弱相关,大模型在语义上表现良好但句法和形态学能力较弱,推理能力有助于提升语法规则理解。

Comments This paper has been accepted for publication in the proceedings of the 15th biennial Language Resources and Evaluation Conference (LREC 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16412 2026-02-24 cs.CV 88%

ReMoRa: Multimodal Large Language Model based on Refined Motion Representation for Long-Video Understanding

ReMoRa:基于精细运动表示的多模态大语言模型用于长视频理解

Daichi Yashima, Shuhei Kurita, Yusuke Oda, Komei Sugiura

机构 * Keio University(庆应大学) NII(日本信息处理学会) NII LLMC(日本信息处理学会语言模型中心)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract)

AI总结 ReMoRa通过精细运动表示实现长视频理解,有效压缩视频数据并提升多模态大语言模型性能。

Comments Accepted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19022 2026-02-24 cs.CV cs.AI 79%

An interpretable framework using foundation models for fish sex identification

基于基础模型的可解释框架用于鱼类性别识别

Zheng Miao, Tien-Chieh Hung

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

AI总结 基于基础模型的可解释框架用于濒危鱼类三角洲虾虎鱼的性别识别,通过原型网络提升鲁棒性与可解释性,实现高准确率识别。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19159 2026-02-24 cs.AI cs.CL cs.LG 78%

Beyond Behavioural Trade-Offs: Mechanistic Tracing of Pain-Pleasure Decisions in an LLM

超越行为权衡:在LLM中疼痛-愉悦决策的机制追溯

Francesca Bianco, Derek Shiller

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.CL、cs.AI、cs.LG

AI总结 研究揭示了LLM在疼痛-愉悦决策中的内部机制,通过机制追溯揭示了价值信号的表示和因果作用,为AI意识和福利的讨论提供了证据基础。

Comments 24 pages, 8+1 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19396 2026-02-24 cs.AI 77%

Hiding in Plain Text: Detecting Concealed Jailbreaks via Activation Disentanglement

明文之中隐藏:通过激活解耦检测隐蔽的 jailbreak

Amirhossein Farzam, Majid Behabahani, Mani Malek, Yuriy Nevmyvaka, Guillermo Sapiro

机构 * Duke University(杜克大学) Princeton University(普林斯顿大学) Google DeepMind(谷歌DeepMind) Apple(苹果公司)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 通过解耦 LLM 激活中的语义因子,提出 FrameShield 异常检测器,提升对隐蔽 jailbreak 的检测能力,并推动 LLM 安全和可解释性研究。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04568 2026-02-24 cs.AI cs.CL cs.IR cs.LG 75%

Neurosymbolic Retrievers for Retrieval-augmented Generation

用于检索增强生成的神经符号检索器

Yash Saxena, Manas Gaur

机构 * Dept. of CSEE University of Maryland Baltimore County, Maryland, USA(电子工程系大学马里兰大学巴尔的摩县) Dept. of CSEE University of Maryland Baltimore County, MD, USA(电子工程系大学马里兰大学巴尔的摩县)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出神经符号 RAG 框架,通过结合知识图谱与神经检索技术,提升检索过程的透明性和生成性能。

Comments 8 pages, 2 Figures, Published in IEEE Intelligent Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11230 2026-02-24 cs.CL 70%

Sparse Autoencoders Can Capture Language-Specific Concepts Across Diverse Languages

稀疏自编码器可以在多种语言中捕捉语言特定的概念

Lyzander Marciano Andrylie, Inaya Rahmanisa, Mahardika Krisna Ihsani, Alfan Farizki Wicaksono, Haryo Akbarianto Wibowo, Alham Fikri Aji

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究利用稀疏自编码器识别语言特定的特征,揭示其在多语言处理中的作用及可解释性优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19043 2026-02-24 cs.CL 70%

Uncovering Context Reliance in Unstructured Knowledge Editing

揭示无结构知识编辑中的上下文依赖性

Zisheng Zhou, Mengqi Zhang, Shiguang Wu, Xiaotian Ye, Chi Zhang, Zhumin Chen, Pengjie Ren

机构 * Shandong University(山东大学) Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出COIN框架,通过减少上下文依赖性提升大型语言模型的无结构知识编辑效果。

Comments 21 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18711 2026-02-24 cs.CV 67%

HIME: Mitigating Object Hallucinations in LVLMs via Hallucination Insensitivity Model Editing

HIME: 通过幻觉不敏感模型编辑缓解LVLMs中的物体幻觉

Ahmed Akl, Abdelwahed Khamis, Ali Cheraghian, Zhe Wang, Sara Khalifa, Kewen Wang

机构 * School of Information and Communication Technology, Griffith University, Australia(信息与通信技术学院,格里菲斯大学) Data61, CSIRO, Australia(Data61,澳大利亚联邦科学与工业研究组织) School of Engineering, Macquarie University, Sydney, Australia(工程学院,麦觉大学) School of Information Systems, Queensland University of Technology, Australia(信息系统学院,昆士兰技术大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract)

AI总结 HIME通过分层加权编辑方法有效抑制LVLMs中的物体幻觉,减少61.8%的幻觉问题,无需额外参数或计算开销。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06199 2026-02-24 cs.LG cs.AI 62%

Benchmarking Pretrained Molecular Embedding Models For Molecular Representation Learning

对预训练分子嵌入模型进行基准测试:用于分子表示学习

Mateusz Praski, Jakub Adamczyk, Wojciech Czech

机构 * Faculty of Computer Science(计算机科学系) AGH University of Krakow(克拉科夫AGH大学)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI、cs.LG

AI总结 本研究对预训练分子嵌入模型进行了全面比较,发现仅CLAMP模型在分子表示学习中表现显著优于其他模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18476 2026-02-24 q-bio.BM cs.AI cs.LG 62%

BioLM-Score: Language-Prior Conditioned Probabilistic Geometric Potentials for Protein-Ligand Scoring

BioLM-Score:基于语言先验的概率几何势用于蛋白质-配体评分

Zhangfan Yang, Baoyun Chen, Dong Xu, Jia Wang, Ruibin Bai, Junkai Ji, Zexuan Zhu

机构 * School of Computer Science, University of Nottingham Ningbo(计算机科学学院,诺丁汉大学宁波分校) School of Artificial Intelligence, Shenzhen University(人工智能学院,深圳大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.AI、cs.LG

AI总结 BioLM-Score结合几何建模与表征学习,提供一种高效、可泛化且可解释的蛋白质-配体评分方法,提升药物发现效率。

Comments 9 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19215 2026-02-24 cs.LG 57%

Understanding Empirical Unlearning with Combinatorial Interpretability

理解经验性反学习与组合可解释性

Shingo Kodama, Niv Cohen, Micah Adler, Nir Shavit

机构 * Middlebury College(中大西洋学院) New York University(纽约大学) MIT(麻省理工学院) Red Hat(红帽公司)

专题命中 知识编辑与模型理解 :foundation model(abstract);分类 cs.LG

AI总结 本文通过组合可解释性框架研究经验性反学习中知识的持续存在机制,揭示反学习方法在移除目标概念知识方面的有效性及知识恢复的可能性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19138 2026-02-24 q-bio.NC cs.AI 57%

CRCC: Contrast-Based Robust Cross-Subject and Cross-Site Representation Learning for EEG

CRCC: 基于对比的跨受试者和跨站点表示学习

Xiaobin Wong, Zhonghua Zhao, Haoran Guo, Zhengyi Liu, Yu Wu, Feng Yan, Zhiren Wang, Sen Song

机构 * Tsinghua Laboratory of Brain(清华大学脑科学实验室) School of Biomedical Engineering, Tsinghua University(清华大学生物医学工程学院) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) University of Chinese Academy of Sciences(中国科学院大学) Weixian College, Tsinghua University(清华大学魏先学院) School of Artificial Intelligence, Beijing University of Posts(北京邮电大学人工智能学院) School of Computer Science(计算机科学学院) Technology, Northwestern Polytechnical University, Xi'an, China(技术,西北工业大学,西安,中国) Beijing Huilongguan Hospital, Capital Medical University(北京回龙观医院,首都医科大学) Peking University Huilongguan Clinical Medical School(北京大学回龙观临床医学院)

专题命中 知识编辑与模型理解 :pretraining(abstract);分类 cs.AI

AI总结 CRCC通过对比学习和对抗优化提升跨站点EEG表示学习的泛化能力,实现10.7个百分点的准确率提升。

Comments First edition

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14934 2026-02-24 stat.ML cs.LG 57%

Activation-Space Uncertainty Quantification for Pretrained Networks

预训练网络中的激活空间不确定性量化

Richard Bergna, Stefan Depeweg, Sergio Calvo-Ordoñez, Jonathan Plenk, Alvaro Cartea, Jose Miguel Hernández-Lobato

机构 * Department of Engineering, University of Cambridge, Cambridge, UK(剑桥大学工程系) Mathematical Institute(数学研究所) Oxford-Man Institute, University of Oxford, Oxford, UK(牛津大学奥克斯曼研究所)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.LG

AI总结 GAPA通过在激活空间中提供闭合形式的epistemic方差,实现了预训练网络的高效不确定性量化,无需采样或反向传播。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18505 2026-02-24 cs.CV 50%

Suppression or Deletion: A Restoration-Based Representation-Level Analysis of Machine Unlearning

抑制或删除:基于恢复的表示层面机器去学习分析

Yurim Jang, Jaeung Lee, Dohyun Kim, Jaemin Jo, Simon S. Woo

机构 * Department of Artificial Intelligence Sungkyunkwan University Suwon Republic of Korea(人工智能系首尔大学水原韩国) Sungkyunkwan University(首尔大学)

专题命中 知识编辑与模型理解 :pretraining(abstract)

AI总结 本文提出了一种基于恢复的分析框架,揭示了现有去学习方法在表示层面保留信息的风险,强调了对新评估标准的需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06084 2026-02-24 cs.CV 50%

Exploring Interpretability for Visual Prompt Tuning with Cross-layer Concepts

探索通过跨层概念实现的视觉提示调优可解释性

Yubin Wang, Xinyang Jiang, De Cheng, Xiangqian Zhao, Zilong Wang, Dongsheng Li, Cairong Zhao

机构 * School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院) Microsoft Research Asia(微软亚洲研究院) School of Telecommunication and Engineering, Xidian University(西安电子科技大学电信与工程学院)

专题命中 知识编辑与模型理解 :foundation model(abstract)

AI总结 本文提出可解释视觉提示调优框架,通过跨层概念原型提升视觉提示的可解释性与性能。

Comments ICLR 2026, 21 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 24 篇

2504.18880 2026-02-24 cs.AI cond-mat.mtrl-sci cs.CL 88%

Reshaping MOFs text mining with a dynamic multi-agents framework of large language model

用大语言模型的动态多智能体框架重塑MOFs文本挖掘

Zuhong Lin, Daoyuan Ren, Kai Ran, Jing Sun, Songlin Yu, Xuefeng Bai, Xiaotian Huang, Haiyang He, Pengxu Pan, Ying Fang, Zhanglin Li, Haipu Li, Jingjing Yao

机构 * Center for Environment and Water Resources, College of Chemistry and Chemical Engineering, Central South University(环境与水资源中心,化学与化工学院,中南大学) Key Laboratory of Hunan Province for Water Environment and Agriculture Product Safety(湖南省水环境与农产品安全重点实验室) School of Resources and Environment, Hunan University of Technology and Business(资源与环境学院,湖南工业大学) School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) Faculty of Data Science, City University of Macau(数据科学学院,澳门城市大学) State Key Laboratory of High Performance Ceramics and Superfine Microstructure, Shanghai Institute of Ceramics, Chinese Academy of Sciences(高性能陶瓷与超细微结构重点实验室,上海陶瓷研究所,中国科学院) Beijing Key Laboratory for Green Catalysis and Separation, Department of Chemical Engineering, College of Materials Science and Engineering, Beijing University of Technology(绿色催化与分离北京市重点实验室,化学工程系,材料科学与工程学院,北京理工大学) State Key Joint Laboratory of Environment Simulation and Pollution Control, School of Environment, Tsinghua University(环境模拟与污染控制国家重点联合实验室,环境学院,清华大学) School of Chemical Engineering and Materials Science, Yueyang University(化学工程与材料科学学院,岳阳大学) School of Computer Science and Engineering, Central South University(计算机科学与工程学院,中南大学) School of Software Engineering, Sun Yat-sen University(软件工程学院,中山大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 MOFh6利用大语言模型的动态多智能体框架,实现MOFs合成条件的高效提取与标准化,提升材料发现的效率和可扩展性。

Comments Accepted by TRAMAT 2 (2026) 100176

Journal ref Transactions of Materials Research, 2026, 2(1), 100176

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19124 2026-02-24 cs.HC 88%

Dark and Bright Side of Participatory Red-Teaming with Targets of Stereotyping for Eliciting Harmful Behaviors from Large Language Models

参与式红队行动的黑暗与光明面:针对刻板印象目标以激发大语言模型有害行为

Sieun Kim, Yeeun Jo, Sungmin Na, Hyunseung Lim, Eunchae Lee, Yu Min Choi, Soohyun Cho, Hwajung Hong

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract)

AI总结 本文通过参与式红队行动研究,探讨如何利用刻板印象目标的亲身经历揭示大语言模型的偏见,同时关注参与者心理福祉与赋权。

Comments 20 pages, 4 tables, 3 figures. Accepted to CHI 2026, April 13-17, 2026, Barcelona, Spain

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18511 2026-02-24 cs.PL cs.AI 86%

Beyond Pass-by-Pass Optimization: Intent-Driven IR Optimization with Large Language Models

超越逐步优化:基于大语言模型的意图驱动的中间表示优化

Lei Qiu, Zi Yang, Fang Lyu, Ming Zhong, Huimin Cui, Xiaobing Feng

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract);分类 cs.AI

AI总结 IntOpt通过显式分离高层次优化意图与低层次转换,实现了更高效的中间表示优化,提升了正确性和性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13033 2026-02-24 cs.CY cs.AI cs.CE cs.CL cs.SI 86%

Buy versus Build an LLM: A Decision Framework for Governments

买还是建一个大语言模型:政府的决策框架

Jiahao Lu, Ziwei Xu, William Tjhi, Junnan Li, Antoine Bosselut, Pang Wei Koh, Mohan Kankanhalli

机构 * National University of Singapore(新加坡国立大学) AI Singapore(AI新加坡) Salesforce AI Research(Salesforce AI研究) EPFL(苏黎世联邦理工学院) University of Washington(华盛顿大学) Allen Institute for AI(人工智能研究院)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出政府在大语言模型决策中应考虑主权、安全、成本等因素的框架,帮助确定购买或建设更适合其需求的方法。

Comments The short version of this document is published as an ACM TechBrief at https://dl.acm.org/doi/epdf/10.1145/3797946, and this document is published as an ACM Technology Policy Council white paper at https://www.acm.org/binaries/content/assets/public-policy/buildvsbuyai.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23465 2026-02-24 cs.CL cs.AI 86%

Role-Aware Language Models for Secure and Contextualized Access Control in Organizations

面向角色的语言模型:用于组织中的安全且上下文化的访问控制

Saeed Almheiri, Yerulan Kongrat, Adrian Santosh, Ruslan Tasmukhanov, Josemaria Loza Vera, Muhammad Dehan Al Kautsar, Fajri Koto

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) Nazarbayev University(纳扎尔拜耶夫大学) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) New York University Abu Dhabi(纽约大学阿布扎克分校)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出面向角色的语言模型,通过三种策略实现基于组织角色的安全访问控制,并通过实验验证其在不同组织结构下的性能和鲁棒性。

Comments AACL 2025 - Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02099 2026-02-24 cs.CR cs.CL cs.LG 86%

A Watermark for Black-Box Language Models

为黑盒语言模型设计的水印

Dara Bahri, John Wieting

机构 * Google DeepMind(谷歌DeepMind)

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出了一种无需白盒访问即可检测LLM输出的水印方案,具备无失真和多密钥嵌套特性,并通过实验验证其优越性。

Comments Published at TMLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18449 2026-02-24 cs.CL cs.AI cs.LG 85%

Prompt Optimization Via Diffusion Language Models

通过扩散语言模型实现提示优化

Shiyu Wang, Haolin Chen, Liangwei Yang, Jielin Qiu, Rithesh Murthy, Ming Zhu, Zixiang Chen, Silvio Savarese, Caiming Xiong, Shelby Heinecke, Huan Wang

专题命中 其他LLM :language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出利用扩散语言模型实现提示优化,通过迭代精炼提升LLM性能。

详情

展开后加载摘要…

URL PDF HTML 收藏