arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing

2026-01-08 至 2026-01-08 共收录 5
2505.20118 2026-01-08 cs.CL cs.CR

TrojanStego: Your Language Model Can Secretly Be A Steganographic Privacy Leaking Agent

TrojanStego: 你的语言模型可能 secretly 成为一个隐写术隐私泄露代理

Dominik Meier, Jan Philip Wahle, Paul Röttger, Terry Ruas, Bela Gipp

机构 * University of Göttingen(哥廷根大学) LKA NRW(北莱茵-威斯特法伦州检察署) Bocconi University(博科尼大学)

AI总结 TrojanStego通过语言隐写术在LLM输出中隐秘泄露敏感信息,展示了一种新型被动且危险的LLM数据外泄攻击方式。

Comments 9 pages, 5 figures To be presented in the Conference on Empirical Methods in Natural Language Processing, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13197 2026-01-08 cs.CL

Text Anomaly Detection with Simplified Isolation Kernel

基于简化隔离核的文本异常检测

Yang Cao, Sikun Yang, Yujiu Yang, Lianyong Qi, Ming Liu

机构 * School of Computing and Information Technology, Great Bay University(大湾大学计算机与信息科技学院) Great Bay Institute for Advanced Study, Great Bay University(大湾大学先进研究学院) Guangdong Provincial Key Laboratory of Mathematical and Neural Dynamical Systems(广东省数学与神经动力系统重点实验室) Dongguan Key Laboratory for Intelligence and Information Technology, China(东莞智能与信息技术重点实验室) Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) China University of Petroleum (East China), China(中国石油大学(华东)) School of IT, Deakin University(德肯大学信息科技学院)

AI总结 本研究提出简化隔离核(SIK)方法,通过将高维嵌入转换为低维稀疏表示,提升文本异常检测的性能与效率。

Comments EMNLP Findings 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13798 2026-01-08 cs.CL

TracSum: A New Benchmark for Aspect-Based Summarization with Sentence-Level Traceability in Medical Domain

TracSum:面向医学领域具有句子级可追溯性的新型摘要基准

Bohao Chu, Meijie Li, Sameh Frihat, Chengyu Gu, Georg Lodde, Elisabeth Livingstone, Norbert Fuhr

机构 * University of Duisburg-Essen(杜伊斯堡-埃森大学) Institute for AI in Medicine (IKIM)(医学人工智能研究所) University Hospital Essen(埃森大学医院)

AI总结 TracSum提出了一种面向医学领域的可追溯摘要基准,通过句子级引用提高摘要准确性,并展示了基于Track-Then-Sum的基线方法及实验结果。

Comments 8 main pages, 12 appendix pages

Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15722 2026-01-08 cs.CL cs.AI

Shared Path: Unraveling Memorization in Multilingual LLMs through Language Similarities

共享路径:通过语言相似性揭示多语言大语言模型中的记忆现象

Xiaoyu Luo, Yiyi Chen, Johannes Bjerva, Qiongxiu Li

机构 * Department of Computer Science(计算机科学系) Department of Electronic Systems(电子系统系)

AI总结 通过语言相似性分析,揭示多语言大语言模型中记忆现象的跨语言关联及其影响因素。

Comments 17 pages, 14 tables, 10 figures

Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pages 19372-19388, Suzhou, China

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.15066 2026-01-08 cs.CL

InsertGNN: Can Graph Neural Networks Outperform Humans in TOEFL Sentence Insertion Problem?

InsertGNN:图神经网络能否在TOEFL句子插入问题中超越人类?

Fang Wu, Stan Z. Li

机构 * Stanford University(斯坦福大学) Westlake University(西拉丘市大学)

AI总结 InsertGNN通过图神经网络在TOEFL句子插入任务中超越人类,实现70%的准确率。

Journal ref EMNLP 2024 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏