arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing

2026-02-11 至 2026-02-11 共收录 3
2602.09312 2026-02-11 cs.CL cs.AI cs.LG

Don't Shoot The Breeze: Topic Continuity Model Using Nonlinear Naive Bayes With Attention

别射流风:使用非线性朴素贝叶斯与注意机制的主题连续性模型

Shu-Ting Pi, Pradeep Bagavan, Yejia Li, Disha, Qun Liu

机构 * Amazon(亚马逊)

AI总结 本文提出了一种基于非线性朴素贝叶斯与注意力机制的主题连续性模型,用于评估对话响应与初始话题的一致性,有效处理长对话并提升可解释性。

Comments EMNLP 2024: Industry Track; 8 pages, 2 figures, 1 table

Journal ref Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing: Industry Track, pages 65-72

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16596 2026-02-11 cs.CL cs.AI

Analyzing the Effects of Supervised Fine-Tuning on Model Knowledge from Token and Parameter Levels

分析监督微调对模型知识的影响:从token和参数层面

Junjie Ye, Yuming Yang, Yang Nan, Shuo Li, Qi Zhang, Tao Gui, Xuanjing Huang, Peng Wang, Zhongchao Shi, Jianping Fan

机构 * Fudan University(复旦大学) Lenovo Research(联想研究) Shanghai Key Lab of Intelligent Information Processing(上海智能信息处理重点实验室) Shanghai Innovation Institute(上海创新研究院)

AI总结 研究发现监督微调对模型知识的影响显著,通过分析token和参数层面发现大部分参数更新不促进知识增强,恢复更新可提升封闭书问题回答性能。

Comments Accepted by EMNLP 2025 Main Conference. Codes for parameter restoration are available at https://github.com/UmeanNever/ParamRestore

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15214 2026-02-11 cs.CL

Self-Guided Function Calling in Large Language Models via Stepwise Experience Recall

通过逐步经验回忆实现大语言模型的自引导函数调用

Sijia Cui, Aiyao He, Shuai Xu, Hongming Zhang, Yanna Wang, Qingyang Zhang, Yajing Wang, Bo Xu

机构 * The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(认知与决策智能复杂系统重点实验室,自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) Nanjing University of Information Science & Technology(南京信息工程大学) Institute of Computing Technology, Chinese Academy of Sciences(计算技术研究所,中国科学院)

AI总结 SEER通过逐步经验回忆方法,提升大语言模型在多步骤工具使用中的准确性和效率。

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏