arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Annual Meeting of the Association for Computational Linguistics · 会议 · Natural Language Processing

2026-08-05 至 2026-08-05 共收录 6
2608.03810 2026-08-05 cs.CL cs.AI 新提交

VIBE: A VAD-Informed Benchmark for Entity-Centered Affective Profiling of Large Language Model Outputs

VIBE:面向大型语言模型输出的以实体为中心的情感分析基准,基于VAD(效价-唤醒度-支配度)

Andrei Chetvergov, Alexander Evseev, Timofei Sivoraksha, Stepan Ukolov, Mikhail Solovev, Danil Sazanakov, Sergey Bolovtsov

AI总结 本文提出基于VAD的VIBE基准,用于以实体为中心分析大型语言模型输出的情感,明确了测量契约,通过三个实证层面验证了相关假设,推动该领域成为规范化实践。

Comments 25 pages, 13 figures, 22 tables. Submitted to ACL Rolling Review, August 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.03411 2026-08-05 cs.CL 新提交

DUD: Decoupled Update Dynamics for Reliable Uncertainty Quantification in Large Language Models

DUD:用于大语言模型可靠不确定性量化的解耦更新动力学

Yixin Bu, Runze Xia, Guanyun Zou, Yupeng Ji, Haodong Liu, Piji Li

AI总结 本研究针对大语言模型不确定性量化的传统方法缺陷,提出DUD框架通过因果干预解耦FFN与注意力的更新贡献,实验表明其在不确定性估计、校准及跨数据集泛化上均优于现有基线。

Comments ACL 2026 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02703 2026-08-05 cs.CL cs.LG 新提交

ARCHead: Activation-Metric Residual Correction for Large Language Model Output Heads

ARCHead:大语言模型输出头的激活度量残差修正

Şuayp Talha Kocabay, Talha Rüzgar Akkuş, Kamer Ali Yuksel

机构 * aiXplain, Inc.(aiXplain公司)

AI总结 ARCHead是一种LM-head压缩器,通过量化低秩核心、分组INT4残差及激活衍生低秩修正,将LM-head存储降3.7-3.9倍,在Qwen3-8B-Base上性能优于朴素INT4,可补充块量化器。

Comments 13 pages, 4 figures. Submitted to ACL Rolling Review (ARR). Code: https://github.com/suayptalha/archead

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02609 2026-08-05 cs.CL cs.LG 新提交

TabletCraft: Bridging a 4,000-Year Cultural Gap with Bidirectional Akkadian NMT and Cuneiform Rendering

TabletCraft:通过双向阿卡德语神经机器翻译(NMT)与楔形文字渲染弥合4000年的文化鸿沟

Zhaohui Wang

机构 * University of Southern California(南加州大学) USC Viterbi School of Engineering(南加州大学维特比工程学院)

AI总结 TabletCraft是首个支持美索不达米亚文字双向交互的开源系统,整合ByT5翻译模型等组件,实现阿卡德语与英语双向翻译及楔形文字渲染,在阿卡德米亚验证集上取得反向翻译的首个定量结果。

Comments 5 pages, 1 figure. Accepted to C3NLP @ ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.30580 2026-08-05 cs.CL cs.LG 版本更新

Speculative Decoding and the Curse of Multilinguality

跨语言的推测解码

Nirajan Paudel, Michael Ginn, Luc De Nardi, Alexis Palmer

机构 * University of Colorado(科罗拉多大学)

AI总结 本文研究了通过微调草稿模型或使用n-gram模型来提高非英语语言中推测解码效率的策略,发现任务特定蒸馏虽能提升效率但泛化性差,而n-gram模型尽管接受率较低,但由于生成速度快,始终能提供显著的加速效果。

Comments 15 pages, 12 figures, submitted to ACL ARR August 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01058 2026-08-05 cs.LG cs.AI cs.CL

LEAP: Layer-wise Exit-Aware Pretraining for Efficient Transformer Inference

LEAP:用于高效Transformer推理的分层退出感知预训练

Shashank Kapadia, Deep Naryan Mishra, Sujal Reddy Alugubelli, Haoan Wang, Saipraveen Vabbilisetty, Rishi Bhatia, Anupriya Sharma

机构 * Walmart Inc.(沃尔玛公司)

AI总结 LEAP通过分层退出感知预训练解决传统蒸馏与早退机制的兼容性问题,实现显著的推理加速与效率提升。

Comments Accepted at ACL 2026 (Industry Track). 14 pages, 5 figures

Journal ref Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (ACL 2026), Industry Track, pages 761-774, San Diego, California, USA

详情

展开后加载摘要…

URL PDF HTML 收藏