arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Annual Meeting of the Association for Computational Linguistics · 会议 · Natural Language Processing

2026-07-13 至 2026-07-13 共收录 11
2607.09263 2026-07-13 cs.CV 新提交

Semantic Hardness Is Not Visual Hardness: Sign-Aware Hard Negative Mining for Sign Language Retrieval

语义难度并非视觉难度:用于手语检索的符号感知硬负样本挖掘

Junmyeong Lee, Chan Hur, ChangSu Choi, Sukmin Cho, Fitsum Gaim, Eui Jun Hwang, Hoyun Song, KyungTae Lim

机构 * School of Computing(计算机学院) Graduate School of Culture Technology(文化技术研究生院) Korea Advanced Institute of Science and Technology(韩国科学技术院) ETRI Medical Informatics Laboratory(电子通信研究院医学信息学实验室)

AI总结 研究手语检索在细粒度场景的问题,提出符号感知硬负样本挖掘方法,通过在嵌入空间基于视觉易混淆性构建硬负样本,实验证明该方法能提升细粒度检索性能且保持粗粒度准确性。

Comments Accepted to ACL 2026 main

Journal ref Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 2026, pages 28262-28277

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.07740 2026-07-13 cs.LG cs.AI 新提交

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE

Jet-Long:使用动态双焦点旋转位置编码的高效长上下文扩展

Haozhan Tang, Zerui Wang, Yuxian Gu, Song Han, Han Cai

机构 * NVIDIA(英伟达)

AI总结 研究针对现代语言模型长上下文应用中零样本上下文扩展问题,提出Jet-Long方法,通过动态双焦点RoPE及相关技术,在推理时开销小,在多种模型和基准测试中表现优异,还能推广到其他架构且超参数弹性强。

Comments added discussion of AdaGroPE and LaMPE (Findings of ACL 2026) with clarified contribution

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.16966 2026-07-13 cs.CR cs.AI

Visual Inception: Compromising Long-term Planning in Agentic Recommenders via Multimodal Memory Poisoning

视觉 inception:通过多模态记忆污染在代理推荐系统中妥协长期规划

Jiachen Qian

机构 * City University of Hong Kong(香港城市大学)

AI总结 本文提出视觉 inception 攻击,通过污染用户上传的图片在代理推荐系统中影响长期规划,提出 CognitiveGuard 防御框架以降低攻击风险。

Comments 17 pages, 6 figures, 16 tables

Journal ref Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 20846-20862, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.16515 2026-07-13 cs.CV cs.CR cs.LG

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

精打细算,却愚弄像素:通过视觉对抗扰动在多模态智能体中绕过价格限制

Jiachen Qian, Zhaolu Kang

机构 * City University of Hong Kong(香港城市大学) Peking University(北京大学)

AI总结 研究发现多模态大语言模型在价格受限环境下易受视觉误导,提出PriceBlind攻击框架,通过语义解耦损失提升图像嵌入精度,实现80%的攻击成功率,并展示鲁棒编码和防御机制对攻击的抑制效果。

Comments 15 pages, 4 figures, 13 tables

Journal ref Findings of the Association for Computational Linguistics: ACL 2026, pages 16059-16073, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09097 2026-07-13 cs.AI 版本更新

Programming over Thinking: Efficient and Robust Multi-Constraint Planning

编程而非思考:高效且稳健的多约束规划

Derrick Goh Xin Deik, Quanyu Long, Zhengyuan Liu, Nancy F. Chen, Wenya Wang

机构 * Nanyang Technological University(南洋理工大学) Agency for Science, Technology and Research(科技研究局)

AI总结 多约束规划面临挑战,现有大语言模型方法有局限。本文提出SCOPE框架,将特定查询推理与通用代码执行分离,产生一致、可重用的求解器函数,降低成本和延迟,性能达最优。

Comments Accepted at ACL 2026 : Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12731 2026-07-13 cs.CL cs.AI

A Shared Geometry of Difficulty in Multilingual Language Models

多语言语言模型中的共同难度几何

Stefano Civelli, Pietro Bernardelle, Nicolò Brunello, Gianluca Demartini

机构 * The University of Queensland(昆士兰大学) Polytechnic University of Milan(米兰理工大学)

AI总结 本研究发现多语言语言模型中问题难度的表示分为两个阶段:浅层表示具有跨语言泛化能力,而深层表示在同语言性能高但跨语言泛化差。

Journal ref Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04476 2026-07-13 cs.CL 版本更新

Probabilistic Textual Time Series Depression Detection

概率性文本时间序列抑郁检测

Fabian Schmidt, Seyedehmoniba Ravan, Vladimir Vlassov

机构 * Department of Computer Science, KTH Royal Institute of Technology(计算机科学系,皇家理工学院) Department of Information Technology, Uppsala University(信息科技系,乌普萨拉大学)

AI总结 研究旨在实现准确且可解释的抑郁严重程度预测,提出PTTSD概率框架,结合多种技术预测PHQ - 8分数,在相关数据集上性能具竞争力,消融实验及分析突出了方法中各部分价值与临床相关性。

Comments 16 pages, 6 figures, 7 tables

Journal ref Findings of the Association for Computational Linguistics, ACL 2026, pages 32574-32589, July 2-7, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03595 2026-07-13 cs.CL 版本更新

Decoupling Task-Solving and Output Formatting in LLM Generation

在大语言模型生成中解耦任务解决与输出格式

Haikang Deng, Po-Nien Kung, Nanyun Peng

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

AI总结 研究针对大语言模型任务指令与格式要求交织致性能下降的问题,提出解耦框架Deco-G,将格式遵循委托给单独模块,引入指令感知蒸馏等三项创新,实验证明其优于基线方法,能保证格式符合。

Comments Update to the latest ACL published version and add a link to the released code

Journal ref Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), Association for Computational Linguistics, 2026, pp. 16764-16781

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15806 2026-07-13 cs.CL cs.AI 版本更新

REAL: REtrieval-reAsoning and Logic-constructed Attention Behaviors for Long-Context KV Cache Compression

REAL:用于长上下文KV缓存压缩的检索推理和逻辑构建注意力行为

Mengjie Li, Yuan Feng, Xike Xie, William J. Song

AI总结 针对大语言模型KV缓存问题,受混淆矩阵启发提出REAL方法,通过注意力行为矩阵全面分析注意力头行为,最大化信噪比,经评估在多模型和基准测试中性能显著,为长上下文建模方法转变提供新思路。

Comments Accepted at ACL 2026 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18043 2026-07-13 cs.CL cs.AI cs.CV 版本更新

GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs

GrAInS:基于梯度的大语言模型和视觉语言模型推理时引导方法

Duy Nguyen, Archiki Prasad, Elias Stengel-Eskin, Mohit Bansal

机构 * UNC Chapel Hill(北卡罗来纳大学教堂山分校) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

AI总结 研究针对推理时引导LLMs和VLMs的方法局限,提出GrAInS。该方法基于梯度归因识别关键令牌构建引导向量,推理时调整激活。实验显示其性能优于微调及现有基线,能在保持模型流畅性和通用能力的同时提升多项指标。

Comments Accepted to ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12446 2026-07-13 cs.CL cs.AI cs.LG 版本更新

Multi-Attribute Steering of Language Models via Targeted Intervention

通过目标干预对语言模型进行多属性引导

Duy Nguyen, Archiki Prasad, Elias Stengel-Eskin, Mohit Bansal

机构 * UNC Chapel Hill(北卡罗来纳大学教堂山分校)

AI总结 研究如何对语言模型进行多属性引导,提出MAT-Steer框架,通过对齐目标学习引导向量,减少属性间冲突,在问答和生成任务中评估,该框架优于现有方法。

Comments ACL 2025 camera-ready, code link: https://github.com/duykhuongnguyen/MAT-Steer

详情

展开后加载摘要…

URL PDF HTML 收藏