arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing

2025-10-30 至 2025-10-30 共收录 13
2510.25612 2025-10-30 cs.AI cs.MA

Counterfactual-based Agent Influence Ranker for Agentic AI Workflows

Amit Giloni, Chiara Picardi, Roy Betser, Shamik Bose, Aishvariya Priya Rathina Sabapathy, Roman Vainshtein

机构 * Fujitsu Research of Europe, UK(富士通欧洲研究中心)

Comments Accepted to EMNLP 2025, 27 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23722 2025-10-30 cs.CL

LLMs are Better Than You Think: Label-Guided In-Context Learning for Named Entity Recognition

Fan Bai, Hamid Hassanzadeh, Ardavan Saeedi, Mark Dredze

机构 * Johns Hopkins University(约翰霍普金斯大学) Optum(奥普姆)

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22586 2025-10-30 cs.CL

Precise In-Parameter Concept Erasure in Large Language Models

Yoav Gur-Arieh, Clara Suslik, Yihuai Hong, Fazl Barez, Mor Geva

机构 * Blavatnik School of Computer Science and AI, Tel Aviv University(塔尔斯基大学计算机科学与人工智能学院) New York University(纽约大学) University of Oxford & WhiteBox(牛津大学及WhiteBox)

Comments Accepted to EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17720 2025-10-30 cs.CL cs.AI

Spontaneous Giving and Calculated Greed in Language Models

Yuxuan Li, Hirokazu Shirado

机构 * School of Computer Science Carnegie Mellon University(计算机科学学院卡内基梅隆大学)

Comments Accepted to EMNLP 2025 main conference and selected as an Oral Presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25427 2025-10-30 cs.CL cs.AI

RLMEval: Evaluating Research-Level Neural Theorem Proving

Auguste Poiroux, Antoine Bosselut, Viktor Kunčak

Comments Accepted to EMNLP 2025 Findings. RLMEval benchmark released: https://github.com/augustepoiroux/RLMEval

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25007 2025-10-30 cs.AI cs.LG

Taming the Real-world Complexities in CPT E/M Coding with Large Language Models

Islam Nassar, Yang Lin, Yuan Jin, Rongxin Zhu, Chang Wei Tan, Zenan Zhai, Nitika Mathur, Thanh Tien Vu, Xu Zhong, Long Duong, Yuan-Fang Li

机构 * Oracle Health & AI(Oracle健康与人工智能)

Comments EMNLP 2025 Industry Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24754 2025-10-30 stat.ML cs.LG

Certainty in Uncertainty: Reasoning over Uncertain Knowledge Graphs with Statistical Guarantees

Yuqicheng Zhu, Jingcheng Wu, Yizhen Wang, Hongkuan Zhou, Jiaoyan Chen, Evgeny Kharlamov, Steffen Staab

机构 * University of Stuttgart(斯图加特大学) Bosch Center for AI(博世人工智能中心) The University of Manchester(曼彻斯特大学) University of Oslo(奥斯陆大学) University of Southampton(南安普顿大学)

Comments Accepted as a main conference paper at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24749 2025-10-30 cs.SE cs.AI

Beyond Function-Level Search: Repository-Aware Dual-Encoder Code Retrieval with Adversarial Verification

Aofan Liu, Shiyuan Song, Haoxuan Li, Cehao Yang, Yiyan Qi

机构 * International Digital Economy Academy (IDEA)(国际数字经济学院) School of Electronic and Computer Engineering, Peking University(电子与计算机工程学院,北京大学) Shenzhen International Graduate School, Tsinghua University(深圳国际研究生院,清华大学)

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04655 2025-10-30 cs.CL

FT-MDT: Extracting Decision Trees from Medical Texts via a Novel Low-rank Adaptation Method

Yuheng Li, Jiechao Gao, Wei Han, Wenwen Ouyang, Wei Zhu, Hui Yi Leong

机构 * Johns Hopkins University(约翰霍普金斯大学) Stanford University(斯坦福大学) Carnegie Mellon University(卡内基梅隆大学) University of Hong Kong(香港大学)

Comments Accepted by EMNLP-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01617 2025-10-30 cs.CL

AMAS: Adaptively Determining Communication Topology for LLM-based Multi-Agent System

Hui Yi Leong, Yuheng Li, Yuqing Wu, Wenwen Ouyang, Wei Zhu, Jiechao Gao, Wei Han

机构 * Johns Hopkins University(约翰霍普金斯大学) Carnegie Mellon University(卡内基梅隆大学) University of Hong Kong(香港大学) Stanford University(斯坦福大学)

Comments Accepted by EMNLP-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15063 2025-10-30 cs.CL

UrduFactCheck: An Agentic Fact-Checking Framework for Urdu with Evidence Boosting and Benchmarking

Sarfraz Ahmad, Hasan Iqbal, Momina Ahsan, Numaan Naeem, Muhammad Ahsan Riaz Khan, Arham Riaz, Muhammad Arslan Manzoor, Yuxia Wang, Preslav Nakov

Comments 15 pages, 4 figures, 5 tables, 6 Listings, Published in Proceeding of The 2025 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11832 2025-10-30 cs.CL cs.AI

OpenFactCheck: A Unified Framework for Factuality Evaluation of LLMs

Hasan Iqbal, Yuxia Wang, Minghan Wang, Georgi Georgiev, Jiahui Geng, Iryna Gurevych, Preslav Nakov

机构 * MBZUAI Monash University(墨尔本大学) Sofia University(索菲亚大学)

Comments 11 pages, 4 Figures, 3 Tables, Published In Proceedings of The 2024 Conference on Empirical Methods in Natural Language Processing

Journal ref In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing: System Demonstrations, pages 219-229, Miami, Florida, USA. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07222 2025-10-30 cs.CL cs.AI cs.LG

Reliable Evaluation and Benchmarks for Statement Autoformalization

Auguste Poiroux, Gail Weiss, Viktor Kunčak, Antoine Bosselut

Comments Accepted to EMNLP 2025. New benchmarks released, see https://github.com/augustepoiroux/RLMEval , https://huggingface.co/datasets/PAug/ProofNetSharp , and https://huggingface.co/datasets/PAug/ProofNetVerif . For code, see https://github.com/augustepoiroux/LeanInteract

详情

展开后加载摘要…

URL PDF HTML 收藏