arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing

共收录 7862
2510.01845 2025-10-03 cs.CL cs.CV

Model Merging to Maintain Language-Only Performance in Developmentally Plausible Multimodal Models

Ece Takmaz, Lisa Bylinina, Jakub Dotlacil

机构 * Utrecht University(乌特勒支大学)

Comments Accepted to the EMNLP 2025 workshop BabyLM: Accelerating language modeling research with cognitively plausible datasets

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01831 2025-10-03 cs.CL

Syntactic Blind Spots: How Misalignment Leads to LLMs Mathematical Errors

Dane Williamson, Yangfeng Ji, Matthew Dwyer

机构 * Department of Computer Science(计算机科学系) University of Virginia(弗吉尼亚大学)

Comments 14 pages, 5 Tables, 9 Figures; Accepted to MathNLP 2025: The 3rd Workshop on Mathematical Natural Language Processing (co-located with EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01659 2025-10-03 cs.CL cs.AI

MDSEval: A Meta-Evaluation Benchmark for Multimodal Dialogue Summarization

Yinhong Liu, Jianfeng He, Hang Su, Ruixue Lian, Yi Nian, Jake Vincent, Srikanth Vishnubhotla, Robinson Piramuthu, Saab Mansour

机构 * AWS AI Labs(AWS人工智能实验室) Language Technology Lab, University of Cambridge(语言技术实验室,剑桥大学)

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01638 2025-10-03 cs.HC cs.AI

Towards Human-Centered RegTech: Unpacking Professionals' Strategies and Needs for Using LLMs Safely

Siying Hu, Yaxing Yao, Zhicong Lu

机构 * City University of Hong Kong(香港城市大学) Johns Hopkins University(约翰霍普金斯大学) George Mason University(乔治·马歇尔大学)

Comments Accepted to the 4th HCI+NLP@EMNLP 2025 Workshop. (Non-archival)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01631 2025-10-03 cs.LG cs.AI cs.CL

Demystifying Synthetic Data in LLM Pre-training: A Systematic Study of Scaling Laws, Benefits, and Pitfalls

Feiyang Kang, Newsha Ardalani, Michael Kuchnik, Youssef Emad, Mostafa Elhoushi, Shubhabrata Sengupta, Shang-Wen Li, Ramya Raghavendra, Ruoxi Jia, Carole-Jean Wu

机构 * FAIR at Meta(Meta 的 FAIR)

Comments Published as a Main Conference paper at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01526 2025-10-03 cs.CL q-fin.CP

One More Question is Enough, Expert Question Decomposition (EQD) Model for Domain Quantitative Reasoning

Mengyu Wang, Sotirios Sabanis, Miguel de Carvalho, Shay B. Cohen, Tiejun Ma

机构 * The University of Edinburgh(爱丁堡大学) National Technical University of Athens(雅典技术大学) Archimedes/Athena Research Centre(阿基米德/雅典研究中心) University of Aveiro(阿维罗大学)

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01270 2025-10-03 cs.CL cs.AI

Think Twice, Generate Once: Safeguarding by Progressive Self-Reflection

Hoang Phan, Victor Li, Qi Lei

机构 * New York University(纽约大学)

Comments Accepted to EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01247 2025-10-03 cs.CL cs.AI

Let's Play Across Cultures: A Large Multilingual, Multicultural Benchmark for Assessing Language Models' Understanding of Sports

Punit Kumar Singh, Nishant Kumar, Akash Ghosh, Kunal Pasad, Khushi Soni, Manisha Jaishwal, Sriparna Saha, Syukron Abu Ishaq Alfarozi, Asres Temam Abagissa, Kitsuchart Pasupa, Haiqin Yang, Jose G Moreno

机构 * Indian Institute of Technology Patna(印度理工学院帕纳布分校) Sardar Patel Institute of Technology(萨达尔·帕特尔技术学院) Universitas Gadjah Mada(加查马大学) King Mongkut’s Institute of Technology Ladkrabang(拉差班国王技术学院) Shenzhen Technology University(深圳技术大学) Université de Toulouse(图卢兹大学)

Comments 52 pages, 56 figures; appearing at EMNLP'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01239 2025-10-03 cs.CL

CIFLEX: Contextual Instruction Flow for Sub-task Execution in Multi-Turn Interactions with a Single On-Device LLM

Juntae Lee, Jihwan Bang, Seunghan Yang, Simyung Chang

Comments accepted at EMNLP 2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01232 2025-10-03 cs.CL cs.AI

Benchmark Profiling: Mechanistic Diagnosis of LLM Benchmarks

Dongjun Kim, Gyuho Shim, Yongchan Chun, Minhyuk Kim, Chanjun Park, Heuiseok Lim

机构 * Department of Computer Science and Engineering, Korea University(韩国大学计算机科学与工程系) School of Software, Soongsil University(顺天大学软件学院)

Comments 16 pages, 5 figures. Accepted to EMNLP 2025 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04032 2025-10-03 cs.CL cs.LG

What if I ask in \textit{alia lingua}? Measuring Functional Similarity Across Languages

Debangan Mishra, Arihant Rastogi, Agyeya Negi, Shashwat Goel, Ponnurangam Kumaraguru

机构 * IIIT Hyderabad(海得拉巴印度理工学院) ELLIS Institute Tübingen(图宾根ELLIS研究所)

Comments Accepted into Multilingual Representation Learning (MRL) Workshop at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18379 2025-10-03 cs.IR

REALM: Recursive Relevance Modeling for LLM-based Document Re-Ranking

Pinhuan Wang, Zhiqiu Xia, Chunhua Liao, Feiyi Wang, Hang Liu

Comments EMNLP 2025 (Main Conference, Oral). 15 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04782 2025-10-03 cs.CL cs.LG

Reason to Rote: Rethinking Memorization in Reasoning

Yupei Du, Philipp Mondorf, Silvia Casola, Yuekun Yao, Robert Litschko, Barbara Plank

机构 * Department of ICS, Utrecht University(信息科学与计算系,乌特勒支大学) MaiNLP, Center for Information and Language Processing, LMU Munich(MaiNLP,信息与语言处理中心,慕尼黑大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) Saarland Informatics Campus, Saarland University(萨尔兰州信息科学校区,萨尔兰州大学)

Comments EMNLP 2025 Main. 21 pages, 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04512 2025-10-03 cs.AI

Schema Generation for Large Knowledge Graphs Using Large Language Models

Bohui Zhang, Yuan He, Lydia Pintscher, Albert Meroño Peñuela, Elena Simperl

机构 * King’s College London(伦敦国王学院) University of Oxford(牛津大学) Wikimedia Deutschland(维基媒体德国) Technical University of Munich(慕尼黑技术大学)

Comments EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24858 2025-10-03 cs.CL cs.LG

MetaFaith: Faithful Natural Language Uncertainty Expression in LLMs

Gabrielle Kaili-May Liu, Gal Yona, Avi Caciularu, Idan Szpektor, Tim G. J. Rudner, Arman Cohan

机构 * Yale University(耶鲁大学) Google Research(谷歌研究) University of Toronto(多伦多大学)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24683 2025-10-03 cs.CL cs.AI

Should I Share this Translation? Evaluating Quality Feedback for User Reliance on Machine Translation

Dayeon Ki, Kevin Duh, Marine Carpuat

机构 * University of Maryland(马里兰大学) Johns Hopkins University(约翰霍普金斯大学)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21315 2025-10-03 cs.CL

Charting the Landscape of African NLP: Mapping Progress and Shaping the Road Ahead

Jesujoba O. Alabi, Michael A. Hedderich, David Ifeoluwa Adelani, Dietrich Klakow

机构 * Saarland University, Saarland Informatics Campus(萨尔兰大学,萨尔兰信息学校区) LMU Munich and Munich Center for Machine Learning(慕尼黑大学及慕尼黑机器学习中心) Mila - Quebec AI Institute, McGill University & Canada CIFAR AI Chair(魁北克人工智能研究所,麦吉尔大学及加拿大 CIFAR 人工智能 chair)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19430 2025-10-03 cs.CL cs.AI

Deriving Strategic Market Insights with Large Language Models: A Benchmark for Forward Counterfactual Generation

Keane Ong, Rui Mao, Deeksha Varshney, Paul Pu Liang, Erik Cambria, Gianmarco Mengaldo

Comments Published at Empirical Methods in Natural Language Processing 2025 (Main Conference) (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12051 2025-10-03 cs.CL

TLUE: A Tibetan Language Understanding Evaluation Benchmark

Fan Gao, Cheng Huang, Nyima Tashi, Xiangxiang Wang, Thupten Tsering, Ban Ma-bao, Renzeg Duojie, Gadeng Luosang, Rinchen Dongrub, Dorje Tashi, Hao Wang Xiao Feng, Yongbin Yu

机构 * University of Electronic Science and Technology of China(电子科技大学) Tibet University(西藏大学) University of Texas Southwestern Medical Center(德克萨斯西南医学中心) Southern Methodist University(南方 Methodist 大学) The State Key Laboratory of Tibetan Intelligence(藏语智能国家重点实验室) University of Connecticut(康涅狄格大学)

Comments Accepted by EMNLP Main Conference (Poster)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14862 2025-10-03 cs.CL cs.AI cs.IR

Interpretable Text Embeddings and Text Similarity Explanation: A Survey

Juri Opitz, Lucas Möller, Andrianos Michail, Sebastian Padó, Simon Clematide

机构 * University of Zurich(苏黎世大学) University of Stuttgart(斯图加特大学)

Comments EMNLP 2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10582 2025-10-03 cs.CL cs.HC

Adapting Large Language Models for Character-based Augmentative and Alternative Communication

Dylan Gaines, Keith Vertanen

机构 * Dylan Gaines Department of Computer Science, Michigan Technological University, Houghton, MI, USA(达西·甘斯 计算机科学系,密歇根技术大学,霍顿,MI,美国) Keith Vertanen(凯斯·维塔内)

Comments To appear in Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11305 2025-10-03 cs.LG cs.AI

QSpec: Speculative Decoding with Complementary Quantization Schemes

Juntao Zhao, Wenhao Lu, Sheng Wang, Lingpeng Kong, Chuan Wu

Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01178 2025-10-02 cs.LG cs.AI

COM-BOM: Bayesian Exemplar Search for Efficiently Exploring the Accuracy-Calibration Pareto Frontier

Gaoxiang Luo, Aryan Deshwal

机构 * University of Minnesota(明尼苏达大学)

Comments Accepted by EMNLP 2025 Main, Code: https://github.com/GaoxiangLuo/COM-BOM

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01165 2025-10-02 cs.CL cs.AI cs.LG

GRAD: Generative Retrieval-Aligned Demonstration Sampler for Efficient Few-Shot Reasoning

Oussama Gabouj, Kamel Charaf, Ivan Zakazov, Nicolas Baldwin, Robert West

机构 * EPFL(苏黎世联邦理工学院)

Comments EMNLP 2025 (findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04408 2025-10-02 cs.CL cs.AI

Unpacking Let Alone: Human-Scale Models Generalize to a Rare Construction in Form but not Meaning

Wesley Scivetti, Tatsuya Aoyama, Ethan Wilcox, Nathan Schneider

机构 * Georgetown University(杰克逊维尔大学)

Comments Empirical Methods for Natural Language Processing (EMNLP) 2025, Camera-Ready Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12950 2025-10-02 cs.CL

GuRE:Generative Query REwriter for Legal Passage Retrieval

Daehee Kim, Deokhyung Kang, Jonghwi Kim, Sangwon Ryu, Gary Geunbae Lee

机构 * Graduate School of Artificial Intelligence, POSTECH, Republic of Korea(人工智能研究生院,POSTECH,大韩民国) AI Future Lab, KT, Republic of Korea(AI未来实验室,KT,大韩民国) Department of Computer Science and Engineering, POSTECH, Republic of Korea(计算机科学与工程系,POSTECH,大韩民国)

Comments NLLP Workshop at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00962 2025-10-02 cs.CL

Analyzing Dialectical Biases in LLMs for Knowledge and Reasoning Benchmarks

Eileen Pan, Anna Seo Gyeong Choi, Maartje ter Hoeve, Skyler Seto, Allison Koenecke

机构 * Department of Information Science, Cornell University(康奈尔大学信息科学系) Apple(苹果公司) Cornell Tech(康奈尔科技)

Comments EMNLP Findings 2025, 12 pages, 11 tables, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00808 2025-10-02 cs.CV cs.AI cs.CL

What You See is What You Ask: Evaluating Audio Descriptions

Divy Kala, Eshika Khandelwal, Makarand Tapaswi

机构 * CVIT, IIIT Hyderabad(计算机视觉与人工智能技术研究所,IIIT海得拉巴)

Comments EMNLP 2025 Main Track Long Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00662 2025-10-02 cs.CL cs.AI

Facilitating Cognitive Accessibility with LLMs: A Multi-Task Approach to Easy-to-Read Text Generation

François Ledoyen, Gaël Dias, Jeremie Pantin, Alexis Lechervy, Fabrice Maurel, Youssef Chahir

机构 * Université Caen Normandie, ENSICAEN, CNRS, Normandie Univ, GREYC UMR 6072(法国卡恩大学、ENSICAEN、CNRS、诺曼底大学、GREYC UMR 6072) Koena SAS

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00567 2025-10-02 cs.CL

Are Large Language Models Chronically Online Surfers? A Dataset for Chinese Internet Meme Explanation

Yubo Xie, Chenkai Wang, Zongyang Ma, Fahui Miao

机构 * Shanghai Maritime University(上海 Maritime 大学) École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院) Xi’an Jiaotong Liverpool University(西安交通大学利物浦大学)

Comments Accepted to EMNLP 2025 Main Conference. 22 pages, 3 figures, 13 tables. GitHub: github.com/yuboxie/chime

详情

展开后加载摘要…

URL PDF HTML 收藏