arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing

共收录 7868
2509.04866 2025-09-08 cs.CL

Memorization $\neq$ Understanding: Do Large Language Models Have the Ability of Scenario Cognition?

Boxiang Ma, Ru Li, Yuanlong Wang, Hongye Tan, Xiaoli Li

机构 * School of Computer and Information Technology, Shanxi University(山西大学计算机与信息学院) Information Systems Technology and Design, Singapore University of Technology and Design(新加坡科技设计大学信息系统技术与设计)

Comments EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04820 2025-09-08 cs.IR

Fishing for Answers: Exploring One-shot vs. Iterative Retrieval Strategies for Retrieval Augmented Generation

Huifeng Lin, Gang Su, Jintao Liang, You Wu, Rui Zhao, Ziyue Li

Comments under Review of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04716 2025-09-08 cs.CL cs.AI cs.IR

KERAG: Knowledge-Enhanced Retrieval-Augmented Generation for Advanced Question Answering

Yushi Sun, Kai Sun, Yifan Ethan Xu, Xiao Yang, Xin Luna Dong, Nan Tang, Lei Chen

Comments Accepted by EMNLP Findings 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03025 2025-09-08 cs.CV cs.AI

Unveiling the Response of Large Vision-Language Models to Visually Absent Tokens

Sohee Kim, Soohyun Ryu, Joonhyung Park, Eunho Yang

机构 * KAIST AI(韩国科学技术院人工智能研究所) AITRICS(人工智能与机器人技术研究所在线)

Comments accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20201 2025-09-08 cs.CL

Social Bias in Multilingual Language Models: A Survey

Lance Calvin Lim Gamboa, Yue Feng, Mark Lee

机构 * School of Computer Science, University of Birmingham(伯明翰大学计算机科学学院) Department of Information Systems and Computer Science, Ateneo de Manila University(马尼拉大学信息系统与计算机科学系)

Comments Accepted into EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15442 2025-09-08 eess.AS cs.AI cs.SD

Mitigating Hallucinations in LM-Based TTS Models via Distribution Alignment Using GFlowNets

Chenlin Liu, Minghui Fang, Patrick Zhang, Wei Zhou, Jie Gao, Jiqing Han

机构 * Harbin Institute of Technology, China(哈尔滨工业大学) Zhejiang University, China(浙江大学) Tsinghua University, Shenzhen, China(清华大学深圳研究院)

Comments Accepted to EMNLP 2025 Main Conference (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07236 2025-09-08 cs.LG cs.AI cs.CL

Simple Yet Effective: An Information-Theoretic Approach to Multi-LLM Uncertainty Quantification

Maya Kruse, Majid Afshar, Saksham Khatwani, Anoop Mayampurath, Guanhua Chen, Yanjun Gao

机构 * University of Colorado Anschutz Medical Campus(科罗拉多大学安舒茨医疗校园) University of Colorado Boulder(科罗拉多大学博尔德分校) University of Wisconsin Madison(威斯康星大学麦迪逊分校)

Comments Accepted to EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05528 2025-09-08 cs.AI cs.CL

Conversational Education at Scale: A Multi-LLM Agent Workflow for Procedural Learning and Pedagogic Quality Assessment

Jiahuan Pei, Fanghua Ye, Xin Sun, Wentao Deng, Koen Hindriks, Junxiao Wang

机构 * Vrije University of Amsterdam(阿姆斯特丹自由大学) University College London(伦敦大学学院) University of Amsterdam(阿姆斯特丹大学) National Institute of Informatics(日本信息处理学会) Shandong University(山东大学) Guangzhou University(广州大学)

Comments 14 pages, accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00008 2025-09-08 cs.AI cs.CV cs.HC

DiMo-GUI: Advancing Test-time Scaling in GUI Grounding via Modality-Aware Visual Reasoning

Hang Wu, Hongkai Chen, Yujun Cai, Chang Liu, Qingwen Ye, Ming-Hsuan Yang, Yiwei Wang

机构 * University of California, Merced(加州大学梅尔德分校) The University of Queensland(昆士兰大学) vivo Mobile Communication Co., Ltd(vivo移动通信有限公司)

Comments EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16422 2025-09-08 cs.CV

Unlocking Smarter Device Control: Foresighted Planning with a World Model-Driven Code Execution Approach

Xiaoran Yin, Xu Luo, Hao Wu, Lianli Gao, Jingkuan Song

Comments Accepted to Findings of EMNLP 2025. This is the camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08613 2025-09-08 cs.CL

Assessing the Sensitivity and Alignment of FOL Closeness Metrics

Ramya Keerthy Thatikonda, Wray Buntine, Ehsan Shareghi

机构 * Department of Data Science & AI, Monash University(数据科学与人工智能系,莫纳什大学) College of Engineering and Computer Science, VinUniversity(工程与计算机科学学院,文大学)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.18416 2025-09-08 cs.CL cs.AI cs.LG

PersonaGym: Evaluating Persona Agents and LLMs

Vinay Samuel, Henry Peng Zou, Yue Zhou, Shreyas Chaudhari, Ashwin Kalyan, Tanmay Rajpurohit, Ameet Deshpande, Karthik Narasimhan, Vishvak Murahari

机构 * University of Maryland, College Park(马里兰大学 College Park分校) University of Illinois Chicago(伊利诺伊大学芝加哥分校) University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校) Georgia Tech(佐治亚理工学院) Princeton University(普林斯顿大学)

Comments EMNLP Findings 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04657 2025-09-08 cs.CL cs.AI cs.DB cs.LG

Evaluating NL2SQL via SQL2NL

Mohammadtaher Safarzadeh, Afshin Oroojlooyjadid, Dan Roth

机构 * Oracle AI

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04602 2025-09-08 cs.CV

Sali4Vid: Saliency-Aware Video Reweighting and Adaptive Caption Retrieval for Dense Video Captioning

MinJu Jeon, Si-Woo Kim, Ye-Chan Kim, HyunGee Kim, Dong-Jin Kim

机构 * Hanyang University(翰阳大学)

Comments Accepted in EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04464 2025-09-08 cs.CL cs.AI

Can Multiple Responses from an LLM Reveal the Sources of Its Uncertainty?

Yang Nan, Pengfei He, Ravi Tandon, Han Xu

机构 * University of Arizona(亚利桑那大学) Michigan State University(密歇根州立大学)

Comments Proceedings of The 2025 Conference on Empirical Methods in Natural Language Processing (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04304 2025-09-05 cs.CL cs.AI

Facts Fade Fast: Evaluating Memorization of Outdated Medical Knowledge in Large Language Models

Juraj Vladika, Mahdi Dhaini, Florian Matthes

机构 * Technical University of Munich(慕尼黑技术大学) School of Computation, Information and Technology(计算、信息与技术学院) Department of Computer Science(计算机科学系)

Comments Accepted to Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04182 2025-09-05 cs.CL

Joint Modeling of Entities and Discourse Relations for Coherence Assessment

Wei Liu, Michael Strube

机构 * Heidelberg Institute for Theoretical Studies gGmbH(海德堡理论研究所)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04011 2025-09-05 cs.IR cs.AI cs.CL

NER Retriever: Zero-Shot Named Entity Retrieval with Type-Aware Embeddings

Or Shachar, Uri Katz, Yoav Goldberg, Oren Glickman

机构 * Computer Science Department, Bar-Ilan University(巴伊兰大学计算机科学系)

Comments Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03995 2025-09-05 cs.CL cs.AI

RTQA : Recursive Thinking for Complex Temporal Knowledge Graph Question Answering with Large Language Models

Zhaoyan Gong, Juan Li, Zhiqiang Liu, Lei Liang, Huajun Chen, Wen Zhang

机构 * Zhejiang University(浙江大学) Ant Group(蚂蚁集团) ZJU-Ant Group Joint Lab of Knowledge Graph(浙江大学-蚂蚁集团知识图谱联合实验室)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03986 2025-09-05 cs.CV cs.AI cs.CL cs.LG

Promptception: How Sensitive Are Large Multimodal Models to Prompts?

Mohamed Insaf Ismithdeen, Muhammad Uzair Khattak, Salman Khan

机构 * Mohamed Bin Zayed University of Artificial Intelligence(莫罕默德·本·扎耶德人工智能大学) Swiss Federal Institute of Technology Lausanne (EPFL)(洛桑联邦理工学院) Australian National University(澳大利亚国立大学)

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03957 2025-09-05 cs.CL cs.AI

CANDY: Benchmarking LLMs' Limitations and Assistive Potential in Chinese Misinformation Fact-Checking

Ruiling Guo, Xinwei Yang, Chen Huang, Tong Zhang, Yong Hu

机构 * School of Cyber Science and Engineering, Sichuan University, China(四川大学网络科学与工程学院) College of Computer Science, Sichuan University, China(四川大学计算机学院) Institute of Data Science, National University of Singapore, Singapore(新加坡国立大学数据科学研究所)

Comments Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01221 2025-09-05 cs.CL cs.AI cs.LG

DaMoC: Efficiently Selecting the Optimal Large Language Model for Fine-tuning Domain Tasks Based on Data and Model Compression

Wei Huang, Huang Wei, Yinggui Wang

机构 * Ant Group, China(蚂蚁集团)

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20038 2025-09-05 cs.CL

Forewarned is Forearmed: Pre-Synthesizing Jailbreak-like Instructions to Enhance LLM Safety Guardrail to Potential Attacks

Sheng Liu, Qiang Sheng, Danding Wang, Yang Li, Guang Yang, Juan Cao

机构 * Sheng Liu Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences University of Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院,中国科学院大学) Qiang Sheng Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院) Danding Wang Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院) Yang Li Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences University of Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院,中国科学院大学) Guang Yang Zhongguancun Laboratory(中关村实验室) Juan Cao Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院)

Comments EMNLP 2025 findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15478 2025-09-05 cs.CL cs.CY cs.PF

SLM-Bench: A Comprehensive Benchmark of Small Language Models on Environmental Impacts--Extended Version

Nghiem Thanh Pham, Tung Kieu, Duc-Manh Nguyen, Son Ha Xuan, Nghia Duong-Trung, Danh Le-Phuoc

机构 * FPT University(FPT大学) Aalborg University(奥尔堡大学) Technische Universität Berlin(柏林技术大学) RMIT University(皇家理工大学) German Research Center for Artificial Intelligence(德国人工智能研究中心) HiveIntel GmbH(HiveIntel公司)

Comments 24 pages. An extended version of "SLM-Bench: A Comprehensive Benchmark of Small Language Models on Environmental Impacts" accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14585 2025-09-05 cs.CL

Context Reasoner: Incentivizing Reasoning Capability for Contextualized Privacy and Safety Compliance via Reinforcement Learning

Wenbin Hu, Haoran Li, Huihao Jing, Qi Hu, Ziqian Zeng, Sirui Han, Heli Xu, Tianshu Chu, Peizhao Hu, Yangqiu Song

机构 * HKUST(香港科技大学) South China University of Technology(华南理工大学) Huawei Technologies(华为技术有限公司)

Comments Accepted to EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08046 2025-09-05 cs.IR

MultiConIR: Towards multi-condition Information Retrieval

Xuan Lu, Sifan Liu, Bochao Yin, Yongqi Li, Xinghao Chen, Hui Su, Yaohui Jin, Wenjun Zeng, Xiaoyu Shen

Comments EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14791 2025-09-05 cs.CL cs.AI cs.LG

Rapid Word Learning Through Meta In-Context Learning

Wentao Wang, Guangyuan Jiang, Tal Linzen, Brenden M. Lake

机构 * New York University(纽约大学) MIT(麻省理工学院)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12065 2025-09-05 cs.CL cs.FL

Autoformalization in the Wild: Assessing LLMs on Real-World Mathematical Definitions

Lan Zhang, Marco Valentino, Andre Freitas

机构 * Department of Computer Science, University of Manchester(曼彻斯特大学计算机科学系) School of Computer Science, University of Sheffield(谢菲尔德大学计算机科学学院) Idiap Research Institute(Idiap研究机构) National Biomarker Centre, CRUK Manchester Institute(国家生物标志物中心、CRUK曼彻斯特研究所)

Comments EMNLP 2025 Camera-Ready Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12736 2025-09-05 cs.CL cs.AI cs.LG cs.SY eess.SY math.OC

ACING: Actor-Critic for Instruction Learning in Black-Box LLMs

Salma Kharrat, Fares Fourati, Marco Canini

机构 * KAUST(卡斯土尼亚大学)

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00634 2025-09-04 cs.LG cs.AI

Efficiently Editing Mixture-of-Experts Models with Compressed Experts

Yifei He, Yang Liu, Chen Liang, Hany Hassan Awadalla

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Microsoft(微软公司)

Comments EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏