arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing

共收录 7861
2507.07229 2025-11-04 cs.CL

SynthTextEval: Synthetic Text Data Generation and Evaluation for High-Stakes Domains

Krithika Ramesh, Daniel Smolyak, Zihao Zhao, Nupoor Gandhi, Ritu Agarwal, Margrét Bjarnadóttir, Anjalie Field

机构 * Johns Hopkins University(约翰霍普金斯大学) University of Maryland, College Park(马里兰大学学院公园分校) Carnegie Mellon University(卡内基梅隆大学)

Comments EMNLP 2025 System Demonstration

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17784 2025-11-04 cs.AI

AnyMAC: Cascading Flexible Multi-Agent Collaboration via Next-Agent Prediction

Song Wang, Zhen Tan, Zihan Chen, Shuang Zhou, Tianlong Chen, Jundong Li

机构 * University of Virginia(弗吉尼亚大学) Arizona State University(亚利桑那州立大学) University of Minnesota Twin Cities(明尼苏达大学双城分校) University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)

Comments EMNLP Main 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07801 2025-11-04 cs.CL cs.AI cs.LG

MultiMatch: Multihead Consistency Regularization Matching for Semi-Supervised Text Classification

Iustin Sirbu, Robert-Adrian Popovici, Cornelia Caragea, Stefan Trausan-Matu, Traian Rebedea

机构 * National University of Science and Technology POLITEHNICA Bucharest(波兰技术大学科学与技术国立大学) University of Illinois Chicago(伊利诺伊大学芝加哥分校) NVIDIA(NVIDIA公司) Renius Technologies(Renius技术公司)

Comments This is the camera-ready version of the paper, accepted for publication in the Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23945 2025-11-04 cs.CL cs.AI

A Closer Look at Bias and Chain-of-Thought Faithfulness of Large (Vision) Language Models

Sriram Balasubramanian, Samyadeep Basu, Soheil Feizi

机构 * Department of Computer Science University of Maryland, College Park(计算机科学系马里兰大学 College Park)

Comments Accepted in EMNLP 2025, 34 pages, 25 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17267 2025-11-04 cs.CL

GreekBarBench: A Challenging Benchmark for Free-Text Legal Reasoning and Citations

Odysseas S. Chlapanis, Dimitrios Galanis, Nikolaos Aletras, Ion Androutsopoulos

机构 * Department of Informatics, Athens University of Economics and Business(信息学院,雅典经济与商业大学) Archimedes, Athena Research Center(阿提卡研究中心-阿基米德) Athena Research Center(阿提卡研究中心) University of Sheffield(谢菲尔德大学)

Comments 19 pages, 17 figures, accepted in EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14393 2025-11-04 cs.CL

Editing Across Languages: A Survey of Multilingual Knowledge Editing

Nadir Durrani, Basel Mousi, Fahim Dalvi

机构 * Qatar Computing Research Institute, HBKU(卡塔尔计算研究所,哈姆扎大学)

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07769 2025-11-04 cs.CV

BiMediX2: Bio-Medical EXpert LMM for Diverse Medical Modalities

Sahal Shaji Mullappilly, Mohammed Irfan Kurpath, Sara Pieri, Saeed Yahya Alseiari, Shanavas Cholakkal, Khaled Aldahmani, Fahad Khan, Rao Anwer, Salman Khan, Timothy Baldwin, Hisham Cholakkal

机构 * Mohamed Bin Zayed University of Artificial Intelligence(Mohamed Bin Zayed大学人工智能学院) Linköping University(林肯大学) Shaikh Tahnoon bin Mohammed Medical City(Shaikh Tahnoon bin Mohammed医疗城) Tawam Hospital(Tawam医院) Sheikh Shakhbout Medical City(Sheikh Shakhbout医疗城) Govt Medical College Kozhikode(科钦政府医学院)

Comments Accepted to EMNLP 2025 (Findings)

Journal ref Findings of the Association for Computational Linguistics: EMNLP 2025, pages 14051-14071

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10584 2025-11-04 cs.AI cs.LG cs.MA

STACKFEED: Structured Textual Actor-Critic Knowledge Base Editing with FeedBack

Shashank Kirtania, Naman Gupta, Priyanshu Gupta, Krishna Kariya, Sumit Gulwani, Arun Iyer, Suresh Parthasarathy, Arjun Radhakrishna, Sriram K. Rajamani, Gustavo Soares

机构 * Microsoft(微软) Microsoft Research India(微软印度研究院)

Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing: Industry Track 2588-2606

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07129 2025-11-04 cs.CL cs.AI

Exploring Large Language Models for Detecting Mental Disorders

Gleb Kuzmin, Petr Strepetov, Maksim Stankevich, Natalia Chudova, Artem Shelmanov, Ivan Smirnov

机构 * AIRI ISP RAS Research Center for Trusted Artificial Intelligence(俄罗斯科学院信息与系统研究信任人工智能研究中心) FRC CSC RAS(俄罗斯科学院应用数学与计算机科学研究所) MIPT(米哈伊尔·伊万诺维奇·普京高等经济学院) RUDN University(俄罗斯人民友谊大学) MBZUAI(马斯喀特大学人工智能研究所)

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18771 2025-11-04 cs.CL

CheckEval: A reliable LLM-as-a-Judge framework for evaluating text generation using checklists

Yukyung Lee, Joonghoon Kim, Jaehee Kim, Hyowon Cho, Jaewook Kang, Pilsung Kang, Najoung Kim

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27672 2025-11-03 cs.CL

Culture Cartography: Mapping the Landscape of Cultural Knowledge

Caleb Ziems, William Held, Jane Yu, Amir Goldberg, David Grusky, Diyi Yang

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27532 2025-11-03 cs.CL

SQLSpace: A Representation Space for Text-to-SQL to Discover and Mitigate Robustness Gaps

Neha Srikanth, Victor Bursztyn, Puneet Mathur, Ani Nenkova

机构 * University of Maryland(马里兰大学) Adobe Research(Adobe研究)

Comments Accepted to EMNLP Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27287 2025-11-03 cs.LG cs.AI

Can LLMs Help You at Work? A Sandbox for Evaluating LLM Agents in Enterprise Environments

Harsh Vishwakarma, Ankush Agarwal, Ojas Patil, Chaitanya Devaguptapu, Mahesh Chandran

机构 * Fujitsu Research(富士通研究所)

Comments Accepted at EMNLP 2025 Main Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27196 2025-11-03 cs.CL cs.AI

MemeArena: Automating Context-Aware Unbiased Evaluation of Harmfulness Understanding for Multimodal Large Language Models

Zixin Chen, Hongzhan Lin, Kaixin Li, Ziyang Luo, Yayue Deng, Jing Ma

机构 * Hong Kong Baptist University(香港 Baptist 大学) Beijing University of Posts and Telecommunications(北京邮电大学) National University of Singapore(新加坡国立大学)

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26606 2025-11-03 cs.AI cs.CL

Normative Reasoning in Large Language Models: A Comparative Benchmark from Logical and Modal Perspectives

Kentaro Ozeki, Risako Ando, Takanobu Morishita, Hirohiko Abe, Koji Mineshima, Mitsuhiro Okada

机构 * Keio University(Keio大学) University of Tokyo(东京大学)

Comments Accepted to the 8th BlackboxNLP Workshop at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25761 2025-11-03 cs.CL

DiagramEval: Evaluating LLM-Generated Diagrams via Graphs

Chumeng Liang, Jiaxuan You

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments EMNLP 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03419 2025-11-03 cs.CL

Curse of Knowledge: When Complex Evaluation Context Benefits yet Biases LLM Judges

Weiyuan Li, Xintao Wang, Siyu Yuan, Rui Xu, Jiangjie Chen, Qingqing Dong, Yanghua Xiao, Deqing Yang

机构 * School of Data Science, Fudan University, 2 Shanghai Key Laboratory of Data Science 3 College of Computer Science and Artificial Intelligence, Fudan University 4 ByteDance Seed, 5 College of Cryptology and Cyber Science, Nankai University(1 数据科学学院,复旦大学 2 上海数据科学实验室 3 计算机科学与人工智能学院,复旦大学 4 字节跳动种子 5 密码学与网络科学学院,南开大学)

Comments EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11664 2025-11-03 cs.AI

VRoPE: Rotary Position Embedding for Video Large Language Models

Zikang Liu, Longteng Guo, Yepeng Tang, Tongtian Yue, Junxian Cai, Kai Ma, Qingbin Liu, Xi Chen, Jing Liu

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) School of Computer Science and Technology, Beijing Jiaotong University(北京交通大学计算机科学与技术学院) Basic Algorithm Center, Tencent(腾讯基础算法中心)

Comments EMNLP 2025 Main Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.05102 2025-11-03 cs.CL cs.AI cs.LG

SparsePO: Controlling Preference Alignment of LLMs via Sparse Token Masks

Fenia Christopoulou, Ronald Cardenas, Gerasimos Lampouras, Haitham Bou-Ammar, Jun Wang

机构 * Huawei Noah’s Ark Lab(华为诺亚实验室) University College London(伦敦大学学院)

Comments 27 pages, 9 figures, 5 tables. Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26799 2025-10-31 cs.CV

Masked Diffusion Captioning for Visual Feature Learning

Chao Feng, Zihao Wei, Andrew Owens

机构 * University of Michigan(密歇根大学) Cornell University(康奈尔大学) University of Maryland(马里兰大学)

Comments EMNLP 2025 (Findings). Project page: https://cfeng16.github.io/mdlm4vfl/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14681 2025-10-31 cs.CL

Massive Supervised Fine-tuning Experiments Reveal How Data, Layer, and Training Factors Shape LLM Alignment Quality

Yuto Harada, Yusuke Yamauchi, Yusuke Oda, Yohei Oseki, Yusuke Miyao, Yu Takagi

机构 * NII LLMC(日本信息处理学会大语言模型中心) The University of Tokyo(东京大学) NAIST(日本科学技术大学) Nagoya Institute of Technology(名古屋技术大学)

Comments Accepted to EMNLP 2025 (Main Conference). Models and evaluation results available at: https://github.com/llm-jp/massive-sft

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09391 2025-10-31 cs.CL

Comparing human and LLM politeness strategies in free production

Haoran Zhao, Robert D. Hawkins

机构 * Department of Linguistics University of Washington(语言学系华盛顿大学) Department of Linguistics Stanford University(语言学系斯坦福大学)

Comments 25 pages, 5 figures | EMNLP 2025 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26446 2025-10-31 cs.CL

1+1>2: A Synergistic Sparse and Low-Rank Compression Method for Large Language Models

Zeliang Zong, Kai Zhang, Zheyang Li, Wenming Tan, Ye Ren, Yiyan Zhai, Jilin Hu

机构 * Hikvision Research Institute(海康威视研究院) School of Data Science and Engineering, East China Normal University(华东师范大学数据科学与工程学院)

Comments 15 pages, 6 figures, EMNLP 2025 findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26193 2025-10-31 cs.CL

RCScore: Quantifying Response Consistency in Large Language Models

Dongjun Jang, Youngchae Ahn, Hyopil Shin

Journal ref EMNLP 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26157 2025-10-31 cs.LG cs.AI

Bridging the Gap Between Molecule and Textual Descriptions via Substructure-aware Alignment

Hyuntae Park, Yeachan Kim, SangKeun Lee

Comments EMNLP 2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26006 2025-10-31 cs.CV cs.CL

CAVE: Detecting and Explaining Commonsense Anomalies in Visual Environments

Rishika Bhagwatkar, Syrielle Montariol, Angelika Romanou, Beatriz Borges, Irina Rish, Antoine Bosselut

机构 * EPFL(瑞士联邦理工学院) MILA(蒙特利尔人工智能研究院)

Journal ref 2025 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25623 2025-10-31 cs.CL

Evaluating the Role of Verifiers in Test-Time Scaling for Legal Reasoning Tasks

Davide Romano, Jonathan Schwarz, Daniele Giofré

机构 * Thomson Reuters(汤姆森·路透)

Comments Accepted to EMNLP - NLLP Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24134 2025-10-31 cs.CV cs.AI cs.CL

VC4VG: Optimizing Video Captions for Text-to-Video Generation

Yang Du, Zhuoran Lin, Kaiqiang Song, Biao Wang, Zhicheng Zheng, Tiezheng Ge, Bo Zheng, Qin Jin

机构 * School of Information, Renmin University of China(中国人民大学信息学院) Taobao & Tmall Group of Alibaba(阿里巴巴淘宝与天猫集团)

Comments Accepted by EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04448 2025-10-31 cs.CV cs.MM

TRUST-VL: An Explainable News Assistant for General Multimodal Misinformation Detection

Zehong Yan, Peng Qi, Wynne Hsu, Mong Li Lee

Comments EMNLP 2025 Oral; Project Homepage: https://yanzehong.github.io/trust-vl/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08410 2025-10-31 cs.CL

Large Language Models Have Intrinsic Meta-Cognition, but Need a Good Lens

Ziyang Ma, Qingyue Yuan, Zhenglin Wang, Deyu Zhou

机构 * School of Computer Science and Engineering, Key Laboratory of Computer Network and Information Integration, Ministry of Education, Southeast University, China(计算机科学与工程学院、计算机网络与信息集成重点实验室、教育部、东南大学,中国) Department of Neurosurgery, Shanghai Tenth People’s Hospital, School of Clinical Medicine of Nanjing Medical University, China(神经外科科、上海第十人民医院、南京医科大学临床医学学院,中国)

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏