发表机构
University of Information Technology; Vietnam National University, Ho Chi Minh city(信息技术大学; 胡志明市越南国家大学)
机构由 AI 辅助整理,请以论文原文为准。AI 中文总结
针对越南法律咨询领域基准缺失的问题,构建了基于真实咨询的大规模基准ViLegalExpert,涵盖172K问题与34个法律领域,实验显示混合检索最优,为可靠越南法律AI提供挑战性基准。
AI 中文摘要
可信赖的法律人工智能要求系统能够在回答法律问题的同时,将其回答基于权威来源。然而,现有的越南法律基准对真实世界法律咨询的覆盖有限。我们提出了ViLegalExpert,一个基于真实的公民与律师咨询构建的大规模基准,包含超过34个法律领域的172K个问题,并配有专业回答和经专家验证的法律证据。ViLegalExpert支持法律信息检索、抽取式问答和生成式问答。使用代表性检索方法和语言模型的实验揭示了在证据检索和基于证据的回答生成方面的重大挑战。虽然预训练模型在问答上表现强劲,但混合检索取得了最佳的检索性能。这些结果表明,将自然表达的法律问题映射到权威条款存在困难,并确立了ViLegalExpert作为可靠越南法律人工智能的一个具有挑战性的基准。
英文摘要
Trustworthy Legal AI requires systems that can answer legal questions while grounding their responses in authoritative sources. However, existing Vietnamese legal benchmarks provide limited coverage of real-world legal consultations. We introduce \textbf{ViLegalExpert}, a large-scale benchmark constructed from authentic citizen--lawyer consultations, containing over \textbf{172K} questions across \textbf{34 legal domains}, together with professional answers and expert-verified legal evidence. ViLegalExpert supports legal information retrieval, extractive QA, and abstractive QA. Experiments with representative retrieval methods and language models reveal substantial challenges in evidence retrieval and grounded answer generation. While pretrained models perform strongly on QA, hybrid retrieval achieves the best retrieval performance. These results demonstrate the difficulty of mapping naturally expressed legal questions to authoritative provisions and establish ViLegalExpert as a challenging benchmark for reliable Vietnamese Legal AI.