arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

RAG / 检索增强生成

检索增强生成、向量检索、知识库问答和面向大模型的搜索系统。

2026-01-29 至 2026-01-29 共收录 4 信号源:cs.IR, cs.CL, cs.AI, cs.DB

1. RAG评测 4 篇

2601.19927 2026-01-29 cs.CL 85%

Attribution Techniques for Mitigating Hallucinated Information in RAG Systems: A Survey

缓解RAG系统中幻觉信息的归因技术:综述

Yuqing Zhao, Ziyao Liu, Yongsen Zheng, Kwok-Yan Lam

机构 * Nanyang Technological University(南洋理工大学)

专题命中 RAG评测 :RAG(title,abstract);retrieval-augmented generation(abstract);retriever(abstract);分类 cs.CL

AI总结 本文综述了RAG系统中缓解幻觉的归因技术,通过分类幻觉类型、统一流程和比较优劣,为实际应用提供指导。

Journal ref The 8th International Conference on Artifcial Intelligence in Information and Communication (ICAIIC 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16033 2026-01-29 cs.HC cs.AI 70%

Helping Johnny Make Sense of Privacy Policies with LLMs

帮助约翰尼理解隐私政策的LLMs

Vincent Freiberger, Arthur Fleig, Erik Buchmann

机构 * Center for Scalable Data Analytics and Artificial Intelligence (ScaDS.AI) Dresden/Leipzig, Leipzig University(可扩展数据分析与人工智能中心(ScaDS.AI)德累斯顿/莱比锡,莱比锡大学)

专题命中 RAG评测 :retrieval-augmented generation(abstract);RAG(abstract);分类 cs.AI

AI总结 PRISMe通过结合LLM的政策评估与交互式界面,帮助用户更有效地理解和参与隐私政策,同时揭示了设计和技术上的挑战。

Comments 21 pages, 3 figures, 3 tables, ACM CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10258 2026-01-29 cs.CV cs.MM 67%

XY-Cut++: Advanced Layout Ordering via Hierarchical Mask Mechanism on a Novel Benchmark

XY-Cut++: 一种基于新型基准的分层遮罩机制的高级布局排序方法

Shuai Liu, Youmeng Li, Jizeng Wei

机构 * College of Intelligence and Computing, Tianjin University(智能与计算学院,天津大学)

专题命中 RAG评测 :retrieval-augmented generation(abstract);RAG(abstract)

AI总结 XY-Cut++通过分层遮罩机制在新型基准上实现高级布局排序,提升文档阅读顺序恢复的准确性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20276 2026-01-29 cs.CL cs.AI 62%

Beyond the Needle's Illusion: Decoupled Evaluation of Evidence Access and Use under Semantic Interference at 326M-Token Scale

超越针的错觉:在32600万token规模下解耦证据访问和使用下的语义干扰评估

Tianwei Lin, Zuyi Zhou, Xinda Zhao, Chenke Wang, Xiaohong Li, Yu Chen, Chuanrui Hu, Jian Pei, Yafeng Deng

机构 * EverMind Shanda Group(Shanda集团) Duke University(杜克大学)

专题命中 RAG评测 :RAG(abstract);分类 cs.CL、cs.AI

AI总结 本文提出EMB-S基准,用于评估长上下文模型在大规模语义干扰下的证据访问与使用能力,揭示语义辨别是主要瓶颈。

详情

展开后加载摘要…

URL PDF HTML 收藏