CommentsThis is the author's version of the work. The definitive version is published in: Proceedings of the 48th European Conference on Information Retrieval (ECIR '26), March 29-April 2, 2026, Delft, The Netherlands
From RAG to Agentic: Validating Islamic-Medicine Responses with LLM Agents
Mohammad Amaan Sayeed, Mohammed Talha Alam, Raza Imam, Shahab Saquib Sohail, Amir Hussain
机构
*
Mohamed bin Zayed University of Artificial Intelligence, UAE(阿布扎克穆罕默德·本·扎耶德人工智能大学)
;
Edinburgh Napier University, UK(爱丁堡纳皮尔大学)
;
VIT Bhopal University, India(比哈尔大学)
SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design
Wenxin Tang, Jingyu Xiao, Wenxuan Jiang, Xi Xiao, Yuhang Wang, Xuxin Tang, Qing Li, Yuehe Ma, Junliang Liu, Shisong Tang, Michael R. Lyu
机构
*
Tsinghua University(清华大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Northeastern University(东北大学)
;
Southwest University(西南大学)
;
Kuaishou Technology(快手科技)
;
Peng Cheng Laboratory(鹏城实验室)
;
BNU-HKBU United International College(北京师范大学-香港 Baptist University联合国际学院)
;
Dalian Maritime University(大连海事大学)
Even Small Reasoners Should Quote Their Sources: Introducing the Pleias-RAG Model Family
Pierre-Carl Langlais, Pavel Chizhov, Mattia Nee, Carlos Rosas Hinostroza, Matthieu Delsart, Irène Girard, Othman Hicheur, Anastasia Stasenko, Ivan P. Yamshchikov
Comments9 pages. Published in the Proceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR 2026)
Journal refProceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR '26), pp. 3464-3472, 2026
Agentic and Generative AI for Open-Source Intelligence and Cyber Investigations: Taxonomy, Evaluation, Challenges, and Future Directions
用于开源情报和网络调查的智能与生成式人工智能:分类法、评估、挑战及未来方向
Eduardo Almeida Palmieri, Mohamed Chahine Ghanem, Dipo Dunsin, Zubair Baig, Ed de Quincey, Kim-Kwang Raymond Choo
机构
*
School of Computer Science and Mathematics, Keele University(基尔大学计算机科学与数学学院)
;
Cybersecurity Institute, University of Liverpool(利物浦大学网络安全研究所)
;
Department of Applied Computing IICL, University of Wales Trinity Saint David(威尔士特里尼达大学应用计算系)
;
Deakin Cyber Research and Innovation Hub, Deakin University(德金大学网络安全研究与创新中心)
;
Department of Information Systems and Cyber Security, The University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校信息系与网络安全系)
IRC-Bench: Recognizing Entities from Contextual Cues in First-Person Reminiscences
IRC-Bench: 从第一人称回忆中的上下文线索识别实体
Yehudit Aperstein, Eden Moran, Alexander Apartsin
机构
*
Intelligent Systems, Afeka Academic College of Engineering(阿法卡学术工程学院智能系统)
;
School of Computer Science, Faculty of Sciences, Holon Institute of Technology(霍隆理工学院计算机科学学院)
DoGMaTiQ: Automated Generation of Question-and-Answer Nuggets for Report Evaluation
DoGMaTiQ:面向报告评估的问答片段自动生成
Bryan Li, William Walden, Yu Hou, Gabrielle Kaili-May Liu, Dawn Lawrie, James Mayfield, Eugene Yang, Chris Callison-Burch, Laura Dietz
机构
*
Google Inc.(谷歌公司)
;
Johns Hopkins University(约翰霍普金斯大学)
;
University of Maryland(马里兰大学)
;
Yale University(耶鲁大学)
;
University of Pennsylvania(宾夕法尼亚大学)
;
University of New Hampshire(新罕布什尔大学)
机构
*
Department of Computer Science and Engineering, Bangladesh University of Engineering and Technology(Bangladesh University of Engineering and Technology计算机科学与工程系)
Journal refProceedings of the Fifteenth Language Resources and Evaluation Conference (LREC 2026), pp. 10457-10466, ELRA, Palma, Mallorca, Spain, May 2026
SkMTEB: Slovak Massive Text Embedding Benchmark and Model Adaptation
SkMTEB:斯洛伐克大规模文本嵌入基准与模型适配
Marek Šuppa, Andrej Ridzik, Daniel Hládek, Natália Kňažeková, Viktória Ondrejová
机构
*
Comenius University in Bratislava(布拉迪斯拉发夸美纽斯大学)
;
Cisco Systems(思科系统)
;
Technical University of Košice(科希策技术大学)
;
Kempelen Institute of Intelligent Technologies(肯佩伦智能技术研究所)
CommentsIn Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining V.2 (KDD '26). 30 pages: 11 pages in main (6 figures, 1 table), 19 pages in appendix (22 figures, 2 tables)
Beyond Static Dialogues: Benchmarking Realistic, Heterogeneous, and Evolving Long-Term Memory
超越静态对话:对现实、异构和演化长期记忆的基准测试
Han Zhang, Zihao Tang, Xin Yu, Xiao Liu, Yeyun Gong, Haizhen Huang, Yan Lu, Weiwei Deng, Feng Sun, Qi Zhang, Hanfang Yang
机构
*
Center for Applied Statistics, Renmin University of China(中国人民大学应用统计中心)
;
School of Statistics, Renmin University of China(中国人民大学统计学院)
;
Microsoft(微软公司)
Overview of the MedHopQA track at BioCreative IX: track description, participation and evaluation of systems for multi-hop medical question answering
BioCreative IX MedHopQA 轨道概述:轨道描述、参与及多跳医学问答系统评估
Rezarta Islamaj, Joey Chan, Robert Leaman, Jongmyung Jung, Hyeongsoon Hwang, Quoc-An Nguyen, Hoang-Quynh Le, Harikrishnan Gurushankar Saisudha, Ganesh Chandrasekar, Rustam R. Taktashov, Nadezhda Yu. Bizyukova, Sofia I. R. Conceição, Paulo R. C. Lopes, Reem Abdel Salam, Mary Adewunmi, Zhiyong Lu
机构
*
National Library of Medicine (NLM), National Institutes of Health (NIH)(美国国家医学图书馆(NLM)、国家卫生研究院(NIH))
;
University of Illinois at Urbana Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Korea University(韩国大学)
;
VNU University of Engineering and Technology, Hanoi, Vietnam(越南河内工程大学)
;
Concordia University, Montreal, QC, CA(蒙特利尔大学)
;
Institute of Biomedical Chemistry (IBMC), 10 bld. 8, Pogodinskaya str., 119121 Moscow, Russia(俄罗斯生物医学化学研究所(IBMC))
;
LASIGE, Departamento de Informática, Faculdade de Ciências, Universidade de Lisboa, 1749-016 Lisbon, Portugal(葡萄牙里斯本大学 LASIGE 实验室)
;
Faculty of Engineering, Computer Engineering Department Cairo University(埃及开罗大学工程学院)
;
Menzies School of Health Research, Charles Darwin University, NT, Australia(澳大利亚查尔斯达尔文大学梅恩兹健康研究中心)
;
CaresAI, Australia(澳大利亚 CaresAI)
Agentic clinical reasoning over longitudinal myeloma records: a retrospective evaluation against expert consensus
基于纵向骨髓瘤记录的代理临床推理:一项回顾性评估与专家共识的对比
Johannes Moll, Jannik Lübberstedt, Christoph Nuernbergk, Jacob Stroh, Luisa Mertens, Anna Purcarea, Christopher Zirn, Zeineb Benchaaben, Fabian Drexel, Hartmut Häntze, Anirudh Narayanan, Friedrich Puttkammer, Andrei Zhukov, Jacqueline Lammert, Sebastian Ziegelmayer, Markus Graf, Marion Högner, Marcus Makowski, Florian Bassermann, Lisa C. Adams, Jiazhen Pan, Daniel Rueckert, Krischan Braitsch, Keno K. Bressem
机构
*
Chair for AI in Healthcare and Medicine, Technical University of Munich (TUM) and TUM University Hospital(人工智能在医疗与健康中的研究所,慕尼黑技术大学(TUM)及慕尼黑技术大学医院)
;
Department of Diagnostic and Interventional Radiology, Klinikum rechts der Isar, TUM University Hospital, School of Medicine and Health, Technical University of Munich(诊断与介入放射科,莱茵河右岸医院,慕尼黑技术大学医院,医学与健康学院,慕尼黑技术大学)
;
Department of Cardiovascular Radiology and Nuclear Medicine, German Heart Center, TUM University Hospital, School of Medicine and Health, Technical University of Munich(心血管放射学与核医学科,德国心脏中心,慕尼黑技术大学医院,医学与健康学院,慕尼黑技术大学)
;
Department of Medicine III, Klinikum rechts der Isar, TUM University Hospital, School of Medicine and Health, Technical University of Munich(第三医学部,莱茵河右岸医院,慕尼黑技术大学医院,医学与健康学院,慕尼黑技术大学)
CUB: Benchmarking Context Utilisation Techniques for Language Models
CUB:语言模型上下文利用技术的基准测试
Lovisa Hagström, Youna Kim, Haeun Yu, Sang-goo Lee, Richard Johansson, Hyunsoo Cho, Isabelle Augenstein
机构
*
Chalmers University of Technology(查尔姆斯理工大学)
;
University of Gothenburg(哥德堡大学)
;
Seoul National University(首尔国立大学)
;
University of Copenhagen(哥本哈根大学)
;
Ewha Womans University(成均馆大学)