How Well Does AI-Generated Feedback Work? Intrinsic and Extrinsic Evaluation across more than 20,000 EFL Essay Drafts
人工智能生成的反馈效果如何?对20000多篇外语作文草稿的内在和外在评估
Steven Coyne, Diana Galvan-Sosa, Ryan Spring, Machi Shimmei, Michael Zock, Keisuke Sakaguchi, Kentaro Inui
机构
*
Tohoku University(东北大学)
;
RIKEN(理化学研究所)
;
ALTA Institute, Computer Laboratory, University of Cambridge(剑桥大学ALTA研究所,计算机实验室)
;
CNRS, LIS, Aix-Marseille University(法国国家科学研究中心,艾克斯-马赛大学语言信息处理实验室)
;
MBZUAI(Mohamed bin Zayed大学人工智能学院)
CommentsPre-review version of DOI https://doi.org/10.1007/978-3-032-29788-4_35, presented at AIED 2026 Late Breaking Results. Readers are encouraged to refer to the published version
MonteRET: AI Agent Enhancing Multimodal LLMs with Multi-granularity Knowledge Retrieval for Chest CT Report Generation
MonteRET:通过多粒度知识检索增强多模态大语言模型以生成胸部CT报告的人工智能代理
Yi Lin, Yihao Ding, Elana Benishay, Elefterios Trikantzopoulos, David Nauheim, Hanley Ong, Jiang Bian, Hua Xu, Yuzhe Yang, George Shih, Yifan Peng
机构
*
Weill Cornell Medicine(威尔康乃尔医学院)
;
University of Western Australia(西澳大利亚大学)
;
Indiana University(印第安纳大学)
;
Regenstrief Institute(瑞根斯特里夫研究所)
;
Yale University(耶鲁大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
Addressing Benchmarking Gaps in Large Language Models for Health and Medicine with Dynamic Red-Teaming
超越基准:动态、自动和系统化的红队代理用于可信的医疗语言模型
Jiazhen Pan, Bailiang Jian, Paul Hager, Yundi Zhang, Che Liu, Friederike Jungmann, Hongwei Bran Li, Julian Canisius, Chenyu You, Junde Wu, Jiayuan Zhu, Fenglin Liu, Yuyuan Liu, Niklas Bubeck, Moritz Knolle, Chen, Chen, Christian Wachinger, Zhenyu Gong, Cheng Ouyang, Georgios Kaissis, Benedikt Wiestler, Daniel Rueckert
机构
*
Technical University of Munich (TUM)(慕尼黑技术大学)
;
University of Oxford(牛津大学)
;
TUM University Hospital(慕尼黑技术大学医院)
;
Imperial College London(伦敦帝国理工学院)
;
Harvard Medical School(哈佛医学院)
;
Stony Brook University(史泰兹布鲁克大学)
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
;
University of Sheffield(谢菲尔德大学)
CommentsThis paper is accepted by IEEE Transactions on Dependable and Secure Computing 2025. The source code is available at \url{https://github.com/shihe98/RAG_Unlearning}
机构
*
School of Data Science, Fudan University(复旦大学数据科学学院)
;
School of Life Sciences, Beijing University of Chinese Medicine(北京中医药大学生命科学学院)
;
Institute of Science and Technology for Brain-Inspired Intelligence, Fudan University(复旦大学脑科学与智能技术研究院)
;
School of Computer Science and Technology, Huazhong University of Science and Technology(华中科技大学计算机科学与技术学院)
;
JD.com, Inc.(京东公司)
;
Key Laboratory of Computational Neuroscience and Brain-Inspired Intelligence, Fudan University, Ministry of Education(复旦大学计算神经科学与类脑智能教育部重点实验室)
;
Department of Neurology, Huashan Hospital, Fudan University(复旦大学附属华山医院神经内科)
Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction
大语言模型能写出可靠的评分标准吗?实验复现的元评估
Hanhua Hong, Yizhi Li, Jiaoyan Chen, Luu Gia Huy, Sophia Ananiadou, Jung-jae Kim, Chenghua Lin
机构
*
The University of Manchester(曼彻斯特大学)
;
Institute for Infocomm Research (I²R), A*STAR(资讯通信研究院(I²R),新加坡科技研究局)
;
IQuest Research(IQuest研究公司)
;
ELLIS Manchester(ELLIS曼彻斯特)
;
University of Information Technology, VNU(越南国家大学信息技术大学)
CommentsThis preprint has not undergone peer review or any post-submission improvements or corrections. The Version of Record of this contribution will be published in Computer Aided Systems Theory - EUROCAST 2026, Lecture Notes in Computer Science, Springer
Lost in Visual Translation: A VLM-Assisted Perceptual-Semantic Coherence Framework for EEG-to-Image Reconstruction
迷失在视觉翻译中:用于脑电到图像重建的基于视觉语言模型的感知语义连贯框架
Sukriti Tiwari, BHVSP Subrahmanyam, Nidhi Goyal, Sai Amrit Patnaik
机构
*
Mahindra University(马欣德拉大学)
;
MU-VT Interdisciplinary Advanced Research Centre for Transformative Technologies, Mahindra University(马欣德拉大学MU-VT变革性技术跨学科高级研究中心)
Operationalising Multi-Dimensional Evaluation for Conversational Agents: A Scalable, Governed Pipeline with Selective Re-evaluation and Model Benchmarking
机构
*
Department of Computer Science, George Mason University(计算机科学系,乔治·马歇尔大学)
;
Department of Engineering Science, University of South Florida(工程科学系,佛罗里达州立大学)
;
Department of Computer Science, Rutgers University(计算机科学系,罗格斯大学)
SheetMind: An End-to-End LLM-Powered Multi-Agent Framework for Spreadsheet Automation
SheetMind:一个由端到端大语言模型驱动的用于电子表格自动化的多智能体框架
Xi Cheng, Ruiyan Zhu, Ke Liu, Rakesh Chowdary Machineni, Lyuhao Chen, Brian Zhu, Daniel Jin, Zheng Qi, Neeraj Parihar, Zhoutian Xu, Oliver Gao
机构
*
Cornell University(康奈尔大学)
;
University of California, Berkeley(加州大学伯克利分校)
;
University of Michigan(密歇根大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Hong Kong University of Science and Technology (GZ)(香港科学与技术大学)