arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing

2025-11-24 至 2025-11-24 共收录 3
2510.12194 2025-11-24 cs.AI

ResearStudio: A Human-Intervenable Framework for Building Controllable Deep-Research Agents

ResearStudio: 一种可人工干预的构建可控深度研究代理的框架

Linyi Yang, Yixuan Weng

机构 * Southern University of Science and Technology(南方科技大学)

AI总结 ResearStudio是一种可人工干预的深度研究代理框架,通过实时人类控制与自动化相结合,在GAIA基准中取得最佳性能。

Comments EMNLP 2025 Demo, Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13804 2025-11-24 cs.CL cs.HC

Beyond Human Judgment: A Bayesian Evaluation of LLMs' Moral Values Understanding

超越人类判断:对LLMs道德价值观理解的贝叶斯评估

Maciej Skorski, Alina Landowska

机构 * University of Luxembourg(卢森堡大学) Kozminski University(科兹明斯基大学) SWPS University(SWPS大学)

AI总结 本文通过贝叶斯方法评估了LLMs在道德价值观理解上的表现,发现AI模型在准确率和敏感性方面优于人类标注者。

Comments Appears in UncertaiNLP@EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01424 2025-11-24 cs.AI cs.CL

From Hypothesis to Publication: A Comprehensive Survey of AI-Driven Research Support Systems

从假设到发表:人工智能驱动研究支持系统的全面综述

Zekun Zhou, Xiaocheng Feng, Lei Huang, Xiachong Feng, Ziyun Song, Ruihan Chen, Liang Zhao, Weitao Ma, Yuxuan Gu, Baoxin Wang, Dayong Wu, Guoping Hu, Ting Liu, Bing Qin

机构 * Harbin Institute of Technology(哈尔滨工业大学) Peng Cheng Laboratory(鹏城实验室) The University of Hong Kong(香港大学) iFLYTEK Research(iFLYTEK研究院)

AI总结 本文综述了人工智能驱动的研究支持系统,涵盖假设生成、验证及论文发表,分析了当前挑战与未来方向,并提供了相关工具和基准的全面概述。

Comments Accepted to EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏