Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification
推理时的验证扩展:通过测试时的评分指南验证实现自我进化深度研究代理
Yuxuan Wan, Tianqing Fang, Zaitang Li, Yintong Huo, Wenxuan Wang, Haitao Mi, Dong Yu, Michael R. Lyu
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Tencent AI Lab(腾讯人工智能实验室)
;
Singapore Management University(新加坡管理学院)
;
The Renmin University of China(中国人民大学)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
A Transformer-Based Cross-Platform Analysis of Public Discourse on the 15-Minute City Paradigm
基于Transformer的跨平台公共 discourse 对15分钟城市范式的分析
Gaurab Chhetri, Darrell Anderson, Boniphace Kutela, Subasish Das
机构
*
1 College of Science
;
Engineering, Texas State University, San Marcos, Texas, USA Email
;
3 Texas A\&M Transportation Institute, Texas A\&M University, Houston, Texas, USA Email
CommentsThis is the author's preprint version of a paper accepted for presentation at the 24th International Conference on Machine Learning and Applications (ICMLA 2025), December 3-5, 2025, Florida, USA. The final published version will appear in the official IEEE proceedings. Conference site: https://www.icmla-conference.org/icmla25/
机构
*
Department of XXX, University of YYY, Location, Country(XXX部门,YYYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
;
Inria (Flowers), University of Bordeaux, France(Inria(Flowers),波尔多大学,法国)
;
MIT, Computational Cognitive Science Lab, Cambridge, MA, USA(麻省理工学院,计算认知科学实验室,马萨诸塞州剑桥,美国)
Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies
结合大语言模型和领域本体的透明网络威胁情报赋能
Luca Cotti, Anisa Rula, Devis Bianchini, Federico Cerutti
机构
*
Department of Information Engineering, University of Brescia, Italy(布雷西亚大学信息工程系)
;
School of Computer Science and Informatics, Cardiff University, United Kingdom(卡迪夫大学计算机科学与信息学学院)
;
Department of Electronics and Computer Science, University of Southampton, United Kingdom(南安普顿大学电子与计算机科学系)
Comments14 pages, 3 figures, 6 tables, presented at the 1st workshops on eXplainable AI, Knowledge Representation and Knowledge Graphs (XAI-KRKG) and User-Centered Explanations in XAI Workshop (UCEX-XAI), October 25-30, 2025, Bologna, Italy
Journal refJoint Proceedings of the 1st workshops on eXplainable AI, Knowledge Representation and Knowledge Graphs (XAI-KRKG) and User-Centered Explanations in XAI Workshop (UCEX-XAI), 2025, CEUR Workshop proceedings, volume 4172, pages 95-108
On Benchmark Hacking in ML Contests: Modeling, Insights and Design
在机器学习竞赛中基准黑客现象:建模、洞察与设计
Xiaoyun Qiu, Yang Yu, Haifeng Xu
机构
*
Department of Economics, Dartmouth College(达特茅斯学院经济系)
;
Sloan & CSAIL, Massachusetts Institute of Technology(麻省理工学院斯隆管理学院与计算机科学与人工智能实验室)
;
Department of Computer Science, University of Chicago(芝加哥大学计算机科学系)
Architectures for Robust Self-Organizing Energy Systems under Information and Control Constraints
面向信息与控制约束的鲁棒自组织能源系统架构
Emilie Frost, Astrid Nieße
机构
*
Department of Computing Science, Carl von Ossietzky Universität Oldenburg(奥尔登堡卡尔·冯·奥西特定律大学计算机科学系)
;
Distributed Artificial Intelligence, OFFIS - Institute for Information Technology(分布式人工智能,OFFIS信息技术研究所)
CommentsThis preprint has not undergone peer review (when applicable) or any post-submission improvements or corrections. The Version of Record of this contribution will be published in Agents and Artificial Intelligence, Lecture Notes in Computer Science, and available online at https://doi.org/10.1007/978-3-032-25029-2_19
机构
*
State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(中国科学技术大学认知智能国家重点实验室)
;
Princeton University(普林斯顿大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Infinite Intelligence Pharma(无限智能制药)
Brief chatbot interactions produce lasting changes in human moral values
简短的聊天机器人互动会产生持久的人类道德价值观变化
Yue Teng, Qianer Zhong, Kim Mai Tich Nguyen Thordsen, Christian Montag, Benjamin Becker
机构
*
Department of Psychology, The University of Hong Kong(香港大学心理学系)
;
Techno-Entrepreneurship Core, The University of Hong Kong(香港大学技术企业家精神核心)
;
Department of Psychology, University of Copenhagen(哥本哈根大学心理学系)
;
Centre for Cognitive and Brain Sciences, Institute of Collaborative Innovation, University of Macau(澳门大学协同创新研究院认知与脑科学中心)
;
Department of Psychology, Faculty of Social Sciences, University of Macau(澳门大学社会科学学院心理学系)
;
Department of Computer and Information Science, Faculty of Science and Technology, University of Macau(澳门大学科技学院计算机与信息科学系)
;
SRT AI, Society & Social Dynamics, Faculty of Social Sciences, The University of Hong Kong(SRT AI,社会与社会动态,香港大学社会科学学院)
机构
*
School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院)
;
Institute of Artificial Intelligence and Future Networks, Beijing Normal University(北京师范大学人工智能与未来网络研究院)
;
Faculty of Arts and Sciences, Beijing Normal University(北京师范大学文理学院)
;
Beijing Normal-Hong Kong Baptist University(北京师范大学-香港 Baptist大学)
SynRXN: An Open Benchmark and Curated Dataset for Computational Reaction Modeling
SynRXN:计算反应建模的开放基准和精心编纂的数据集
Tieu-Long Phan, Nhu-Ngoc Nguyen Song, Peter F. Stadler
机构
*
Bioinformatics Group, Department of Computer Science \& Interdisciplinary Center for Bioinformatics \& School for Embedded
;
Composite Artificial Intelligence (SECAI), Leipzig University, H \"a rtelstra e 16–18, D-04107 Leipzig, Germany
;
School of Pharmacy, University of Medicine