机构
*
University of Pittsburgh(匹兹堡大学)
;
Johns Hopkins University(约翰霍普金斯大学)
;
University of Notre Dame(诺特丹大学)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
University of Washington(华盛顿大学)
;
Allen Institute for Artificial Intelligence(人工智能研究院)
;
University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
CommentsAccepted by Transactions on Machine Learning Research (TMLR 2024)
Journal refTransactions on Machine Learning Research (TMLR), 2024; Presented at The Thirteenth International Conference on Learning Representations (ICLR 2025), Singapore
Rethinking On-policy Optimization for Query Augmentation
重新思考查询增强的在线策略优化
Zhichao Xu, Shengyao Zhuang, Xueguang Ma, Bingsen Chen, Yijun Tian, Fengran Mo, Tao Li, Jie Cao, Vivek Srikumar
机构
*
University of Utah(犹他大学)
;
The University of Queensland(昆士兰大学)
;
University of Waterloo(滑铁卢大学)
;
New York University(纽约大学)
;
University of Notre Dame(圣母大学)
;
Université de Montréal(蒙特利尔大学)
;
Google DeepMind(谷歌DeepMind)
;
University of Oklahoma(俄克拉荷马大学)
机构
*
Nara Institute of Science and Technology(奈良先端科学技术大学院大学)
;
Kyoto University(京都大学)
;
Ateneo de Manila University(马尼拉雅典耀大学)
;
UNI-President Information Philippines Corporation(统一信息菲律宾公司)
;
The Chinese University of Hong Kong(香港中文大学)
Learning Robust Penetration Testing Policies under Partial Observability: A systematic evaluation
学习部分可观测下的鲁棒渗透测试策略:系统评估
Raphael Simon, Pieter Libin, Wim Mees
机构
*
Cyber Defence Lab, CISS Department Royal Military Academy(国防网络安全实验室,信息与系统科学系皇家军事学院)
;
AI Lab, Department of Computer Science Vrije Universiteit Brussel(人工智能实验室,计算机科学系自由大学布鲁塞尔)
Laurens Samson, Nimrod Barazani, Sennay Ghebreab, Yuki M. Asano
机构
*
Socially-Intelligent Artificial Systems Group, University of Amsterdam(智能社会人工智能系统组,阿姆斯特丹大学)
;
University of Amsterdam(阿姆斯特丹大学)
;
Fundamental AI Lab, University of Technology Nuremberg(基础人工智能实验室,纽伦堡技术大学)
机构
*
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Pennsylvania(宾夕法尼亚大学)
;
Okinawa Institute of Science and Technology(冲绳科学技术大学院大学)
DP-FedSOFIM: Differentially Private Federated Stochastic Optimization using Regularized Fisher Information Matrix
DP-FedSOFIM: 使用正则化Fisher信息矩阵的差分隐私联邦随机优化
Sidhant Nair, Tanmay Sen, Mrinmay Sen, Sayantan Banerjee
机构
*
Department of Mechanical Engineering, Indian Institute of Technology Delhi(印度理工学院德里机械工程系)
;
SQC & OR Unit, Indian Statistical Institute Kolkata(印度统计研究院科钦SQC与OR单位)
;
Department of Artificial Intelligence, Indian Institute of Technology Hyderabad(印度理工学院海得拉巴人工智能系)
;
Operations Management & Quantitative Techniques Area, Indian Institute of Management, Indore(印度管理学院印地尔运营管理和定量技术领域)
机构
*
University of Southern California(南加州大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Mohamed Bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
CommentsProceedings of the 39th Annual Conference on Neural Information Processing Systems, ARLET Workshop (Aligning Reinforcement Learning Experimentalists and Theorists)
Journal refTransactions on Machine Learning Research, Vol. 2026, June 2026
FORGE: Foundational Optimization Representations from Graph Embeddings
FORGE:基于图嵌入的基础优化表示
Zohair Shafi, Serdar Kadioglu
机构
*
Khoury College of Computer Science Northeastern University(诺埃弗大学计算机科学学院)
;
AI Center of Excellence, Fidelity Investments(富达投资人工智能卓越中心)
;
Department of Computer Science, Brown University(布朗大学计算机科学系)