Grounding Large Language Models as Generalizable Policies in Network Control
大语言模型作为网络优化的通用策略
Duo Wu, Linjia Kang, Zhimin Wang, Fangxin Wang, Wei Zhang, Chongbo Sun, Xuefeng Tao, Wei Yang, Le Zhang, Wenwu Zhu, Peng Cui, Zhi Wang
机构
*
Bytedance(字节跳动)
;
Shenzhen International Graduate School(深圳国际研究生院)
;
Tsinghua University(清华大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Department of Computer Science and Technology(计算机科学与技术系)
Goedel-Code-Prover: Hierarchical Proof Search for Open State-of-the-Art Code Verification
Goedel-Code-Prover:面向开放状态的最新代码验证的分层证明搜索
Zenan Li, Ziran Yang, Deyuan He, Haoyu Zhao, Andrew Zhao, Shange Tang, Kaiyu Yang, Aarti Gupta, Zhendong Su, Chi Jin
机构
*
ETH Zürich(苏黎世联邦理工学院)
;
Princeton Language and Intelligence(普林斯顿语言与智能实验室)
;
Department of Computer Science, Princeton University(普林斯顿大学计算机科学系)
;
MiroMind
Comments17 pages, 3 figures, 6 tables (9-page main text). Ancillary file hidden-automata-rl-code.zip contains reproduction code and the complete per-run data behind every table and figure. v3: author name corrected, title revised, text rewritten for clarity, new robustness checks (nonlinear and off-policy probes) added; results and conclusions unchanged