Write, Execute, Refine: From Skill Followers to Skill Optimizers via Reinforcement Learning from Execution Feedback
编写、执行、优化:通过执行反馈强化学习从技能跟随者到技能优化器
Kang Peng, Zhiwei Zhang, Yichen Zhang, Zezhong Wang, Yiming Du, Geng Tu, Baojun Wang, Bin Liang, Ruifeng Xu, Kam-Fai Wong
机构
*
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
The Chinese University of Hong Kong(香港中文大学)
;
Huawei Technologies Co., Ltd.(华为技术有限公司)
;
Harbin Institute of Technology(哈尔滨工业大学)
Analytical Provisioning for Attention-FFN Disaggregated LLM Serving under Stochastic Workloads
注意力-前馈网络解耦大语言模型服务的分析资源配置
Chendong Song, Meixuan Wang, Hang Zhou, Hong Liang, Yuan Lyu, Zixi Chen, Yuwei Fan, Zijie Zhou
机构
*
Dept. of Industrial Engineering and Decision Analytics HKUST(工业工程与决策分析系香港科技大学)
;
Dept. of Computer Science and Technology Tsinghua University(计算机科学与技术系清华大学)
;
IIIS Tsinghua University(清华大学信息学院)
;
Huawei Hong Kong Research Center(华为香港研发中心)
;
School of Mathematical Sciences Peking University(北京大学数学科学学院)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
South China University of Technology(南方科技大学)
;
Beihang University(北航)
;
Huawei Technologies Ltd(华为技术有限公司)
;
Shanghai AI Lab(上海人工智能实验室)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
InfiniVerse: Occupancy Guided Unbounded Scene Generation for Autonomous Driving
InfiniVerse: 面向自动驾驶的占用引导无界场景生成
Xiaoyu Ye, Leheng Li, Xinyu Ji, Yingjie Cai, Hongda He, Xu Yan, Guanyi Zhao, Ying-Cong Chen, Bingbing Liu, Shuguang Cui, Zhen Li
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Nanyang Technological University(南洋理工大学)
;
Huawei Technologies Ltd.(华为技术有限公司)
;
University of New South Wales(新南威尔士大学)
机构
*
School of Computer Science and Technology, East China Normal University(东华大学计算机科学与技术学院)
;
Labs, Huawei Technologies Co., LTD(华为技术有限公司2012实验室)
;
School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(上海交通大学自动化与智能感知学院)
机构
*
Huawei(华为)
;
University of Science and Technology of China(中国科学技术大学)
;
Zhejiang University(浙江大学)
;
Tsinghua University(清华大学)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东省人工智能与数字经济实验室(深圳))
Comments16 pages, 12 figures, 8 tables. Joint first authors: Yuhang Wei and Chuqin Zhou. Corresponding author: Guo Lu. To appear in Proceedings of the 34th ACM International Conference on Multimedia (MM '26), November 10-14, 2026, Rio de Janeiro, Brazil