机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Zhongguancun Academy(中关村科学院)
Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling
长度值模型:用于令牌级长度建模的可扩展值预训练
Zhen Zhang, Changyi Yang, Zijie Xia, Zhen Yang, Chengzhi Liu, Zhaotiao Weng, Yepeng Liu, Haobo Chen, Jin Pan, Chenyang Zhao, Yuheng Bu, Alkesh Patel, Zhe Gan, Xin Eric Wang
机构
*
University of California, Santa Barbara(加州大学圣巴巴拉分校)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
LMSYS Org(LMSYS组织)
;
Apple Inc.(苹果公司)
机构
*
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
National University of Singapore(新加坡国立大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Hong Kong Polytechnic University(香港理工大学)
机构
*
Beihang University(北京航空航天大学)
;
State Key Laboratory of Space Information System and Integrated Application(空间信息系统与集成应用国家重点实验室)
;
Nanjing University(南京大学)
;
Harbin Engineering University(哈尔滨工程大学)
;
Beijing Zhongguancun Academy(北京中关村科学城)
机构
*
Department of Landscape Architecture and Rural Systems Engineering, Seoul National University, Seoul, South Korea(景观建筑与乡村系统工程系,首尔国立大学,首尔,韩国)
;
Interdisciplinary Program in Landscape Architecture, Seoul National University, South Korea(景观建筑跨学科项目,首尔国立大学,韩国)
;
Integrated Major in Smart City Global Convergence, Seoul National University, Seoul, Republic of Korea(智能城市全球融合整合专业,首尔国立大学,首尔,韩国)
;
Research Institute of Agriculture and Life Sciences, Seoul National University, Seoul, Republic of Korea(农业与生命科学研究院,首尔国立大学,首尔,韩国)
机构
*
Nanjing University(南京大学)
;
Siemens Data and AI Research(西门子数据与人工智能研究院)
;
Nanjing University – Siemens Joint Research Center on Industrial AI(南京大学-西门子工业人工智能联合研究中心)
Mitigating Errors in LLM-Generated Web API Invocations via Retrieval-Augmented Generation and Constrained Decoding
通过检索增强生成和约束解码减轻大语言模型生成的Web API调用中的错误
Daniel Maninger, Leon Chemnitz, Jannis Brugger, Tushar Lamba, Amir Molzam Sharifloo, Mira Mezini
机构
*
Technische Universität Darmstadt(德累斯顿技术大学)
;
Hessian Center for Artificial Intelligence (hessian.AI)(黑森人工智能中心)
;
Pariton AI
;
National Research Center for Applied Cybersecurity ATHENE(应用网络安全国家研究中心ATHENE)
PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies
面向通用视觉-语言-动作策略的通用姿态预训练
Haitao Lin, Hanyang Yu, Jingshun Huang, He Zhang, Yonggen Ling, Ping Tan, Xiangyang Xue, Yanwei Fu
机构
*
Tencent Robotics X(腾讯机器人X)
;
Futian Laboratory(福田实验室)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Fudan University(复旦大学)
;
Shanghai Innovation Institute(上海创新研究院)
机构
*
Spatial Design Intelligence Lab, BitInf Ltd.(BitInf有限公司空间设计智能实验室)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
;
College of Computer and Information Engineering, Nanjing Tech University(南京工业大学计算机与信息工程学院)
A Controlled Counterexample to Strong Proxy-Based Explanations of OOD Performance: in a Fixed Pretraining-and-Probing Setup
对强代理基于解释的OOD性能的受控反例:在固定预训练和探测设置中
Hongmin Li
机构
*
School of Life Science and Technology, Institute of Science Tokyo(生命科学与技术学院,科学东京研究所)
;
Department of Computational Biology and Medical Sciences, Graduate School of Frontier Sciences(计算生物学与医学科学系,前沿科学研究生院)