XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations
XR-1:通过学习统一的视觉-运动表示实现多功能的视觉-语言-动作模型
Shichao Fan, Kun Wu, Zhengping Che, Xinhua Wang, Di Wu, Fei Liao, Ning Liu, Yixue Zhang, Zhen Zhao, Zhiyuan Xu, Meng Li, Qingjie Liu, Shanghang Zhang, Min Wan, Jian Tang
机构
*
Beijing Innovation Center of Humanoid Robotics, Beijing, China(北京人形机器人创新中心,北京,中国)
;
School of Mechanical Engineering and Automation, Beihang University, Beijing, China(北京航空航天大学机械工程及自动化学院,北京,中国)
;
State Key Laboratory of Virtual Reality Technology and Systems, SCSE, Beihang University, Beijing, China(虚拟现实技术与系统国家重点实验室,SCSE,北京航空航天大学,北京,中国)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University, Beijing, China(多媒体信息处理国家重点实验室,计算机科学学院,北京大学,北京,中国)
Retrieval-Augmented Detection of Potentially Abusive Clauses in Chilean Terms of Service
智利服务条款中潜在滥用条款的检索增强检测
Christoffer Loeffler, Tomás Rey Pizarro, Daniel Ignacio Miranda Vásquez, Andrea Martínez Freile
机构
*
School of Computer Engineering, Pontificia Universidad Católica de Valparaíso(Pontificia Universidad Católica de Valparaíso计算机工程学院)
;
Faculty of Law, Universidad Adolfo Ibáñez(Adolfo Ibáñez大学法学院)
Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency
可预测的编造:大型语言模型的事实回忆能力随模型大小和主题频率而增加
Matthew L. Smith, Jonathan P. Shock, Samuel T. Segun, Iyiola E. Olatunji, Tegawendé F. Bissyandé
机构
*
International Development Research Centre Canada(加拿大国际发展研究中心)
;
University of Cape Town(开普敦大学)
;
Global Center on AI Governance(人工智能治理全球中心)
;
SnT, University of Luxembourg(卢森堡大学SnT分校)
;
CITADEL AI Centre of Excellence, Burkina Faso(布基纳法索CITADEL人工智能卓越中心)
专题命中
预训练与数据
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
EntropyScan: Towards Model-level Backdoor Detection in LVLMs via Visual Attention Entropy
EntropyScan: 向通过视觉注意力熵实现LVLMs的模型级后门检测
Xuanyu Ge, Zhongqi Wang, Jie Zhang, Shiguang Shan, Xilin Chen
机构
*
China University of Geosciences(中国地质大学)
;
University of the Chinese Academy of Sciences(中国科学院大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
专题命中
预训练与数据
:LLM(abstract);large language model(abstract);language model(abstract)
When and Why SignSGD Outperforms SGD: A Theoretical Study Based on $\ell_1$-norm Lower Bounds
何时以及为何SignSGD优于SGD:基于ℓ1范数下界的一个理论研究
Hongyi Tao, Dingzhi Yu, Lijun Zhang
机构
*
State Key Laboratory of Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
EmbodiedHead: Real-Time Listening and Speaking Avatar for Conversational Agents
EmbodiedHead:面向对话代理的实时听与说的虚拟形象
Yu Zhang, Kaiyuan Shen, Yang Li
机构
*
School of Computer Science and Technology, East China Normal University, Shanghai, China(东华大学计算机科学与技术学院,上海,中国)
;
Garabido Shanghai Technology Co., Ltd., Shanghai, China(Garabido上海科技有限公司,上海,中国)
Comments\c{opyright} 2026 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works
Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training
数据混合代理:学习重新加权领域以实现持续预训练
Kailai Yang, Xiao Liu, Lei Ji, Hao Li, Xiao Liang, Zhiwei Liu, Yeyun Gong, Peng Cheng, Mao Yang
机构
*
The University of Manchester(曼彻斯特大学)
;
Microsoft Research(微软研究院)
;
Imperial College London(伦敦帝国学院)
;
University of California, Los Angeles(加利福尼亚大学洛杉矶分校)
专题命中
预训练与数据
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG