CommentsRevised version prepared in response to the editorial assessment. The main manuscript is 32 pages; detailed mathematical derivations have been moved to the accompanying Supplementary Material. The principal results and contributions are strengthened and clarified
Automatic, Debiased, and Invariant Counterfactual Generation under General Interventions
通用干预下的自动、去偏和不变反事实生成
Raphael C Kim, Jingsen Zhu, Ramin Zabih, Michele Santacatterina
机构
*
Cornell Tech(康奈尔科技)
;
Cornell University(康奈尔大学)
;
Department of Biostatistics, Department of Population Health(生物统计学系、人口健康系)
;
New York University Grossman School of Medicine(纽约大学格罗斯曼医学院)
机构
*
Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学)
;
School of Robotics and Advanced Manufacture, Harbin Institute of Technology(机器人与先进制造学院,哈尔滨工业大学)
;
School of Computation, Information and Technology, Technical University of Munich(计算、信息与技术学院,慕尼黑技术大学)
;
College of Computer Science & Visual Computing and Intelligent Perception Lab, Nankai University(计算机科学与视觉计算学院及智能感知实验室,南开大学)
;
School of Cyber Science and Technology, Sun Yat-sen University(网络科学与技术学院,中山大学)
;
School of Automation, Southeast University(自动化学院,东南大学)
Comments8 pages, 5 figures. Introduces ForeTime-VLA, a causal future-token distillation method for conveyor-belt manipulation from a frozen world action model teacher
Ludi${}_{\scriptscriptstyle 0.1}$: An Agentic System for Socially Intelligent Robots
Ludi₀.₁:面向社交智能机器人的智能体系统
Wooseong Chung, William Cong, Jakub Dworakowski, Ethan Ewer, Tri Wahyu Guntara, Yeonwoo Jeong, Tianchong Jiang, Chaewon Kim, Hyunseo Kim, Jinwoo Kim, Jinyeon Kim, Yea-Seul Kim, Jack Kunde, Kangwook Lee, Sangheon Lee, Robert Nowak, Junha Roh
机构
*
Ludo Robotics(乐动机器人)
专题命中
机器人基础模型
:manipulation(abstract);navigation(abstract);robot foundation model(abstract);分类 cs.RO
机构
*
Department of Computer Science, University College London(计算机科学系,伦敦大学学院)
;
Department of Mechanical Engineering, University College London(机械工程系,伦敦大学学院)
Q-VGM: Q-Value-Gradient Matching for Offline-to-Online Reinforcement Learning of Flow-Matching VLA
Q-VGM: 基于Q引导的值梯度匹配的流匹配VLA策略
Ziqian Wang, Yitian Liu, Xingjian Mao, Minqian Wang, Yao Mu
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
University of Michigan, Ann Arbor(密歇根大学安娜堡分校)
;
University of Electronic Science and Technology of China(电子科技大学)
RARM: Confidence-Gated Progress Reward Modeling for RL in Manipulation
RARM:基于置信度门控的进展奖励建模用于操作中的强化学习
Pengzhi Yang, Xinyu Wang, Pengyu Jing, Kehan Wen, Yiduo Qu, Zhenhao Huang, Minghao Fu, Xin Liu, Yaheng Shen, Fan Shi
机构
*
NUS Human-Centered Robotic Lab(新加坡国立大学人机共融机器人实验室)
;
University of Cambridge(剑桥大学)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
WorldToken: Time-First Sequence Modeling for Robotic Imitation Learning
WorldToken:面向机器人模仿学习的时间优先序列建模
Chunkai Yang, Andong Yang, Chao Gao
机构
*
School of Remote Sensing and Information Engineering, Wuhan University(武汉大学遥感信息工程学院)
;
Tsinghua University(清华大学)
;
Institute for AI Industry Research, Tsinghua University(清华大学人工智能产业研究院)