VizAnchor: Decoding Manipulation Intent from Tampering Visualizations via Dual-Anchor Reasoning
VizAnchor:通过双锚推理从篡改可视化中解码操纵意图
Xiaotian Zhang, Huayuan Ye, Haiyang Zhang, Chenhui Li, Changbo Wang, Sicheng Song
机构
*
School of Data Science and Engineering, East China Normal University(华东师范大学数据科学与工程学院)
;
Division of Emerging Interdisciplinary Areas, The Hong Kong University of Science and Technology(香港科技大学新兴跨学科领域分部)
;
School of Computer Science and Technology, East China Normal University(华东师范大学计算机科学与技术学院)
机构
*
Tuojing Intelligence(拓境智能)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Tsinghua University(清华大学)
;
Simple AI(简智人工智能)
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
;
Shanghai Jiao Tong University(上海交通大学)
;
Zhejiang University(浙江大学)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
The University of Hong Kong(香港大学)
机构
*
Tsinghua University(清华大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
Tsingscribe Medical Ltd.(清scribe医疗有限公司)
;
Peking University School and Hospital of Stomatology(北京大学口腔医学院)
;
Beijing National Research Center for Information Science and Technology(北京信息科学与技术国家研究中心)
机构
*
Peking University(北京大学)
;
Beijing Innovation Center of Humanoid Robotics(北京人形机器人创新中心)
;
New York University(纽约大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Nanyang Technological University(南洋理工大学)
;
The Chinese University of Hong Kong(香港中文大学)
CommentsWe are currently building Gaming World Model and data engine that transfers game dynamics to robotics. Feel free to contact Dongping Chen (dongpingchen0612@gmail.com) if you are interested in research collaboration or financial support
Platonic Representation Hypothesis on World Models
世界模型的柏拉图式表征假说
Wenhow Li (1), Chengwei MA (1), Hui Xiong (1), Ying-Cong Chen (1), Lei Zhang (1) ((1) The Hong Kong University of Science and Technology (Guangzhou), Guangzhou, China)
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
Learning to Act While Waiting: RL Finetuning of Generalist Robot Policies Under Inference Latency
等待时学习行动:考虑推理延迟的通用机器人策略的强化学习微调
Brian Zhu (1), Momen Khalil (1), E Harrison (2), Emanuele Poggi (1), Philipp Schmitt (1), Bernd Kast (1), Philine Meister (1), Pranav Atreya (2), Qiyang Li (2), Finn Ferchau (1), Cesar Colmenero (1), Yash Shahapurkar (1), Gokul Narayanan (1), Melih Erdogan (1), Kai Wurm (1), Georg von Wichert (1), Oier Mees (3 and 4 and 2), Eugen Solowjow (1), Andrew Wagenmaker (2), Sergey Levine (2) ((1) Siemens (2) UC Berkeley (3) Microsoft (4) ETH Zurich)
ConsensusTAS: Self-Supervised Temporal Action Segmentation for Long-Horizon Construction Videos
ConsensusTAS:面向长时程施工视频的自监督时序动作分割
Xiaoshan Zhou, Yafei Sun
机构
*
School of Project Management, University of Sydney(悉尼大学项目管理学院)
;
School of Civil and Environmental Engineering, University of New South Wales(新南威尔士大学土木与环境工程学院)