arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

2026-08-27 至 2026-08-27 共收录 6
2608.23720 2026-08-27 cs.CV 版本更新

Platonic Representation Hypothesis on World Models

世界模型的柏拉图式表征假说

Wenhow Li, Chengwei MA, Hui Xiong, Ying-Cong Chen, Lei Zhang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 该研究针对世界模型的表征性质,提出预测一致性假设,通过DINO-WM实验发现性能良好的世界模型会形成几何相似的内部结构,且模型特征可跨模型映射,证实预测一致性能促进共享潜在结构的形成。

Comments 18 pages, 10 figures, 2 tables. Wenhow Li and Chengwei MA contributed equally. Project page: this https URL (https://sellerbubble.github.io/platonic-representation-hypothesis-on-world-models/). Corrected author metadata formatting and updated the project-page link presentation in the abstract; manuscript content and results unchanged

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.22991 2026-08-27 cs.CV 版本更新

Training-Free Interaction-Aligned Visual Token Pruning for Efficient Embodied Manipulation

VLA-IAP: 通过交互对齐实现无训练视觉令牌剪枝用于视觉-语言-动作模型

Jintao Cheng, Weibin Li, Haozhe Wang, Gang Wang, Yipu Zhang, Xiaoyu Tang, Jin Wu, Xieyuanli Chen, Yunhui Liu, Wei Zhang

机构 * Hong Kong University of Science and Technology(香港科学与技术大学) South China Normal University(华南师范大学) The Chinese University of Hong Kong(香港中文大学) University of Science and Technology Beijing(北京科技大学) National University of Defense Technology(国防科技大学)

AI总结 本文提出VLA-IAP方法,通过引入几何先验机制和动态调度策略,实现无训练的视觉令牌剪枝,提升VLA模型在资源受限平台上的推理效率和稳定性,实验显示在LIBERO基准上成功率达97.8%且速度提升1.25倍。

Comments 27 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00686 2026-08-27 cs.RO 版本更新

Learning to Accelerate Vision-Language-Action Models through Adaptive Visual Token Caching

通过自适应视觉令牌缓存学习加速视觉-语言-动作模型

Yujie Wei, Jiahan Fan, Jiyu Guo, Ruichen Zhen, Rui Shao, Xiu Su, Zeke Xie, Hongxun Yao, Shuo Yang

机构 * Harbin Institute of Technology(哈尔滨工业大学) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Meituan Academy of Robotics Shenzhen, Meituan(美团机器人深圳研究院) Central South University(中南大学) HKUST(GZ)(香港科技大学(广州))

AI总结 本文提出通过自适应视觉令牌缓存学习加速VLA模型,提升推理效率并提高任务成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06472 2026-08-27 cs.LG 版本更新

StablePDENet: Enhancing Neural Operator Stability through Physics-Informed Residual-Sensitivity Regularization

StablePDENet: 提高求解微分方程的算子学习稳定性

Chutian Huang, Chang Ma, Kaibo Wang, Yang Xiang

机构 * Department of Mathematics, The Hong Kong University of Science and Technology(香港科技大学数学系) Algorithms of Machine Learning and Autonomous Driving Research Lab(机器学习与自动驾驶算法研究实验室) HKUST Shenzhen-Hong Kong Collaborative Innovation Research Institute(香港科技大学深圳-香港协同创新研究院)

AI总结 StablePDENet通过对抗训练提升神经算子学习的稳定性,确保在正常和对抗条件下均能保持高精度求解微分方程。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03542 2026-08-27 cs.CL cs.AI 版本更新

Layer-Order Inversion: Rethinking Latent Multi-Hop Reasoning in Large Language Models

层序倒置:重新思考大语言模型中的潜在多跳推理

Xukai Liu, Ye Liu, Jipeng Zhang, Yanghai Zhang, Kai Zhang, Qi Liu

机构 * State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室) University of Science and Technology of China(中国科学技术大学) The Hong Kong University of Science and Technology(香港科技大学)

AI总结 本文提出"层序倒置"现象,通过概率性回忆与提取框架解释大语言模型中多跳推理的机制,揭示了层序倒置与总跳步数的关系,并提供了多跳失败的诊断方法。

Comments 16 pages, 18 figures, EMNLP 2026 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24209 2026-08-27 cs.CV 版本更新

Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse-view Videos

Forge4D:基于未校准稀疏视图视频的前馈式4D人体重建与插值

Yingdong Hu, Yisheng He, Jinnan Chen, Weihao Yuan, Kejie Qiu, Zehong Lin, Siyu Zhu, Zilong Dong, Steven Hoi, Jun Zhang

机构 * HKUST(香港科技大学) Tongyi Lab, Alibaba Group(阿里云实验室) NUS(新加坡国立大学) FDU(福建大学)

AI总结 提出Forge4D模型,将4D人体重建与插值简化为流式3D高斯重建和稠密运动预测任务,结合自监督损失实现高效重建与插值,在多数据集上验证有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏