Towards Open Environments and Instructions: General Vision-Language Navigation via Fast-Slow Interactive Reasoning
迈向开放环境和指令:通过快速-缓慢交互推理实现通用视觉-语言导航
Yang Li, Aming Wu, Zihao Zhang, Yahong Han
机构
*
School of Artificial Intelligence, College of Intelligence and Computing, Tianjin University, China(人工智能学院,智能与计算学院,天津大学,中国)
;
School of Computer Science and Information Engineering, Hefei University of Technology, China(计算机科学与信息工程学院,合肥工业大学,中国)
Bodhi VLM: Privacy-Alignment Modeling for Hierarchical Visual Representations in Vision Backbones and VLM Encoders via Bottom-Up and Top-Down Feature Search
机构
*
Auckland University of Technology(奥克兰技术大学)
;
Resideo Technologies, Inc.(Resideo技术公司)
;
Guilin University of Electronic Technology(桂林电子科技大学)
;
Universidad de Chile(智利大学)
ExGS: Extreme 3D Gaussian Compression with Diffusion Priors
ExGS:基于扩散先验的极端3D高斯压缩
Jiaqi Chen, Xinhao Ji, Yuanyuan Gao, Hao Li, Yuning Gong, Yifei Liu, Dan Xu, Zhihang Zhong, Dingwen Zhang, Xiao Sun
机构
*
Northwestern Polytechnical University(北western工业大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
Hong Kong University of Science and Technology(香港科技大学)
Analytical Expression for Spherically Symmetric Photoacoustic Sources: A Unified General Solution (Theoretical Analysis and Derivation)
球对称光声源的解析表达式:统一的通用解(理论分析与推导)
Shuang Li, Yibing Wang, Yu Zhang, Changhui Li
机构
*
Department of Biomedical Engineering, College of Future Technology, Peking University, Beijing, China(北京大学未来技术学院生物医学工程系)
;
National Biomedical Imaging Center, Peking University, Beijing, China(北京大学国家生物医学成像中心)