ShapeR: Robust Conditional 3D Shape Generation from Casual Captures
ShapeR: 从随意捕获序列中生成鲁棒的条件3D形状
Yawar Siddiqui, Duncan Frost, Samir Aroudj, Armen Avetisyan, Henry Howard-Jenkins, Daniel DeTone, Pierre Moulon, Qirui Wu, Zhengqin Li, Julian Straub, Richard Newcombe, Jakob Engel
机构
*
Meta Reality Labs Research(Meta现实实验室)
;
Simon Fraser University(西蒙弗雷泽大学)
Can Vision-Language Models Understand Construction Workers? An Exploratory Study
视觉-语言模型能理解建筑工人吗?一项探索性研究
Hieu Bui, Nathaniel E. Chodosh, Arash Tavakoli
机构
*
Department of Electrical and Computer Engineering(电气与计算机工程系)
;
Villanova University(维拉诺瓦大学)
;
Department of Computing Sciences(计算科学系)
;
Department of Civil and Environmental Engineering(土木与环境工程系)
Attention Debiasing for Token Pruning in Vision Language Models
视觉语言模型中用于标记剪枝的注意力去偏
Kai Zhao, Wubang Yuan, Yuchen Lin, Liting Ruan, Xiaofeng Lu, Deng-Ping Fan, Ming-Ming Cheng, Dan Zeng
机构
*
School of Communication and Information Engineering, Shanghai University(通信与信息工程学院,上海大学)
;
School of Computer Science and Engineering, Shanghai University(计算机科学与工程学院,上海大学)
;
School of Computer Science and Engineering, Nankai University(计算机科学与工程学院,南开大学)
专题命中
VLM训练与架构
:vision language model(title);vision-language model(abstract);VLM(abstract);分类 cs.CV
CommentsThis is an earlier version of the work released in May 2025. The version accepted at CHI 2026 is available as a separate preprint at arXiv:2511.04366
机构
*
Tsinghua University(清华大学)
;
Institute for Interdisciplinary Information Sciences(交叉信息研究院)
;
Shanghai Qi Zhi Institute(上海启智研究院)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Huazhong University of Science and Technology(华中科技大学)
机构
*
School of Integrated Circuits, Zhejiang University(浙江大学集成电路学院)
;
Shanghai Innovation Institute(上海创新研究院)
;
School of Electrical and Electronic Engineering (EEE), Nanyang Technological University(南洋理工大学电子与电气工程学院)
;
School of Data Science, Fudan University(复旦大学数据科学学院)