GeoWorld-VLM: Geometry from World Models for Vision-Language Models
GeoWorld-VLM:从世界模型中获取几何结构用于视觉-语言模型
Renjie Gu, Kaichen Zhou, Yan Luo, Mengyu Wang
机构
*
Harvard AI and Robotics Lab(哈佛人工智能与机器人实验室)
;
Kempner Institute for the Study of Natural and Artificial Intelligence(凯普纳自然与人工智能研究 institute)
;
Harvard University(哈佛大学)
Earth-OneVision: Extending Remote Sensing Multimodal Large Language Models to More Sensor Modalities and Tasks
Earth-OneVision:将遥感多模态大语言模型扩展到更多传感器模态和任务
Miaoxin Cai, Guanqun Wang, Wei Zhang, Guangyao Zhou, Yin Zhuang, Tong Zhang, Hao Wang, He Chen, Jun Li
机构
*
National Key Laboratory of Science and Technology on Space-Born Intelligent Information Processing (SBIIP), Beijing Institute of Technology(北京理工大学空间智能信息处理国家重点实验室)
;
Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院空天信息创新研究院)
;
Key Laboratory of Technology in Geo-Spatial Information Processing and Application System, Chinese Academy of Sciences(中国科学院地理空间信息处理与应用系统技术重点实验室)
;
Advanced Research Institute of Multidisciplinary Sciences, Beijing Institute of Technology(北京理工大学前沿交叉科学研究院)
;
School of Mechatronical Engineering, Beijing Institute of Technology(北京理工大学机电学院)
;
School of Earth and Space Sciences, Peking University(北京大学地球与空间科学学院)
;
School of Electronics, Peking University(北京大学电子学院)
;
School of Computer Science and Hubei Key Laboratory of Intelligent Geo-Information Processing(华中科技大学计算机科学与技术学院&湖北省智能地理信息处理重点实验室)
机构
*
School of Artificial Intelligence, Beijing Normal University, Beijing, China(北京师范大学人工智能学院)
;
Engineering Research Center of Intelligent Technology(智能技术与教育应用工程研究中心)
;
Beijing Key Laboratory of Artificial Intelligence for Education, Beijing, China(北京人工智能教育重点实验室)
;
Baidu, Beijing, China(百度)
GVC-Seg: Training-Free 3D Instance Segmentation via Geometric Visual Correspondence
GVC-Seg: 基于几何视觉对应的免训练3D实例分割
Liang Xu, Fangjing Wang, Jinyu Yang, Feng Zheng
机构
*
Victoria University of Wellington(惠灵顿维多利亚大学)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Southern University of Science and Technology(南方科技大学)
SAGE: Shape-Adapting Gated Experts for Adaptive Histopathology Image Segmentation
SAGE:适应性组织病理图像分割的形状自适应门控专家
Gia Huy Thai, Hoang-Nguyen Vu, Anh-Minh Phan, Quang-Thinh Ly, Thi-Ngoc-Truc Nguyen, Nhat Ho
机构
*
University of Science, VNU-HCM(越南国家大学科学学院)
;
Trivita AI
;
University of Technology, VNU-HCM(越南国家大学技术学院)
;
Michigan State University, USA(美国密歇根州立大学)
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography
RadAgent:一种用于胸部CT逐步解读的工具型AI智能体
Mélanie Roschewitz, Kenneth Styppa, Yitian Tao, Jiwoong Sohn, Jean-Benoit Delbrouck, Benjamin Gundersen, Nicolas Deperrois, Christian Bluethgen, Julia E. Vogt, Bjoern Menze, Farhad Nooralahzadeh, Michael Krauthammer, Michael Moor
机构
*
Department of Biosystems Science and Engineering, ETH Zurich(生物系统科学与工程系,苏黎世联邦理工学院)
;
ETH AI Center, Zurich(ETH人工智能中心,苏黎世)
;
Department of Computer Science, ETH Zurich(计算机科学系,苏黎世联邦理工学院)
;
Faculty of Computer Science and Mathematics, Heidelberg University(计算机科学与数学学院,海德堡大学)
;
Stanford Center for Artificial Intelligence in Medicine and Imaging, Stanford University(斯坦福大学人工智能在医学和影像中的中心)
;
Department of Radiology, Stanford University(放射科,斯坦福大学)
;
Department of Quantitative Biomedicine, University of Zurich(定量生物医学系,苏黎世大学)
;
Institute of Computer Science, Zurich University of Applied Sciences(应用科学大学计算机科学研究所)
A Deep Learning Model of Mental Rotation Informed by Interactive VR Experiments
基于交互式VR实验的心理旋转深度学习模型
Raymond Khazoum, Daniela Fernandes, Aleksandr Krylov, Qin Li, Stephane Deny
机构
*
Department of Computer Science, Aalto University, Espoo, Finland(奥卢大学计算机科学系,芬兰埃斯波)
;
Department of Neuroscience and Biomedical Engineering, Aalto University, Espoo, Finland(奥卢大学神经科学与生物医学工程系,芬兰埃斯波)