arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-02-03 至 2026-02-03 共收录 38 信号源:cs.CV, cs.GR, cs.RO

1. 点云 12 篇

2602.00110 2026-02-03 cs.CV cs.LG 57%

Observing Health Outcomes Using Remote Sensing Imagery and Geo-Context Guided Visual Transformer

利用遥感图像和地理上下文引导的视觉变换器观察健康结果

Yu Li, Guilherme N. DeSouza, Praveen Rao, Chi-Ren Shyu

专题命中 点云 :spatial understanding(abstract);分类 cs.CV

AI总结 本文提出一种结合地理空间信息的视觉变换器模型,通过引导注意力机制提升遥感图像处理效果,有效预测疾病流行率。

Comments Submitted to IEEE Transactions on Geoscience and Remote Sensing

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00973 2026-02-03 cs.HC 50%

Exploration of Radar-based Obstacle Visualizations to Support Safety and Presence in Camera-Free Outdoor VR

雷达基于的障碍可视化探索以支持无摄像头户外VR中的安全与存在

Avinash Ajit Nargund, Andrew L. Huard, Tobias Höllerer, Misha Sra

专题命中 点云 :point cloud(abstract)

AI总结 本研究探索了雷达基于的障碍可视化技术,通过三种不同方法比较,发现其在户外VR中平衡沉浸、安全与隐私方面存在不同效果和局限。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22432 2026-02-03 cs.LG cs.CG 50%

The Flood Complex: Large-Scale Persistent Homology on Millions of Points

洪水复杂性:在数百万点上的大规模持久同调

Florian Graf, Paolo Pellizzoni, Martin Uray, Stefan Huber, Roland Kwitt

专题命中 点云 :point cloud(abstract)

AI总结 本文提出洪水复形,用于高效计算大规模点云的持久同调,适用于下游机器学习任务,尤其在处理复杂几何或拓扑结构时表现更优。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 空间理解 1 篇

2602.00551 2026-02-03 cs.RO cs.CV 62%

APEX: A Decoupled Memory-based Explorer for Asynchronous Aerial Object Goal Navigation

APEX:一种解耦的记忆型探索者用于异步空中目标导航

Daoxuan Zhang, Ping Chen, Xiaobo Xia, Xiu Su, Ruichen Zhen, Jianqiang Xiao, Shuo Yang

机构 * Harbin Institute of Technology(哈尔滨工业大学) National University of Singapore(新加坡国立大学) Central South University(中南大学) Meituan Academy of Robotics(美团机器人研究院)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.CV、cs.RO

AI总结 APEX通过分层异步架构实现高效空中目标导航,提升探索效率与目标识别精度。

Comments 15 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. SLAM与定位 3 篇

2602.02430 2026-02-03 cs.RO 57%

3D Foundation Model-Based Loop Closing for Decentralized Collaborative SLAM

基于3D基础模型的分布式协作SLAM中的回环闭合

Pierre-Yves Lajoie, Benjamin Ramtoula, Daniele De Martini, Giovanni Beltrame

机构 * Department of Computer and Software Engineering, Polytechnique Montréal(计算机与软件工程系,蒙特利尔理工学院) Oxford Robotics Institute, Dept. of Engineering Science, University of Oxford(牛津大学机器人研究所,工程科学系)

专题命中 SLAM与定位 :3D reconstruction(abstract);分类 cs.RO

AI总结 本文提出基于3D基础模型的分布式协作SLAM回环闭合方法,通过整合基础模型提升多机器人建图的鲁棒性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09509 2026-02-03 cs.RO 57%

SMapper: A Multi-Modal Data Acquisition Platform for SLAM Benchmarking

SMapper: 一种用于SLAM基准测试的多模态数据采集平台

Pedro Miguel Bastos Soares, Ali Tourani, Miguel Fernandez-Cortizas, Asier Bikandi-Noya, Holger Voos, Jose Luis Sanchez-Lopez

机构 * Automation and Robotics Research Group (ARG), Interdisciplinary Centre for Security, Reliability, and Trust (SnT), University of Luxembourg(卢森堡大学自动化与机器人研究组(ARG)、安全、可靠与信任跨学科研究中心(SnT))

专题命中 SLAM与定位 :3D reconstruction(abstract);分类 cs.RO

AI总结 SMapper是一种用于SLAM基准测试的多模态数据采集平台,通过开放硬件设计和可重复的数据收集,提供高精度的SLAM数据集和基准测试结果,推动SLAM算法的发展和评估。

Comments 13 pages, 5 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01095 2026-02-03 cs.CV 57%

PandaPose: 3D Human Pose Lifting from a Single Image via Propagating 2D Pose Prior to 3D Anchor Space

PandaPose: 通过将2D姿态先验传播到3D锚空间实现单张图像的3D人体姿态提升

Jinghong Zheng, Changlong Jiang, Yang Xiao, Jiaqi Li, Haohong Kuang, Hang Xu, Ran Wang, Zhiguo Cao, Min Du, Joey Tianyi Zhou

机构 * School of Journalism and Information Communication, Huazhong University of Science and Technology(华中科技大学新闻与信息传播学院) ByteDance Inc.(字节跳动公司) Institute of High Performance Computing, Agency for Science, Technology and Research, Singapore(科技研究局高性能计算研究所)

专题命中 SLAM与定位 :3D vision(abstract);分类 cs.CV

AI总结 PandaPose通过将2D姿态先验传播到3D锚空间,提升单张图像的3D人体姿态估计精度,有效缓解自遮挡问题并减少误差。

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 其他3D视觉 1 篇

2602.01200 2026-02-03 cs.CV 57%

Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis

Med3D-R1: 促进3D医学视觉-语言模型的临床推理

Haoran Lai, Zihang Jiang, Kun Zhang, Qingsong Yao, Rongsheng Wang, Zhiyang He, Xiaodong Tao, Wei Wei, Shaohua Kevin Zhou

机构 * School of Biomedical Engineering, Division of Life Sciences and Medicine, University of Science and Technology of China(生物医学工程学院,生命科学与医学系,中国科学技术大学) Suzhou Institute for Advanced Research, University of Science and Technology of China(先进研究所,中国科学技术大学) Stanford University(斯坦福大学) Medical Business Department, iFlytek Co.Ltd(iFlytek公司医学业务部) The First Affiliated Hospital of USTC, Division of Life Sciences and Medicine University of Science and Technology of China(中国科学技术大学第一附属医院,生命科学与医学系)

专题命中 其他3D视觉 :3D vision(abstract);分类 cs.CV

AI总结 Med3D-R1通过两阶段训练框架提升3D医学视觉-语言模型的临床推理能力,实现更准确的异常诊断。

详情

展开后加载摘要…

URL PDF HTML 收藏