EgoInstruct: An Egocentric Video Dataset of Face-to-face Instructional Interactions with Multi-modal LLM Benchmarking
机构 * The University of Tokyo(东京大学)
Comments Accepted to the I-HFM Workshop at ICCV 2025
期刊&会议
International Conference on Computer Vision · 会议 · Computer Vision
机构 * The University of Tokyo(东京大学)
Comments Accepted to the I-HFM Workshop at ICCV 2025
机构 * Dept. of Computer Science(计算机科学系) ; Gauhati University(果阿大学)
Comments 7 pages, 3 figures and 1 table. 2024 IEEE International Conference on Computer Vision and Machine Intelligence (CVMI). IEEE, 2024
机构 * Tongji University(同济大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; Shanghai Jiao Tong University(上海交通大学) ; State Key Laboratory of Autonomous Intelligent Unmanned Systems(自主智能无人系统国家重点实验室)
Comments Accepted by ICCV 2025
机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; Cornell University(康奈尔大学) ; University of Michigan(密歇根大学) ; Columbia University(哥伦比亚大学)
Comments Accepted by 2025 6th International Conference on Computer Vision, Image and Deep Learning
Journal ref Proceedings of the 2025 6th International Conference on Computer Vision, Image and Deep Learning (CVIDL), 2025, pp. 693-696
机构 * ETH Zürich - DALAB(苏黎世联邦理工学院- DALAB) ; Technical University of Munich(慕尼黑技术大学) ; Google Switzerland(瑞士谷歌)
Comments Accepted to ICCV'25. Project page: https://uip2p.github.io/