Fine-grained Spatiotemporal Grounding on Egocentric Videos
细粒度眼动视频时空定位
机构 * The Chinese University of Hong Kong(香港中文大学)
专题命中 视频数据与评测 :video understanding(abstract);分类 cs.CV
AI总结 本文提出EgoMask基准和EgoMask-Train数据集,针对眼动视频细粒度时空定位的挑战,通过微调提升模型性能。
Comments Accepted by ICCV 2025