RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
机构 * Institute for Anthropomatics and Robotics, Karlsruhe Institute of Technology(人机化研究所,卡尔斯鲁厄技术大学) ; RISE Research Institutes of Sweden(瑞典RISE研究机构) ; KTH Royal Institute of Technology(皇家理工学院) ; School of Artificial Intelligence and Robotics(人工智能与机器人学院) ; National Engineering Research Center of Robot Visual Perception and Control Technology(机器人视觉感知与控制技术国家工程研究中心) ; Chinese University of Hong Kong(香港中文大学) ; Shanghai AI Lab(上海人工智能实验室)
Comments Extended version of ECCV 2024 paper arXiv:2407.01872. The dataset and code are released at https://github.com/KPeng9510/refAVA2