ECCV 2022poster29 citations

Learning Spatial-Preserved Skeleton Representations for Few-Shot Action Recognition

Ning Ma, Hongyi Zhang, Xuhui Li, Sheng Zhou, Zhen Zhang, Jun Wen, Haifeng Li, Jingjun Gu

Abstract

"Few-shot action recognition aims to recognize few-labeled novel action classes and attracts growing attentions due to practical significance. Human skeletons provide explainable and data-efficient representation for this problem by explicitly modeling spatial-temporal relations among skeleton joints. However, existing skeleton-based spatial-temporal models tend to deteriorate the positional distinguishability of joints, which leads to fuzzy spatial matching and poor explainability. To address these issues, we propose a novel spatial matching strategy consisting of spatial disentanglement and spatial activation. The motivation behind spatial disentanglement is that we find more spatial information for leaf nodes (e.g., the “hand” joint ) is beneficial to increase representation diversity for skeleton matching. To achieve spatial disentanglement, we encourage the skeletons to be represented in a full rank space with rank maximization constraint. Finally, an attention based spatial activation mechanism is introduced to incorporate the disentanglement, by adaptively adjusting the disentangled joints according to matching pairs. Extensive experiments on three skeleton benchmarks demonstrate that the proposed spatial matching strategy can be effectively inserted into existing temporal alignment frameworks, achieving considerable performance improvements as well as inherent explainability."

BibTeX
@inproceedings{eccv2022_learningspatialp,
  title = {Learning Spatial-Preserved Skeleton Representations for Few-Shot Action Recognition},
  author = {Ning Ma and Hongyi Zhang and Xuhui Li and Sheng Zhou and Zhen Zhang and Jun Wen and Haifeng Li and Jingjun Gu and Jiajun Bu},
  booktitle = {ECCV 2022},
  year = {2022}
}
Learning Spatial-Preserved Skeleton Representations for Few-Shot Action Recognition · ECCV 2022