2022
Disentangled Action Recognition with Knowledge Bases
NAACL 2022long
Action in video usually involves the interaction of human with objects. Action labels are typically composed of various combinations of verbs and nouns, but we may not have training data for all possible combinations. In this paper, we aim to improve the generalization ability of the compositional a…