Efficient Posenet with Coarse to Fine Transformer
In recent years, Transformers have been widely applied in human pose estimation by converting image features into token forms as inputs. However, redundant information in images burdens the network and can even negatively impact training as noise. Thus, we propose a coarse-to-fine Transformer called…