2023
PSVT: End-to-End Multi-Person 3D Pose and Shape Estimation With Progressive Video Transformers
CVPR 2023poster
Existing methods of multi-person video 3D human Pose and Shape Estimation (PSE) typically adopt a two-stage strategy, which first detects human instances in each frame and then performs single-person PSE with temporal model. However, the global spatio-temporal context among spatial instances can not…