2026
SUGAR: Learning Skeleton Representation with Visual-Motion Knowledge for Action Recognition
AAAI 2026technical
Large Language Models (LLMs) hold rich implicit knowledge and powerful transferability. In this paper, we explore the combination of LLMs with the human skeleton to perform action classification and description. However, when treating LLM as a recognizer, two questions arise: 1) How can LLMs underst