← Search

Zhiyong Cheng

4 accepted papers

2026

Exploring Adaptive Masked Reconstruction for Self-Supervised Skeleton-Based Action Recognition

CVPR 2026

Recently, masked skeleton reconstruction models have emerged as strong action representation learners, driving significant progress in self-supervised skeleton-based action recognition. However, existing state-of-the-art methods must predict an exceedingly large number of spatiotemporal patches, sig

Cited by 0SourcecodeScholar
2025

Towards Efficient General Feature Prediction in Masked Skeleton Modeling

ICCV 2025poster

Recent advances in the masked autoencoder (MAE) paradigm have significantly propelled self-supervised skeleton-based action recognition. However, most existing approaches limit reconstruction targets to raw joint coordinates or their simple variants, resulting in computational redundancy and limited…

Cited by 0SourcePDFScholar
2024

Causality-Inspired Invariant Representation Learning for Text-Based Person Retrieval

AAAI 2024technical

Text-based Person Retrieval (TPR) aims to retrieve relevant images of specific pedestrians based on the given textual query. The mainstream approaches primarily leverage pretrained deep neural networks to learn the mapping of visual and textual modalities into a common latent space for cross-modalit…

Cited by 17SourcePDFScholar
2021

AdaVQA: Overcoming Language Priors with Adapted Margin Cosine Loss

IJCAI 2021poster

A number of studies point out that current Visual Question Answering (VQA) models are severely affected by the language prior problem, which refers to blindly making predictions based on the language shortcut. Some efforts have been devoted to overcoming this issue with delicate models. However, the…