← Search

Sanyi Zhang

5 accepted papers

2026

MAG: Multi-Modal Aligned Autoregressive Co-Speech Gesture Generation without Vector Quantization

ICASSP 2026oral

This work focuses on full-body co-speech gesture generation. Existing methods typically employ an autoregressive model accompanied by vector-quantized tokens for gesture generation, which results in information loss and compromises the realism of the generated gestures. To address this, inspired by…

Cited by 0SourcePDFScholar
2022

PMP-NET: Rethinking Visual Context for Scene Graph Generation

ICASSP 2022accepted

Scene graph generation aims to describe the contents in scenes by identifying the objects and their relationships. In previous works, visual context is widely utilized in message passing networks to generate the representations for classification. However, the noisy estimation of visual context limi…

Cited by 0SourceScholar
2021

Progressive Contour Regression for Arbitrary-Shape Scene Text Detection

CVPR 2021poster

State-of-the-art scene text detection methods usually model the text instance with local pixels or components from the bottom-up perspective and, therefore, are sensitive to noises and dependent on the complicated heuristic post-processing especially for arbitrary-shape texts. To relieve these two i…

Cited by 141PDFcodeScholar