← Search

Tongkai Shi

3 accepted papers

2025

GReg: Geometry-Aware Region Refinement for Sign Language Video Generation

ICCV 2025poster

Sign Language Video Generation (SLVG) aims to transform sign language sequences into natural and fluent sign language videos. Existing SLVG methods lack geometric modeling of human anatomical structures, leading to anatomically implausible and temporally inconsistent generation. To address these cha…

Cited by 0SourcePDFScholar
2024

Deep Correlated Prompting for Visual Recognition with Missing Modalities

NeurIPS 2024poster

Large-scale multimodal models have shown excellent performance over a series of tasks powered by the large corpus of paired multimodal training data. Generally, they are always assumed to receive modality-complete inputs. However, this simple assumption may not always hold in the real world due to p…

2024

Pose Guided Fine-Grained Sign Language Video Generation

ECCV 2024poster

"Sign language videos are an important medium for spreading and learning sign language. However, most existing human image synthesis methods produce sign language images with details that are distorted, blurred, or structurally incorrect. They also produce sign language video frames with poor tempor…

Cited by 1SourcePDFScholar