← Search

Jing-Yao Zhang

1 accepted papers

2026

SAM2Text: Towards Prompt-Free and Multi-Resolution Video Scene Text Segmentation

CVPR 2026

We introduce a novel method for video Scene Text Segmentation (STS), a task critical for understanding dynamic visual content. Despite the success of foundation models like Segment Anything Model 2 (SAM2) in generic segmentation, their application to video STS is hindered by the reliance on external

Cited by 0SourcecodeScholar