2026
SAM2Text: Towards Prompt-Free and Multi-Resolution Video Scene Text Segmentation
CVPR 2026
We introduce a novel method for video Scene Text Segmentation (STS), a task critical for understanding dynamic visual content. Despite the success of foundation models like Segment Anything Model 2 (SAM2) in generic segmentation, their application to video STS is hindered by the reliance on external