← Search

Xiaoge Song

3 accepted papers

2025

Conditional Convolutions for End-to-End Single-Stage Video Text Detection

ICASSP 2025accepted

We propose a simple yet effective single-stage video text detection framework, termed CVTD (Conditional convolutions for Video Text Detection), which, to the best of our knowledge, is the first end-to-end single-stage video text detection framework.Most existing video text detection methods adopt te…

Cited by 0SourceScholar
2020

PolarMask: Single Shot Instance Segmentation With Polar Representation

CVPR 2020oral

In this paper, we introduce an anchor-box free and single shot instance segmentation method, which is conceptually simple, fully convolutional and can be used by easily embedding it into most off-the-shelf detection methods. Our method, termed PolarMask, formulates the instance segmentation problem…

Cited by 764PDFcodeScholar
2019

Efficient and Accurate Arbitrary-Shaped Text Detection With Pixel Aggregation Network

ICCV 2019poster

Scene text detection, an important step of scene text reading systems, has witnessed rapid development with convolutional neural networks. Nonetheless, two main challenges still exist and hamper its deployment to real-world applications. The first problem is the trade-off between speed and accuracy.…

Cited by 666PDFcodeScholar