2025
Conditional Convolutions for End-to-End Single-Stage Video Text Detection
ICASSP 2025accepted
We propose a simple yet effective single-stage video text detection framework, termed CVTD (Conditional convolutions for Video Text Detection), which, to the best of our knowledge, is the first end-to-end single-stage video text detection framework.Most existing video text detection methods adopt te…