← Search

Zhihang Yu

1 accepted papers

2022

Towards Video Text Visual Question Answering: Benchmark and Baseline

NeurIPS 2022accept

There are already some text-based visual question answering (TextVQA) benchmarks for developing machine's ability to answer questions based on texts in images in recent years. However, models developed on these benchmarks cannot work effectively in many real-life scenarios (e.g. traffic monitoring,…