← Search

Shi-Xue Zhang

4 accepted papers

2026

VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation

AAAI 2026technical

Video captions play a crucial role in text-to-video generation tasks, as their quality directly influences the semantic coherence and visual fidelity of the generated videos. Although large vision-language models (VLMs) have demonstrated significant potential in caption generation, existing benchmar

Cited by 0SourcePDFScholar
2023

Learning Correction Filter via Degradation-Adaptive Regression for Blind Single Image Super-Resolution

ICCV 2023poster

Although existing image deep learning super-resolution (SR) methods achieve promising performance on benchmark datasets, they still suffer from severe performance drops when the degradation of the low-resolution (LR) input is not covered in training. To address the problem, we propose an innovative…

Cited by 32PDFcodeScholar
2021

Adaptive Boundary Proposal Network for Arbitrary Shape Text Detection

ICCV 2021poster

Arbitrary shape text detection is a challenging task due to the high complexity and variety of scene texts. In this work, we propose a novel adaptive boundary proposal network for arbitrary shape text detection, which can learn to directly produce accurate boundary for arbitrary shape text without a…

Cited by 121PDFcodeScholar
2020

Deep Relational Reasoning Graph Network for Arbitrary Shape Text Detection

CVPR 2020oral

Arbitrary shape text detection is a challenging task due to the high variety and complexity of scenes texts. In this paper, we propose a novel unified relational reasoning graph network for arbitrary shape text detection. In our method, an innovative local graph bridges a text proposal model via Con…

Cited by 281PDFcodeScholar