← Search

Jinwen Zhong

5 accepted papers

2026

False Positives Matter: Multidimensional Localization Evaluation and Training-Free Explainable Adversarial Patch Defense

AAAI 2026technical

Adversarial patch attacks pose a significant threat to visual systems. While current patch purification-based defense methods enhance core metrics of visual perception models, they overlook the critical issue of false positive patches, severely compromising image usability. This paper reveals the in

Cited by 0SourcePDFScholar
2025

Arbitrary Reading Order Scene Text Spotter with Local Semantics Guidance

AAAI 2025technical

Scene text spotting has attracted the enthusiasm of relative researchers in recent years. Most existing scene text spotters follow the detection-then-recognition paradigm, where the vanilla detection module hardly determines the reading order and leads to failure recognition. After rethinking the au…

Cited by 2SourcePDFScholar
2025

LDP: Generalizing to Multilingual Visual Information Extraction by Language Decoupled Pretraining

AAAI 2025technical

Visual Information Extraction (VIE) plays a crucial role in the comprehension of semi-structured documents, and several pre-trained models have been developed to enhance performance. However, most of these works are monolingual (usually English). Due to the extremely unbalanced quantity and quality…

Cited by 2SourcePDFScholar
2025

SSTAG: Structure-Aware Self-Supervised Learning Method for Text-Attributed Graphs

NeurIPS 2025poster

Large-scale pre-trained models have revolutionized Natural Language Processing (NLP) and Computer Vision (CV), showcasing remarkable cross-domain generalization abilities. However, in graph learning, models are typically trained on individual graph datasets, limiting their capacity to transfer knowl…

Cited by 0SourceScholar