← Search

Juntao Chen

5 accepted papers

2026

SEA-Vision: A Multilingual Benchmark for Comprehensive Document and Scene Text Understanding in Southeast Asia

CVPR 2026

Multilingual document and scene text understanding plays an important role in applications such as search, finance, and public services. However, most existing benchmarks focus on high-resource languages and fail to evaluate models in realistic multilingual environments. In Southeast Asia, the diver

Cited by 0SourcecodeScholar
2026

Stackelberg Coupling of Online Representation Learning and Reinforcement Learning

ICLR 2026poster

Deep Q-learning jointly learns representations and values within monolithic networks, promising beneficial co-adaptation between features and value estimates. Although this architecture has attained substantial success, the coupling between representation and value learning creates instability as re…

Cited by 0SourceScholar
2025

ATP: Adaptive Threshold Pruning for Efficient Data Encoding in Quantum Neural Networks

CVPR 2025poster

Quantum Neural Networks (QNNs) offer promising capabilities for complex data tasks, but are often constrained by limited qubit resources and high entanglement, which can hinder scalability and efficiency. In this paper, we introduce Adaptive Threshold Pruning (ATP), an encoding method that reduces e…

Cited by 0SourcePDFScholar
2025

Leveraging Debiased Cross-modal Attention Maps and Code-based Reasoning for Zero-shot Referring Expression Comprehension

ICCV 2025poster

Zero-shot Referring Expression Comprehension (REC) aims at locating an object described by a natural language query without training on task-specific datasets. Current approaches often utilize Vision-Language Models (VLMs) to perform region-text matching based on region proposals. However, this may…

Cited by 0SourcePDFScholar