← Search

Soochahn Lee

4 accepted papers

2026

Contamination Detection for VLMs Using Multi‑Modal Semantic Perturbations

ICLR 2026poster

Recent advances in Vision–Language Models (VLMs) have achieved state-of-the-art performance on numerous benchmark tasks. However, the use of internet-scale, often proprietary, pretraining corpora raises a critical concern for both practitioners and users: inflated performance due to \emph{test-set l…

Cited by 0SourcecodeScholar
2026

DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents

ICML 2026poster

Multimodal Large Language Models (MLLMs) have achieved strong performance on structured visual understanding tasks such as chart and document question answering. However, existing benchmarks typically evaluate these domains in isolation, overlooking realistic settings where numerical evidence in cha…

Cited by 0SourceScholar
2015

Random Tree Walk Toward Instantaneous 3D Human Pose Estimation

CVPR 2015poster

The availability of accurate depth cameras have made real-time human pose estimation possible; however, there are still demands for faster algorithms on low power processors. This paper introduces 1000 frames per second pose estimation method on a single core CPU. A large computation gain is achieve…

Cited by 124SourcePDFScholar