← Search

Xiaoxuan Fan

3 accepted papers

2026

InteractBench: Benchmarking LLMs on Competitive Programming under Unrevealed Information

ICML 2026poster

Competitive programming is increasingly being used to evaluate the algorithmic reasoning capabilities of large language models (LLMs). However, existing benchmarks primarily focus on full-information tasks where all problem inputs are provided upfront. This overlooks a critical dimension of algorith…

Cited by 0SourceScholar
2025

MM-OPERA: Benchmarking Open-ended Association Reasoning for Large Vision-Language Models

NeurIPS 2025poster

Large Vision-Language Models (LVLMs) have exhibited remarkable progress. However, deficiencies remain compared to human intelligence, such as hallucination and shallow pattern matching. In this work, we aim to evaluate a fundamental yet underexplored intelligence: association, a cornerstone of human…

Cited by 0SourcecodeScholar
2025

Reproducible Vision-Language Models Meet Concepts Out of Pre-Training

CVPR 2025poster

Contrastive Language-Image Pre-training (CLIP) models as a milestone of modern multimodal intelligence, its generalization mechanism grasped massive research interests in the community. While existing studies limited in the scope of pre-training knowledge, hardly underpinned its generalization to co…

Cited by 0SourcePDFScholar