← Search

Junbo Zou

3 accepted papers

2026

SportR: A Benchmark for Multimodal Large Language Model Reasoning in Sports

ICLR 2026poster

Artificial Intelligence brings powerful new tools to sports, from automated officiating to tactical analysis, but these applications all depend on a core reasoning capability. Deeply understanding sports requires an intricate blend of fine-grained visual perception and rule-based reasoning—a challe…

Cited by 0SourcecodeScholar
2026

VideoBrain: Learning Adaptive Frame Sampling for Long Video Understanding

ICML 2026poster

Long-form video understanding remains challenging for Vision-Language Models (VLMs) due to the inherent tension between computational constraints and the need to capture information distributed across thousands of frames. Existing approaches either sample frames uniformly (risking information loss) …

Cited by 5SourceScholar
2025

SPORTU: A Comprehensive Sports Understanding Benchmark for Multimodal Large Language Models

ICLR 2025poster

Multimodal Large Language Models (MLLMs) are advancing the ability to reason about complex sports scenarios by integrating textual and visual information. To comprehensively evaluate their capabilities, we introduce SPORTU, a benchmark designed to assess MLLMs across multi-level sports reasoning tas…