← Search

Jinkun Hou

1 accepted papers

2026

GameVerse: Can Vision-Language Models Learn from Video-based Reflection?

ICML 2026poster

Human gameplay is a visually grounded interaction loop in which players act, reflect on failures, and watch tutorials to refine strategies. Can Vision-Language Models (VLMs) also learn from video-based reflection? We present **GameVerse**, a comprehensive video game benchmark that enables a *reflect…

Cited by 0SourceScholar