2026
JRDB-Reasoning: A Difficulty-Graded Benchmark for Visual Reasoning in Robotics
AAAI 2026technical
Recent advances in Vision-Language Models (VLMs) and large language models (LLMs) have greatly enhanced visual reasoning, a key capability for embodied AI agents like robots. However, existing visual reasoning benchmarks often suffer from several limitations: they lack a clear definition of reasonin