2020
CLEVRER: Collision Events for Video Representation and Reasoning
ICLR 2020spotlight
The ability to reason about temporal and causal events from videos lies at the core of human intelligence. Most video reasoning benchmarks, however, focus on pattern recognition from complex visual and language input, instead of on causal structure. We study the complementary problem, exploring the…