Greedy Hierarchical Variational Autoencoders for Large-Scale Video Prediction
A video prediction model that generalizes to diverse scenes would enable intelligent agents such as robots to perform a variety of tasks via planning with the model. However, while existing video prediction models have produced promising results on small datasets, they suffer from severe underfittin…