2023
Improving Visual Representation Learning Through Perceptual Understanding
CVPR 2023poster
We present an extension to masked autoencoders (MAE) which improves on the representations learnt by the model by explicitly encouraging the learning of higher scene-level features. We do this by: (i) the introduction of a perceptual similarity term between generated and real images (ii) incorporati…