← Search

George Toderici

9 accepted papers

2025

Towards flexible perception with visual memory

ICML 2025poster

Training a neural network is a monolithic endeavor, akin to carving knowledge into stone: once the process is completed, editing the knowledge in a network is nearly impossible, since all information is distributed across the network's weights. We here explore a simple, compelling alternative by mar…

2023

Multi-Realism Image Compression With a Conditional Generator

CVPR 2023poster

By optimizing the rate-distortion-realism trade-off, generative compression approaches produce detailed, realistic images, even at low bit rates, instead of the blurry reconstructions produced by rate-distortion optimized models. However, previous methods do not explicitly control how much detail is…

Cited by 72SourcePDFScholar
2022

Neural Video Compression Using GANs for Detail Synthesis and Propagation

ECCV 2022poster

"We present the first neural video compression method based on generative adversarial networks (GANs). Our approach significantly outperforms previous neural and non-neural video compression methods in a user study, setting a new state-of-the-art in visual quality for neural methods. We show that th…

Cited by 51SourcePDFScholar
2022

VCT: A Video Compression Transformer

NeurIPS 2022accept

We show how transformers can be used to vastly simplify neural video compression. Previous methods have been relying on an increasing number of architectural biases and priors, including motion prediction and warping operations, resulting in complex models. Instead, we independently map input frames…

2020

Scale-Space Flow for End-to-End Optimized Video Compression

CVPR 2020poster

Despite considerable progress on end-to-end optimized deep networks for image compression, video coding remains a challenging task. Recently proposed methods for learned video compression use optical flow and bilinear warping for motion compensation and show competitive rate-distortion performance r…

Cited by 382PDFScholar
2018

AVA: A Video Dataset of Spatio-Temporally Localized Atomic Visual Actions

CVPR 2018poster

This paper introduces a video dataset of spatio-temporally localized Atomic Visual Actions (AVA). The AVA dataset densely annotates 80 atomic visual actions in 437 15-minute video clips, where actions are localized in space and time, resulting in 1.59M action labels with multiple labels per person o…

Cited by 1319SourcePDFScholar
2018

Improved Lossy Image Compression With Priming and Spatially Adaptive Bit Rates for Recurrent Networks

CVPR 2018poster

We propose a method for lossy image compression based on recurrent, convolutional neural networks that outper- forms BPG (4:2:0), WebP, JPEG2000, and JPEG as mea- sured by MS-SSIM. We introduce three improvements over previous research that lead to this state-of-the-art result us- ing a single model…

Cited by 483SourcePDFScholar
2017

Full Resolution Image Compression With Recurrent Neural Networks

CVPR 2017oral

This paper presents a set of full-resolution lossy image compression methods based on neural networks. Each of the architectures we describe can provide variable compression rates during deployment without requiring retraining of the network: each network need only be trained once. All of our archit…

Cited by 1111PDFScholar
2015

Beyond Short Snippets: Deep Networks for Video Classification

CVPR 2015poster

Convolutional neural networks (CNNs) have been exten- sively applied for image recognition problems giving state- of-the-art results on recognition, detection, segmentation and retrieval. In this work we propose and evaluate several deep neural network architectures to combine image infor- mation ac…

Cited by 3205SourcePDFScholar