← Search

Jonathan Brandt

7 accepted papers

2024

TurboEdit: Real-time text-based disentangled real image editing

ECCV 2024poster

"We address the challenges of precise image inversion and disentangled image editing in the context of few-step diffusion models. We introduce an encoder based iterative inversion technique. The inversion network is conditioned on the input image and the reconstructed image from the previous step, a…

Cited by 0SourcePDFScholar
2021

AESOP: Abstract Encoding of Stories, Objects, and Pictures

ICCV 2021poster

Visual storytelling and story comprehension are uniquely human skills that play a central role in how we learn about and experience the world. Despite remarkable progress in recent years in synthesis of visual and textual content in isolation and learning effective joint visual-linguistic representa…

Cited by 19PDFcodeScholar
2021

StreamHover: Livestream Transcript Summarization and Annotation

EMNLP 2021main

With the explosive growth of livestream broadcasting, there is an urgent need for new summarization technology that enables us to create a preview of streamed content and tap into this wealth of knowledge. However, the problem is nontrivial due to the informal nature of spoken language. Further, the…

2017

Spatial-Semantic Image Search by Visual Feature Synthesis

CVPR 2017spotlight

The performance of image retrieval has been improved tremendously in recent years through the use of deep feature representations. Most existing methods, however, aim to retrieve images that are visually similar or semantically relevant to the query, irrespective of spatial configuration. In this pa…

Cited by 52PDFcodeScholar
2016

A Multi-Level Contextual Model For Person Recognition in Photo Albums

CVPR 2016poster

In this work, we present a new framework for person recognition in photo albums that exploits contextual cues at multiple levels, spanning individual persons, individual photos, and photo groups. Through experiments, we show that the information available at each of these distinct contextual levels…

Cited by 39PDFScholar
2016

Shortlist Selection With Residual-Aware Distance Estimator for K-Nearest Neighbor Search

CVPR 2016poster

In this paper, we introduce a novel shortlist computation algorithm for approximate, high-dimensional nearest neighbor search. Our method relies on a novel distance estimator: the residual-aware distance estimator, that accounts for the residual distances of data points to their respective quantized…

Cited by 13PDFScholar
2015

A Convolutional Neural Network Cascade for Face Detection

CVPR 2015poster

In real-world face detection, large visual variations, such as those due to pose, expression, and lighting, demand an advanced discriminative model to accurately differentiate faces from the backgrounds. Consequently, effective models for the problem tend to be computationally prohibitive. To addre…

Cited by 1844SourcePDFScholar