← Search

Santosh Divvala

4 accepted papers

2019

Video Relationship Reasoning Using Gated Spatio-Temporal Energy Graph

CVPR 2019poster

Visual relationship reasoning is a crucial yet challenging task for understanding rich interactions across visual concepts. For example, a relationship \ man, open, door\ involves a complex relation \ open\ between concrete entities \ man, door\ . While much of the existing work has studied this p…

Cited by 127PDFcodeScholar
2018

DOCK: Detecting Objects by transferring Common-sense Knowledge

ECCV 2018poster

We present a scalable approach for Detecting Objects by transferring Common-sense Knowledge (DOCK) from source to target categories. In our setting, the training data for the source categories have bounding box annotations, while those for the target categories only have image-level annotations. Cur…

Cited by 42SourcePDFScholar
2017

Asynchronous Temporal Fields for Action Recognition

CVPR 2017poster

Actions are more than just movements and trajectories: we cook to eat and we hold a cup to drink from it. A thorough understanding of videos requires going beyond appearance modeling and necessitates reasoning about the sequence of activities, as well as the higher-level constructs such as intention…

Cited by 212PDFcodeScholar
2016

You Only Look Once: Unified, Real-Time Object Detection

CVPR 2016oral

We present YOLO, a new approach to object detection. Prior work on object detection repurposes classifiers to perform detection. Instead, we frame object detection as a regression problem to spatially separated bounding boxes and associated class probabilities. A single neural network predicts bound…

Cited by 61920PDFcodeScholar