← Search

Shashank Shekhar

5 accepted papers

2023

PUG: Photorealistic and Semantically Controllable Synthetic Data for Representation Learning

NeurIPS 2023poster

Synthetic image datasets offer unmatched advantages for designing and evaluating deep neural networks: they make it possible to (i) render as many data samples as needed, (ii) precisely control each scene and yield granular ground truth labels (and captions), (iii) precisely control distribution shi…

2022

Beyond neural scaling laws: beating power law scaling via data pruning

NeurIPS 2022accept

Widely observed neural scaling laws, in which error falls off as a power of the training set size, model size, or both, have driven substantial performance improvements in deep learning. However, these improvements through scaling alone require considerable costs in compute and energy. Here we focus…

2021

Context-Aware Scene Graph Generation With Seq2Seq Transformers

ICCV 2021poster

Scene graph generation is an important task in computer vision aimed at improving the semantic understand- ing of the visual world. In this task, the model needs to detect objects and predict visual relationships between them. Most of the existing models predict relationships in parallel assuming th…

Cited by 98PDFcodeScholar
2021

Improved Knowledge Modeling and Its Use for Signaling in Multi-Agent Planning with Partial Observability

AAAI 2021technical

Collaborative Multi-Agent Planning (MAP) problems with uncertainty and partial observability are often modeled as Dec-POMDPs. Yet, in deterministic domains, Qualitative Dec-POMDPs can scale up to much larger problem sizes. The best current QDec solver (QDec-FP) reduces MAP problems to multiple singl…

Cited by 5SourcePDFScholar
2019

From Strings to Things: Knowledge-Enabled VQA Model That Can Read and Reason

ICCV 2019oral

Text present in images are not merely strings, they provide useful cues about the image. Despite their utility in better image understanding, scene texts are not used in traditional visual question answering (VQA) models. In this work, we present a VQA model which can read scene texts and perform re…

Cited by 64PDFScholar