← Search

Sachit Menon

9 accepted papers

2025

MINERVA: Evaluating Complex Video Reasoning

ICCV 2025poster

Multimodal LLMs are turning their focus to video benchmarks, however most video benchmarks only provide outcome supervision, with no intermediate or interpretable reasoning steps. This makes it challenging to assess if models are truly able to combine perceptual and temporal information to reason ab…

2023

Doubly Right Object Recognition: A Why Prompt for Visual Rationales

CVPR 2023poster

Many visual recognition models are evaluated only on their classification accuracy, a metric for which they obtain strong performance. In this paper, we investigate whether computer vision models can also provide correct rationales for their predictions. We propose a "doubly right" object recognitio…

2023

What You Can Reconstruct From a Shadow

CVPR 2023poster

3D reconstruction is a fundamental problem in computer vision, and the task is especially challenging when the object to reconstruct is partially or fully occluded. We introduce a method that uses the shadows cast by an unobserved object in order to infer the possible 3D volumes under occlusion. We…

Cited by 3SourcePDFScholar
2020

PULSE: Self-Supervised Photo Upsampling via Latent Space Exploration of Generative Models

CVPR 2020poster

The primary aim of single-image super-resolution is to construct a high-resolution (HR) image from a corresponding low-resolution (LR) input. In previous approaches, which have generally been supervised, the training objective typically measures a pixel-wise average distance between the super-resolv…

Cited by 679PDFcodeScholar