← Search

Anish Madan

3 accepted papers

2025

Roboflow100-VL: A Multi-Domain Object Detection Benchmark for Vision-Language Models

NeurIPS 2025poster

Vision-language models (VLMs) trained on internet-scale data achieve remarkable zero-shot detection performance on common objects like car, truck, and pedestrian. However, state-of-the-art models still struggle to generalize to out-of-distribution classes, tasks and imaging modalities not typically…

Cited by 0SourcecodeScholar
2024

Revisiting Few-Shot Object Detection with Vision-Language Models

NeurIPS 2024poster

The era of vision-language models (VLMs) trained on web-scale datasets challenges conventional formulations of “open-world" perception. In this work, we revisit the task of few-shot object detection (FSOD) in the context of recent foundational VLMs. First, we point out that zero-shot predictions fro…