← Search

Wamiq Reyaz Para

4 accepted papers

2026

SigLino: Efficient Multi-Teacher Distillation for Agglomerative Vision Foundation Models

CVPR 2026

Vision foundation models trained via multi-teacher distillation offer a promising path toward unified visual representations, yet the learning dynamics and data efficiency of such approaches remain underexplored. In this paper, we systematically study multi-teacher distillation for vision foundation

Cited by 0SourcecodeScholar
2026

VisRes Bench: On Evaluating the Visual Reasoning Capabilities of VLMs

CVPR 2026

Vision-Language Models (VLMs) have achieved remarkable progress across tasks such as visual question answering and image captioning. Yet, the extent to which these models perform visual reasoning as opposed to relying on linguistic priors remains unclear. To address this, we introduce VisRes Bench,

Cited by 0SourcecodeScholar
2025

Vision-Language Models Can't See the Obvious

ICCV 2025poster

We present Saliency Benchmark (SalBench), a novel benchmark designed to assess the capability of Large Vision-Language Models (LVLM) in detecting visually salient features that are readily apparent to humans, such as a large circle amidst a grid of smaller ones. This benchmark focuses on low-level f…

Cited by 0SourcePDFScholar
2021

SketchGen: Generating Constrained CAD Sketches

NeurIPS 2021poster

Computer-aided design (CAD) is the most widely used modeling approach for technical design. The typical starting point in these designs is 2D sketches which can later be extruded and combined to obtain complex three-dimensional assemblies. Such sketches are typically composed of parametric primitive…

Cited by 82SourcePDFScholar