← Search

Anshul Shah

6 accepted papers

2026

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

ICML 2026poster

Reinforcement learning (RL) fine-tuning is now widely used to improve LLM reasoning, and recent work has begun extending it to vision-language models (VLMs). While RL-tuned VLMs can improve visual reasoning benchmark performance, they can still suffer from weak visual grounding, hallucinations, and …

Cited by 0SourceScholar
2023

HaLP: Hallucinating Latent Positives for Skeleton-Based Self-Supervised Learning of Actions

CVPR 2023poster

Supervised learning of skeleton sequence encoders for action recognition has received significant attention in recent times. However, learning such encoders without labels continues to be a challenging problem. While prior works have shown promising results by applying contrastive learning to pose s…

2023

STEPs: Self-Supervised Key Step Extraction and Localization from Unlabeled Procedural Videos

ICCV 2023poster

We address the problem of extracting key steps from unlabeled procedural videos, motivated by the potential of Augmented Reality (AR) headsets to revolutionize job training and performance. We decompose the problem into two steps: representation learning and key steps extraction. We propose a traini…

Cited by 11PDFcodeScholar
2022

FeLMi : Few shot Learning with hard Mixup

NeurIPS 2022accept

Learning from a few examples is a challenging computer vision task. Traditionally, meta-learning-based methods have shown promise towards solving this problem. Recent approaches show benefits by learning a feature extractor on the abundant base examples and transferring these to the fewer novel exam…

Cited by 34SourcePDFScholar