← Search

Khoi Duc Nguyen

4 accepted papers

2025

Learning to Inference Adaptively for Multimodal Large Language Models

ICCV 2025poster

Multimodal Large Language Models (MLLMs) have shown impressive capabilities in visual reasoning, yet come with substantial computational cost, limiting their deployment in resource-constrained settings. Despite recent effort on improving the efficiency of MLLMs, prior solutions fall short in respond…

Cited by 0SourcePDFScholar
2025

PAVE: Patching and Adapting Video Large Language Models

CVPR 2025poster

We present PAVE, a framework for adapting pre-trained video large language models (Video-LLMs) to downstream tasks that incorporate side-channel signals, such as audio, camera pose, or high frame rate videos. PAVE introduces a lightweight adaptation strategy called "patching", which adds a small num…

2024

ESCAPE: Encoding Super-keypoints for Category-Agnostic Pose Estimation

CVPR 2024poster

In this paper we tackle the task of category-agnostic pose estimation (CAPE) which aims to predict poses for objects of any category with few annotated samples. Previous works either rely on local matching between features of support and query samples or require support keypoint identifier. The form…

2021

POODLE: Improving Few-shot Learning via Penalizing Out-of-Distribution Samples

NeurIPS 2021poster

In this work, we propose to use out-of-distribution samples, i.e., unlabeled samples coming from outside the target classes, to improve few-shot learning. Specifically, we exploit the easily available out-of-distribution samples to drive the classifier to avoid irrelevant features by maximizing the…