← Search

Tae-Hyun Oh*

2 accepted papers

2024

BEAF: Observing BEfore-AFter Changes to Evaluate Hallucination in Vision-language Models

ECCV 2024poster

"Vision language models (VLMs) perceive the world through a combination of a visual encoder and a large language model (LLM). The visual encoder, pre-trained on large-scale vision-text datasets, provides zero-shot generalization to visual data, and the LLM endows its high reasoning ability to VLMs.…

2024

Learning-based Axial Video Motion Magnification

ECCV 2024poster

"Video motion magnification amplifies invisible small motions to be perceptible, which provides humans with a spatially dense and holistic understanding of small motions in the scene of interest. This is based on the premise that magnifying small motions enhances the legibility of motions. In the re…