← Search

Donghoon Han

3 accepted papers

2024

MERLIN: Multimodal Embedding Refinement via LLM-based Iterative Navigation for Text-Video Retrieval-Rerank Pipeline

EMNLP 2024industry

The rapid expansion of multimedia content has made accurately retrieving relevant videos from large collections increasingly challenging. Recent advancements in text-video retrieval have focused on cross-modal interactions, large-scale foundation model training, and probabilistic modeling, yet often…

2023

MixNeRF: Modeling a Ray With Mixture Density for Novel View Synthesis From Sparse Inputs

CVPR 2023poster

Neural Radiance Field (NeRF) has broken new ground in the novel view synthesis due to its simple concept and state-of-the-art quality. However, it suffers from severe performance degradation unless trained with a dense set of images with different camera poses, which hinders its practical applicatio…

2022

Few-Shot Image Generation with Mixup-Based Distance Learning

ECCV 2022poster

"Producing diverse and realistic images with generative models such as GANs typically requires large scale training with vast amount of images. GANs trained with limited data can easily memorize few training samples and display undesirable properties like ""stairlike"" latent space where interpolati…