← Search

Inho Won

3 accepted papers

2025

VLR-Bench: Multilingual Benchmark Dataset for Vision-Language Retrieval Augmented Generation

COLING 2025main

We propose the VLR-Bench, a visual question answering (VQA) benchmark for evaluating vision language models (VLMs) based on retrieval augmented generation (RAG). Unlike existing evaluation datasets for external knowledge-based VQA, the proposed VLR-Bench includes five input passages. This allows tes…

Cited by 1SourcePDFScholar
2024

Optimizing Language Augmentation for Multilingual Large Language Models: A Case Study on Korean

COLING 2024main

Large language models (LLMs) use pretraining to predict the subsequent word; however, their expansion requires significant computing resources. Numerous big tech companies and research institutes have developed multilingual LLMs (MLLMs) to meet current demands, overlooking less-resourced languages (…

2024

X-LLaVA: Optimizing Bilingual Large Vision-Language Alignment

NAACL 2024findings

The impressive development of large language models (LLMs) is expanding into the realm of large multimodal models (LMMs), which incorporate multiple types of data beyond text. However, the nature of multimodal models leads to significant expenses in the creation of training data. Furthermore, constr…