← Search

Jinbae Im

4 accepted papers

2025

MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models

ACL 2025finding

This paper introduces MMRefine, a MultiModal Refinement benchmark designed to evaluate the error refinement capabilities of Multimodal Large Language Models (MLLMs). As the emphasis shifts toward enhancing reasoning during inference, MMRefine provides a framework that evaluates MLLMs’ abilities to d…

2024

EGTR: Extracting Graph from Transformer for Scene Graph Generation

CVPR 2024poster

Scene Graph Generation (SGG) is a challenging task of detecting objects and predicting relationships between objects. After DETR was developed one-stage SGG models based on a one-stage object detector have been actively studied. However complex modeling is used to predict the relationship between ob…

2021

Self-Supervised Multimodal Opinion Summarization

ACL 2021long

Recently, opinion summarization, which is the generation of a summary from multiple reviews, has been conducted in a self-supervised manner by considering a sampled review as a pseudo summary. However, non-text data such as image and metadata related to reviews have been considered less often. To us…