← Search

Minsu Park

4 accepted papers

2026

MAVIS: A Benchmark for Multimodal Source Attribution in Long-form Visual Question Answering

AAAI 2026technical

Source attribution aims to enhance the reliability of AI-generated answers by including references for each statement, helping users validate the provided answers. However, existing work has primarily focused on text-only scenario and largely overlooked the role of multimodality. We introduce MAVIS,

Cited by 0SourcePDFScholar
2025

Large Language Models Are Natural Video Popularity Predictors

ACL 2025finding

Predicting video popularity is often framed as a supervised learning task, relying heavily on meta-information and aggregated engagement data. However, video popularity is shaped by complex cultural and social factors that such approaches often overlook. We argue that Large Language Models (LLMs), w…

2025

Locally Convex Global Loss Network for Decision-Focused Learning

AAAI 2025technical

In decision-making problems under uncertainty, predicting unknown parameters is often considered independent of the optimization part. Decision-focused learning (DFL) is a task-oriented framework that integrates prediction and optimization by adapting the predictive model to give better decisions fo…

2025

Solving Copyright Infringement on Short Video Platforms: Novel Datasets and an Audio Restoration Deep Learning Pipeline

IJCAI 2025

Short video platforms like YouTube Shorts and TikTok face significant copyright compliance challenges, as infringers frequently embed arbitrary background music (BGM) to obscure original soundtracks (OST) and evade content originality detection. To tackle this issue, we propose a novel pipeline that

Cited by 0SourcePDFScholar