← Search

Joonhan Park

1 accepted papers

2021

Attend What You Need: Motion-Appearance Synergistic Networks for Video Question Answering

ACL 2021long

Video Question Answering is a task which requires an AI agent to answer questions grounded in video. This task entails three key challenges: (1) understand the intention of various questions, (2) capturing various elements of the input video (e.g., object, action, causality), and (3) cross-modal gro…