← Search

Austin Myers

5 accepted papers

2024

Distribution Aware Metrics for Conditional Natural Language Generation

COLING 2024main

Traditional automated metrics for evaluating conditional natural language generation rely on pairwise comparisons between a single generated text and the best-matching gold-standard reference. This method is effective when ground truth data diversity can be attributed to noise, however, it falls sho…

Cited by 7SourcePDFScholar
2024

Streaming Dense Video Captioning

CVPR 2024poster

An ideal model for dense video captioning -- predicting captions localized temporally in a video -- should be able to handle long input videos predict rich detailed textual descriptions and be able to produce outputs before processing the entire video. Current state-of-the-art models however process…

2023

IC3: Image Captioning by Committee Consensus

EMNLP 2023long main

If you ask a human to describe an image, they might do so in a thousand different ways. Traditionally, image captioning models are trained to generate a single "best" (most like a reference) image caption. Unfortunately, doing so encourages captions that are "informationally impoverished," and focus…

Cited by 0SourcecodeScholar
2019

VideoBERT: A Joint Model for Video and Language Representation Learning

ICCV 2019poster

Self-supervised learning has become increasingly important to leverage the abundance of unlabeled data available on platforms like YouTube. Whereas most existing approaches learn low-level representations, we propose a joint visual-linguistic model to learn high-level features without any explicit s…

Cited by 1568PDFScholar
2015

Affordance detection of tool parts from geometric features

ICRA 2015poster

As robots begin to collaborate with humans in everyday workspaces, they will need to understand the functions of tools and their parts. To cut an apple or hammer a nail, robots need to not just know the tool's name, but they must localize the tool's parts and identify their functions. Intuitively, t…

Cited by 386SourceScholar