← Search

Maxwell Forbes

9 accepted papers

2022

Is GPT-3 Text Indistinguishable from Human Text? Scarecrow: A Framework for Scrutinizing Machine Text

ACL 2022long

Modern neural language models can produce remarkably fluent and grammatical text. So much, in fact, that recent work by Clark et al. (2021) has reported that conventional crowdsourcing can no longer reliably distinguish between machine-authored (GPT-3) and human-authored writing. As errors in machin…

2021

CLIPScore: A Reference-free Evaluation Metric for Image Captioning

EMNLP 2021main

Image captioning has conventionally relied on reference-based automatic evaluations, where machine captions are compared against captions written by humans. This is in contrast to the reference-free manner in which humans assess caption quality. In this paper, we report the surprising empirical find…

2021

Edited Media Understanding Frames: Reasoning About the Intent and Implications of Visual Misinformation

ACL 2021long

Understanding manipulated media, from automatically generated ‘deepfakes’ to manually edited ones, raises novel research challenges. Because the vast majority of edited or manipulated images are benign, such as photoshopped images for visual enhancements, the key challenge is to understand the compl…

Cited by 17SourcePDFScholar
2021

Moral Stories: Situated Reasoning about Norms, Intents, Actions, and their Consequences

EMNLP 2021main

In social settings, much of human behavior is governed by unspoken rules of conduct rooted in societal norms. For artificial systems to be fully integrated into social environments, adherence to such norms is a central prerequisite. To investigate whether language generation models can serve as beha…

2021

MultiTalk: A Highly-Branching Dialog Testbed for Diverse Conversations

AAAI 2021technical

We study conversational dialog in which there are many possible responses to a given history. We present the MultiTalk Dataset, a corpus of over 320,000 sentences of written conversational dialog that balances a high branching factor (10) with several conversation turns (6) through selective branch…

Cited by 11SourcePDFScholar
2021

Paragraph-level Commonsense Transformers with Recurrent Memory

AAAI 2021technical

Human understanding of narrative texts requires making commonsense inferences beyond what is stated in the text explicitly. A recent model, COMET, can generate such inferences along several dimensions such as pre- and post-conditions, motivations, and mental states of the participants. However, COME…

Cited by 46SourcePDFScholar
2015

Robot Programming by Demonstration with situated spatial language understanding

ICRA 2015poster

Robot Programming by Demonstration (PbD) allows users to program a robot by demonstrating the desired behavior. Providing these demonstrations typically involves moving the robot through a sequence of states, often by physically manipulating it. This requires users to be co-located with the robot an…

Cited by 82SourceScholar