← Search

Lewis Griffin

3 accepted papers

2022

How to Stay Curious while avoiding Noisy TVs using Aleatoric Uncertainty Estimation

ICML 2022spotlight

When extrinsic rewards are sparse, artificial agents struggle to explore an environment. Curiosity, implemented as an intrinsic reward for prediction errors, can improve exploration but it is known to fail when faced with action-dependent noise sources (‘noisy TVs’). In an attempt to make exploring…

2022

Identifying Human Strategies for Generating Word-Level Adversarial Examples

EMNLP 2022finding

Adversarial examples in NLP are receiving increasing research attention. One line of investigation is the generation of word-level adversarial examples against fine-tuned Transformer models that preserve naturalness and grammaticality. Previous work found that human- and machine-generated adversaria…

Cited by 4SourcePDFScholar
2021

Contrasting Human- and Machine-Generated Word-Level Adversarial Examples for Text Classification

EMNLP 2021main

Research shows that natural language processing models are generally considered to be vulnerable to adversarial attacks; but recent work has drawn attention to the issue of validating these adversarial inputs against certain criteria (e.g., the preservation of semantics and grammaticality). Enforcin…