← Search

Esma Balkir

2 accepted papers

2023

This prompt is measuring <mask>: evaluating bias evaluation in language models

ACL 2023findings

Bias research in NLP seeks to analyse models for social biases, thus helping NLP practitioners uncover, measure, and mitigate social harms. We analyse the body of work that uses prompts and templates to assess bias in language models. We draw on a measurement modelling framework to create a taxonomy…

Cited by 34SourcePDFScholar
2022

Necessity and Sufficiency for Explaining Text Classifiers: A Case Study in Hate Speech Detection

NAACL 2022long

We present a novel feature attribution method for explaining text classifiers, and analyze it in the context of hate speech detection. Although feature attribution models usually provide a single importance score for each token, we instead provide two complementary and theoretically-grounded scores…