← Search

Genevieve Smith

3 accepted papers

2026

Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to Models

ICLR 2026poster

Vision-language models trained on large-scale multimodal datasets show strong demographic biases, but the role of training data in producing these biases remains unclear. A major barrier has been the lack of demographic annotations in web-scale datasets such as LAION-400M. We address this gap by cre…

Cited by 0SourceScholar
2024

Linguistic Bias in ChatGPT: Language Models Reinforce Dialect Discrimination

EMNLP 2024main

We present a large-scale study of linguistic bias exhibited by ChatGPT covering ten dialects of English (Standard American English, Standard British English, and eight widely spoken non-”standard” varieties from around the world). We prompted GPT-3.5 Turbo and GPT-4 with text by native speakers of e…

Cited by 24SourcePDFScholar
2024

Position: Near to Mid-term Risks and Opportunities of Open-Source Generative AI

ICML 2024oral

In the next few years, applications of Generative AI are expected to revolutionize a number of different areas, ranging from science & medicine to education. The potential for these seismic changes has triggered a lively debate about potential risks and resulted in calls for tighter regulation, in p…

Cited by 9SourcePDFScholar