← Search

Jonathan Tu

1 accepted papers

2023

Understanding the Inner-workings of Language Models Through Representation Dissimilarity

EMNLP 2023short main

As language models are applied to an increasing number of real-world applications, understanding their inner workings has become an important issue in model trust, interpretability, and transparency. In this work we show that representation dissimilarity measures, which are functions that measure t…

Cited by 0SourceScholar