← Search

Nikita Soni

6 accepted papers

2025

Evaluation of LLMs-based Hidden States as Author Representations for Psychological Human-Centered NLP Tasks

NAACL 2025findings

Like most of NLP, models for human-centered NLP tasks—tasks attempting to assess author-level information—predominantly use rep-resentations derived from hidden states of Transformer-based LLMs. However, what component of the LM is used for the representation varies widely. Moreover, there is a need…

2025

Residualized Similarity for Faithfully Explainable Authorship Verification

EMNLP 2025

Responsible use of Authorship Verification (AV) systems not only requires high accuracy but also interpretable solutions. More importantly, for systems to be used to make decisions with real-world consequences requires the model’s prediction to be explainable using interpretable features that can be

2025

Systematic Evaluation of Auto-Encoding and Large Language Model Representations for Capturing Author States and Traits

ACL 2025finding

Large Language Models (LLMs) are increasingly used in human-centered applications, yet their ability to model diverse psychological constructs is not well understood. In this study, we systematically evaluate a range of Transformer-LMs to predict psychological variables across five major dimensions:…

Cited by 0SourcePDFScholar
2024

Large Human Language Models: A Need and the Challenges

NAACL 2024long

As research in human-centered NLP advances, there is a growing recognition of the importance of incorporating human and social factors into NLP models. At the same time, our NLP systems have become heavily reliant on LLMs, most of which do not model authors. To build NLP systems that can truly under…

Cited by 9SourcePDFScholar
2021

MeLT: Message-Level Transformer with Masked Document Representations as Pre-Training for Stance Detection

EMNLP 2021finding

Much of natural language processing is focused on leveraging large capacity language models, typically trained over single messages with a task of predicting one or more tokens. However, modeling human language at higher-levels of context (i.e., sequences of messages) is under-explored. In stance de…