← Search

Marcin Gruza

3 accepted papers

2023

Massively Multilingual Corpus of Sentiment Datasets and Multi-faceted Sentiment Classification Benchmark

NeurIPS 2023poster

Despite impressive advancements in multilingual corpora collection and model training, developing large-scale deployments of multilingual models still presents a significant challenge. This is particularly true for language tasks that are culture-dependent. One such example is the area of multilingu…

Cited by 20SourcePDFScholar
2023

PALS: Personalized Active Learning for Subjective Tasks in NLP

EMNLP 2023long main

For subjective NLP problems, such as classification of hate speech, aggression, or emotions, personalized solutions can be exploited. Then, the learned models infer about the perception of the content independently for each reader. To acquire training data, texts are commonly randomly assigned to us…

Cited by 0SourcecodeScholar
2021

Controversy and Conformity: from Generalized to Personalized Aggressiveness Detection

ACL 2021long

There is content such as hate speech, offensive, toxic or aggressive documents, which are perceived differently by their consumers. They are commonly identified using classifiers solely based on textual content that generalize pre-agreed meanings of difficult problems. Such models provide the same r…