2025
Beyond Distribution: Investigating Language Models’ Understanding of Sino-Korean Morphemes
EMNLP 2025
We investigate whether Transformer-based language models, trained solely on Hangul text, can learn the compositional morphology of Sino-Korean (SK) morphemes, which are fundamental to Korean vocabulary. Using BERT_BASE and fastText, we conduct controlled experiments with target words and their “real