2025
FicSim: A Dataset for Multi-Faceted Semantic Similarity in Long-Form Fiction
EMNLP 2025
As language models become capable of processing increasingly long and complex texts, there has been growing interest in their application within computational literary studies. However, evaluating the usefulness of these models for such tasks remains challenging due to the cost of fine-grained annot