← Search

Gaifan Zhang

2 accepted papers

2025

Annotating Training Data for Conditional Semantic Textual Similarity Measurement using Large Language Models

EMNLP 2025

Semantic similarity between two sentences depends on the aspects considered between those sentences. To study this phenomenon, Deshpande et al. (2023) proposed the Conditional Semantic Textual Similarity (C-STS) task and annotated a human-rated similarity dataset containing pairs of sentences compar

2024

Evaluating Unsupervised Dimensionality Reduction Methods for Pretrained Sentence Embeddings

COLING 2024main

Sentence embeddings produced by Pretrained Language Models (PLMs) have received wide attention from the NLP community due to their superior performance when representing texts in numerous downstream applications. However, the high dimensionality of the sentence embeddings produced by PLMs is problem…

Cited by 4SourcePDFScholar