← Search

Iqra Ali

2 accepted papers

2025

HLU: Human Vs LLM Generated Text Detection Dataset for Urdu at Multiple Granularities

COLING 2025main

The rise of large language models (LLMs) generating human-like text has raised concerns about misuse, especially in low-resource languages like Urdu. To address this gap, we introduce the HLU dataset, which consists of three datasets: Document, Paragraph, and Sentence level. The document-level datas…

Cited by 0SourcePDFScholar
2024

Monolingual Paraphrase Detection Corpus for Low Resource Pashto Language at Sentence Level

COLING 2024main

Paraphrase detection is a task to identify if two sentences are semantically similar or not. It plays an important role in maintaining the integrity of written work such as plagiarism detection and text reuse detection. Formerly, researchers focused on developing large corpora for English. However,…

Cited by 3SourcePDFScholar