2025
SynDARin: Synthesising Datasets for Automated Reasoning in Low-Resource Languages
COLING 2025main
Question Answering (QA) datasets have been instrumental in developing and evaluating Large Language Model (LLM) capabilities. However, such datasets are scarce for languages other than English due to the cost and difficulties of collection and manual annotation. This means that producing novel model…