2025
TounsiBench: Benchmarking Large Language Models for Tunisian Arabic
EMNLP 2025
In this work, we introduce the first benchmark for evaluating the capabilities of large language models (LLMs) in understanding and generating responses in Tunisian Arabic. To achieve this, we construct a dataset of Tunisian Arabic instructions and prompt ten widely-used LLMs that claim to support A