2025
RESPIN-S1.0: A read speech corpus of 10000+ hours in dialects of nine Indian Languages
NeurIPS 2025poster
We introduce **RESPIN-S1.0**, the largest publicly available dialect-rich read-speech corpus for Indian languages, comprising more than 10,000 hours of validated audio across nine major languages: Bengali, Bhojpuri, Chhattisgarhi, Hindi, Kannada, Magahi, Maithili, Marathi, and Telugu. Indian languag…