This dataset contains the texts of the Arabic and Latin Corpus. It only contains texts that, to our knowledge, were free of copyright at the time of the last update. Unlike many of the files published
The ROBIN Technical Acquisition Speech Corpus (ROBINTASC) was developed within the ROBIN project. Its main purpose was to improve the behaviour of a conversational agent, allowing human-machine intera
VINKO is a spoken corpus based on crowd-sourced audio recordings that has been designed to provide relevant linguistic information about the minority languages and dialects...