SPCS Speech Corpus
License agreement
By downloading this resource I accept and agree to the terms of use and the associated license conditions under which the resource is distributed.
Download
MD5: 2c2b367ba1811e4024b52b95bf85acf8
License agreement
By downloading this resource I accept and agree to the terms of use and the associated license conditions under which the resource is distributed.
Collections
- Resource Catalogue [350]
- Resource Index [412]
Author(s)
Modipa, T. I.
Davel, M. H.
De Wet, F.
Metadata
Show full item recordDescription
Broadband speech corpus of approximately 10 hours and the corresponding transcriptions.
The development process of the corpus involved the recording and transcribing of radio broadcasts. The transcriptions were used to generate the Sepedi code-switched prompts to re-record speech from multiple speakers.
The following sub-directories are found in this directory:
Audio: Audio files for all the recorded code-switched speech
Transcriptions: The corresponding orthographic transcriptions
Metadata: Information about the speakers and the transcriptions
Documentation: The directory structure and the Sepedi prompt list
Contact person
Ulrike JankeContact person's e-mail address
ulrike.must@gmail.comPublisher(s)
Council for Scientific and Industrial Research
North-West University