Shuwa Arabic voice dataset
Voice datasets are structured collections of audio recordings paired with corresponding text transcriptions, metadata, and annotations. These datasets serve as the foundation for training and evaluating speech recognition systems, text-to-speech engines, and other voice-enabled applications. High-quality voice datasets are essential for developing accurate and robust speech technologies. Voice models are machine learning systems trained on […]
Kanuri TTS and ASR models
Voice datasets are structured collections of audio recordings paired with corresponding text transcriptions, metadata, and annotations. These datasets serve as the foundation for training and evaluating speech recognition systems, text-to-speech engines, and other voice-enabled applications. High-quality voice datasets are essential for developing accurate and robust speech technologies. Voice models are machine learning systems trained on […]
Chichewa synthetic voice dataset, TTS models, ASR models
Below is a curated collection of open resources for text-to-speech (TTS), automatic speech recognition (ASR), and synthetic voice datasets in the Chichewa language. Text-to-Speech (TTS) Models Explore our collection of Chichewa TTS models, including XTTS and other multilingual models fine-tuned for natural-sounding Chichewa speech. View Chichewa TTS Models on Hugging Face Automatic Speech Recognition (ASR) Models […]
Hausa synthetic voice dataset, TTS models, ASR models
Below is a curated collection of open resources for text-to-speech (TTS), automatic speech recognition (ASR), and synthetic voice datasets in the Hausa language. Text-to-Speech (TTS) Models Explore our collection of Hausa TTS models, including XTTS and other multilingual models fine-tuned for natural-sounding Hausa speech. View Hausa TTS Models on Hugging Face Automatic Speech Recognition (ASR) […]
Dholuo synthetic voice dataset, TTS models, ASR models
Below is a curated collection of open resources for text-to-speech (TTS), automatic speech recognition (ASR), and synthetic voice datasets in the Dholuo language. Text-to-Speech (TTS) Models Explore our collection of Dholuo TTS models, including XTTS and other multilingual models fine-tuned for natural-sounding Dholuo speech. View Dholuo TTS Models on Hugging Face Automatic Speech Recognition (ASR) […]
Marma TTS and text data resources
Text data This dataset contains sentences in the Marma language (ISO code: rmz), with both original and normalized forms. The dataset is designed to support language technology development for the Marma language, a Tibeto-Burman language spoken primarily by the Marma people in Bangladesh and Myanmar. This dataset was created as part of a project funded […]