All projects
Open SourceAI

YorubaVoice

Speech recognition for Yoruba, trained on community data

2.8K views
YorubaVoice cover

Overview

YorubaVoice is an open speech-to-text model and dataset for Yoruba, built with community contributors.

The problem

Voice interfaces exclude tens of millions of Yoruba speakers because no usable open model exists.

The solution

A community recording app collects consented, labelled speech, and a fine-tuned model is released openly with benchmarks.

Who it's for

Developers building voice products for West African markets.

How it works

Contributors read prompts in the app, recordings are validated by peers, and validated audio feeds a fine-tuning pipeline with public evaluation.

What's next

Expand to Igbo and Hausa, and publish a hosted inference API.

Screenshots

Overview
Overview

Ship log

Progress updates from the builder.

    Feedback

    Sign in to leave feedback for this builder.