Introducing Monsoon ASR ⚡️
@voicearena_ai
Speech recognition does not have a model problem anymore. It has a data problem.
The best ASR systems are approaching human-level performance in English. But move into the long tail of the world's languages, especially real, conversational speech and error rates can still be 5-10X higher.
Today, we're releasing Monsoon ASR: a new generation of training data built specifically to close that gap.
50 languages. 100,000+ hours. Dense spontaneous speech.
And one goal: Single-digit WER across the world's languages.🧵