โ† All examples

ModelEthio-ASR Amharic (600M)4Ethio-ASR Multi (300M)8Ethio-ASR Multi (600M)5Ethio-ASR Multi (94M)7Meta MMS-1B12OmniASR-CTC (1B)7OmniASR-CTC (300M)8OmniASR-CTC (3B)3OmniASR-LLM (1B)7OmniASR-LLM (300M)0OmniASR-LLM (3B)0OpenAI Whisper-small12
Audio qualityClean Studio12Ambient Room Noise12Harsh Environment12Distant / Low-Volume Mic12Wideband Phone (VoLTE)12Randomized Phone Channel12Narrowband 2G + Packet Loss12

What did the model get wrong?

Meta MMS-1B Clean Studio   Example 9 of 20  ยท  speaker: Female

๐ŸŽง Listen โ€” Clean Studio
Clear, studio-quality recording. Best-case audio.
โœ… What was actually said
แ‰แŒฅแˆฉ แ‹œแˆฎ แ‹˜แŒ แŠ แŠ แŠ•แ‹ต แˆแˆˆแ‰ต แŠ แˆแˆตแ‰ต แˆตแ‹ตแˆตแ‰ต แˆฐแ‰ฃแ‰ต แˆตแˆแŠ•แ‰ต แ‹˜แŒ แŠ แ‹œแˆฎ แАแ‹
๐Ÿค– What this model heard
แ‰ณแ‰ แˆ˜ แ‹ˆ
Correct
Wrong (different word)
Missed (skipped)
Extra (added)
0
Correct
2
Wrong
10
Missed
0
Extra
0%
Words right
Out of 12 spoken words, this model got 0 right (0%). Errors: 2 words wrong; 10 missed.

Try switching Audio quality โ€” see the same model degrade as the audio gets noisier.