โ† All examples

ModelEthio-ASR Amharic (600M)0Ethio-ASR Multi (300M)2Ethio-ASR Multi (600M)1Ethio-ASR Multi (94M)3Meta MMS-1B8OmniASR-CTC (1B)1OmniASR-CTC (300M)3OmniASR-CTC (3B)0OmniASR-LLM (1B)0OmniASR-LLM (300M)2OmniASR-LLM (3B)1OpenAI Whisper-small8
Audio qualityClean Studio8Ambient Room Noise8Harsh Environment8Distant / Low-Volume Mic8Wideband Phone (VoLTE)8Randomized Phone Channel8Narrowband 2G + Packet Loss8

What did the model get wrong?

Meta MMS-1B Clean Studio   Example 15 of 20  ยท  speaker: Female

๐ŸŽง Listen โ€” Clean Studio
Clear, studio-quality recording. Best-case audio.
โœ… What was actually said
แŠฅแˆฝ แŠ แˆแŠ• แŠฅแŠ•แ‹ฐแ‰ต แАแ‹ แ‹จแˆแŒ แ‰€แˆ˜แ‹ แŠ แ‹ตแˆต แ‰แŒฅแˆญ แŠ แˆแ‹ฐแˆจแˆฐแŠแˆ
๐Ÿค– What this model heard
แˆญแ‹ แ‰ตแˆ แ‰ฝ แ‰  แ‰ตแ‰ฝแˆญ แ‰ณแˆ˜ แ‰ต
Correct
Wrong (different word)
Missed (skipped)
Extra (added)
0
Correct
7
Wrong
1
Missed
0
Extra
0%
Words right
Out of 8 spoken words, this model got 0 right (0%). Errors: 7 words wrong; 1 missed.

Try switching Audio quality โ€” see the same model degrade as the audio gets noisier.