โ† All examples

ModelEthio-ASR Amharic (600M)8Ethio-ASR Multi (300M)16Ethio-ASR Multi (600M)11Ethio-ASR Multi (94M)10Meta MMS-1B12OmniASR-CTC (1B)7OmniASR-CTC (300M)13OmniASR-CTC (3B)9OmniASR-LLM (1B)5OmniASR-LLM (300M)11OmniASR-LLM (3B)10OpenAI Whisper-small12
Audio qualityClean Studio3Ambient Room Noise0Harsh Environment0Distant / Low-Volume Mic0Wideband Phone (VoLTE)0Randomized Phone Channel0Narrowband 2G + Packet Loss9

What did the model get wrong?

OmniASR-CTC (3B) Narrowband 2G + Packet Loss   Example 20 of 20  ยท  speaker: Male

๐ŸŽง Listen โ€” Narrowband 2G + Packet Loss
Heavy background noise plus phone-call compression โ€” worst-case stress test.
โœ… What was actually said
แ‰แŒฅแˆฉ แ‹œแˆฎ แ‹˜แŒ แŠ แˆแˆˆแ‰ต แ‹œแˆฎ แŠ แˆแˆตแ‰ต แˆตแ‹ตแˆตแ‰ต แˆฐแ‰ฃแ‰ต แˆตแˆแŠ•แ‰ต แ‹˜แŒ แŠ แ‹œแˆฎ แАแ‹
๐Ÿค– What this model heard
tr แ‹œแˆฎ ot t แ‹œแˆฎ mmn ot แ‹œแˆฎ n
Correct
Wrong (different word)
Missed (skipped)
Extra (added)
3
Correct
6
Wrong
3
Missed
0
Extra
25%
Words right
Out of 12 spoken words, this model got 3 right (25%). Errors: 6 words wrong; 3 missed.

Try switching Audio quality โ€” see the same model degrade as the audio gets noisier.