โ† All examples

ModelEthio-ASR Amharic (600M)1Ethio-ASR Multi (300M)4Ethio-ASR Multi (600M)2Ethio-ASR Multi (94M)3Meta MMS-1B9OmniASR-CTC (1B)4OmniASR-CTC (300M)2OmniASR-CTC (3B)2OmniASR-LLM (1B)0OmniASR-LLM (300M)1OmniASR-LLM (3B)0OpenAI Whisper-small13
Audio qualityClean Studio13Ambient Room Noise18Harsh Environment9Distant / Low-Volume Mic13Wideband Phone (VoLTE)9Randomized Phone Channel9Narrowband 2G + Packet Loss9

What did the model get wrong?

OpenAI Whisper-small Clean Studio   Example 7 of 20  ยท  speaker: Female

๐ŸŽง Listen โ€” Clean Studio
Clear, studio-quality recording. Best-case audio.
โœ… What was actually said
แ‹จแŠขแŠ•แ‰ฐแˆญแŠ”แ‰ต แŒฅแ‰…แˆ แ‹ˆแ‹ญแˆ แ“แŠฌแŒ… แˆ˜แŒแ‹›แ‰ต แŠฅแ‰ฝแˆ‹แˆˆแˆ แ‹‹แŒ‹แ‹ แˆตแŠ•แ‰ต แАแ‹
๐Ÿค– What this model heard
yen terny i me ttehane itdikal yem package mpakich mxatich mxatich mxatich lalu
Correct
Wrong (different word)
Missed (skipped)
Extra (added)
0
Correct
9
Wrong
0
Missed
4
Extra
0%
Words right
Out of 9 spoken words, this model got 0 right (0%). Errors: 9 words wrong; 4 extra.

Try switching Audio quality โ€” see the same model degrade as the audio gets noisier.