@MSFTResearch: Today, we announce the second release of PazaBench, our benchmark for evaluating Automatic Speech Recognition (ASR) mod…

X AI KOLs Timeline Tools

Summary

Microsoft Research announces the second release of PazaBench, a benchmark for evaluating automatic speech recognition models across African languages.

Today, we announce the second release of PazaBench, our benchmark for evaluating Automatic Speech Recognition (ASR) models across African languages. https://t.co/RLb5SnRQCu
Original Article
View Cached Full Text

Cached at: 07/21/26, 08:48 PM

Today, we announce the second release of PazaBench, our benchmark for evaluating Automatic Speech Recognition (ASR) models across African languages. https://t.co/RLb5SnRQCu

Similar Articles

Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German

arXiv cs.CL

This paper presents a benchmark evaluating five commercial ASR systems on code-switching speech across Arabic-English, Persian-English, and German-English pairs, using a two-stage pipeline to select 300 samples per pair and assessing performance with WER and BERTScore. ElevenLabs Scribe v2 achieves the lowest overall WER (13.2%) and highest BERTScore (0.936), with public dataset available.

BlasBench: An Open Benchmark for Irish Speech Recognition

arXiv cs.CL

BlasBench introduces an open evaluation benchmark for Irish speech recognition with Irish-aware text normalization that preserves linguistic features like fadas, lenition, and eclipsis. The paper benchmarks 12 ASR systems across four architecture families, revealing significant generalization gaps and showing that existing multilingual systems struggle with Irish due to inadequate normalization.

Introducing the FFASR Leaderboard: Benchmarking ASR in the Real World

Hugging Face Blog

Introduces the FFASR Leaderboard, an open, community-driven benchmark for evaluating automatic speech recognition models under realistic far-field acoustic conditions, highlighting the significant performance gap between near-field and far-field scenarios.