Catalogue of Tools & Metrics for Trustworthy AI

These tools and metrics are designed to help AI actors develop and use trustworthy AI systems and applications that respect human rights and are fair, transparent, explainable, robust, secure and safe.

ASR-FAIRBENCH



ASR-FairBench is an online leaderboard that assesses both the accuracy and the fairness of automatic speech recognition (ASR) systems, which convert speech into text. ASR systems often perform less well for some speakers than others, depending on factors such as accent, gender, age or first language. Existing ASR leaderboards rank models only on overall accuracy, measured by word error rate (WER). A model can therefore rank highly while performing poorly for underrepresented groups. ASR-FairBench addresses this gap by evaluating fairness alongside accuracy. It was developed by researchers at IIT Kharagpur and presented at Interspeech 2025.

The leaderboard uses a stratified 10% sample of Meta's Fair-Speech dataset, which contains recordings of voice assistant commands from 593 US participants. Participants reported their own age, gender, ethnicity, socioeconomic background and first language. The sample preserves the demographic balance of the full dataset while greatly reducing evaluation time.

A statistical model estimates how error rates differ across demographic groups. Each demographic category receives a fairness score from 0 to 100, which is reduced when differences between groups are statistically significant. These scores are combined into an overall fairness score. This is then merged with WER into a single Fairness-Adjusted ASR Score (FAAS), and models are rated on a five-level scale from "severely biased" to "exemplarily fair". Users can submit their own ASR models through a web interface and receive a fairness audit within minutes.

Auto-discovered on 2026-09-23 by OECD Catalogue Automation

Use Cases

There is no use cases for this tool yet.

Would you like to submit a use case for this tool?

If you have used this tool, we would love to know more about your experience.

Add use case
Partnership on AI

Disclaimer: The tools and metrics featured herein are solely those of the originating authors and are not vetted or endorsed by the OECD or its member countries. The Organisation cannot be held responsible for possible issues resulting from the posting of links to third parties' tools and metrics on this catalogue. More on the methodology can be found at https://oecd.ai/catalogue/faq.