> ML_BENCHMARK // LIBRISPEECH-WORD-ERROR-RATE_v1.0
LibriSpeech ASR Word Error Rate (WER)
Speech Recognition · task-speech-recognition · near-saturation
Speech Recognitionnear-saturationmoderate
Evaluation Protocol
Standard greedy or beam search decoding evaluated on test-clean and noisy test-other splits without language model rescoring.
Baseline & Metrics
Canonical Baseline:Kaldi (hybrid): 3.8% clean | wav2vec 2.0: 1.8% clean / 3.9% other | Whisper Large-v3: 1.5% clean / 2.7% other
Evaluated Metrics:
Word Error Rate (WER % on test-clean and test-other)
Contamination & Leakage Risks
Clean/other splits are strictly speaker-independent.
Reproducibility Concerns
Text normalization rules (capitalization, punctuation, number-to-words) alter WER by up to 2%.
CONNECTED DATASETLibriSpeech ASR Corpus (Panayotov et al. 2015)Speech Recognition & Acoustic Modeling · 1,000 hours of 16kHz read English audiobooks with aligned text transcripts
View Dataset Spec