Word Error Rate

Turkish equivalent: Kelime hata oranıDomain: Speech Processing

An ASR accuracy metric defined as substitutions plus deletions plus insertions divided by the number of reference words.

Speech Processing Context

Word Error Rate is computed as (S + D + I) / N, where substitutions, deletions, and insertions are measured against the number of reference words. Text normalization, punctuation, number formatting, casing, and tokenization must be fixed before benchmark results are compared.

Evaluation Boundary

A lower WER does not imply a better system for every application. Speaker attribution, timestamp quality, latency, domain-critical words, and error severity may need separate metrics.

Related technical article: Automatic Speech Recognition.

Related technical publications

Publications whose title or summary directly references this concept.

IBM POWER9 AC922 for Digital Forensics and Artificial Intelligence

My IBM POWER9 AC922 has been a long-running engineering platform since 2019 for digital forensics and AI work involving AltiVec/VSX, OpenMP, CUDA, Tesla V100, dlib, MXNet, FAISS, ArcFace, Kaldi, Vosk, whisper.cpp, OCR, file carving, and current CTranslate2/faster-whisper experiments.

Why Delphi Became an Old Technology

A personal engineering perspective from someone who has worked with Delphi, examining how it moved from highly productive desktop RAD to a narrower position in modern open-source, web, mobile, Linux and server-side software ecosystems.