What is WER?

Word Error Rate — the standard accuracy metric for speech recognition. Lower is better.

WER is calculated by comparing the system's output against the reference transcript word by word, counting substitutions, deletions and insertions, then dividing that total by the word count.

One caution when buying: the WER a supplier quotes is usually measured on their own test set. To judge the data, ask under what conditions it was measured.

Related terms

  • ASR

    Automatic Speech Recognition — the task of turning recorded speech into text.

  • Transcription

    Writing down what is said in a recording. It is the core annotation task in any speech dataset.

Keep reading

Submit a sourcing request

Tell us the language, the hours, and what the data needs to look like. You will get a real number and a real timeline — not a range. If we cannot source it well, we will tell you that instead.

  • Pilot batch before the full run, so problems surface early.
  • Consent documentation delivered with the data.
  • No medical or clinical data. No recorded telephone calls.

We reply within two business days. Your details are used only to answer this request. See our privacy policy.

Contact

Talk to a human

Send a specification and we will come back with a real number and timeline.

Submit a sourcing request

Or email hello@linguacorpus.com