Skip to main content
Transcription accuracy varies by language. This page tells you what to expect from each one and how much review to budget for. For the full list of languages CoLoop accepts, see Supported languages.

How the pipeline works

CoLoop processes multilingual research in three steps: transcription, then translation, then analysis. Analysis runs on the original-language transcript, even after you translate a file to English and even when the output you read is in English. Themes and insights come from the source transcript rather than from a translation of it, so nothing drifts in the translation step. Direct participant quotes stay in the language they were spoken in.
Audio quality affects accuracy in every language. Record in a quiet room with a good microphone wherever you can.

What to expect by language

Ratings describe the transcript CoLoop produces. These are the languages most common in international research programs. Every language CoLoop transcribes can also be translated to English.

How to plan your project

  • 🟢 High: run these end to end without special handling.
  • 🟡 Good: have a native speaker check a sample of transcripts before you run full analysis.
  • 🟠 Moderate and 🔴 Lower: correct transcripts before you analyze them, and consider human transcription where the audio is poor or the interview is unstructured.
Enter your key words and phrases at the transcription stage, in any language. Brand names, product names, and domain vocabulary are then corrected rather than guessed at.

Medical and clinical research

Medical Mode adds a correction pass over the terminology general-purpose models most often get wrong: medication names, procedures, conditions, and dosages. It covers English, Spanish, German, and French, and you can combine it with your own key phrases for terminology specific to your study.
Medical Mode is off by default and enabled per account. For clinical research in other languages, standard transcription applies, so enter your key phrases.

Reading the accuracy bands

The bands above are based on Word Error Rate (WER), the standard measure for transcription accuracy. WER counts the corrections, insertions, and deletions needed to turn an automated transcript into a perfect one, as a percentage of total words. A high WER does not mean every other word is wrong. Errors cluster around proper nouns, technical terms, strong accents, and fast speech, while the surrounding context stays intact. WER tells you how much researcher review to budget for.

How files are routed

CoLoop picks the transcription and translation service that handles each language best. Routing is automatic. Regional dialects such as Quebecois French, Brazilian Portuguese, and Mexican and Argentine Spanish are recognized without a separate setting. For the list of providers that handle your data, see the CoLoop subprocessor list.
Cantonese is a distinct spoken language from Mandarin. Select Cantonese, not Chinese, when you set up the file, and review a sample before you analyze it.

What CoLoop does to limit errors