Supported languages
Language support for Medallia Speech
Speech supports transcription of voice data in several languages. Speech uses two separate transcription engines to maximize language support: Medallia's in-house transcription engine (Engine 1) and Amazon Transcribe (Engine 2).
Supported language matrix
Medallia Speech supports a variety of languages and regional dialects, with different transcription engines and features available for different languages.
Languages listed as Engine 1 use Medallia's in-house transcription engine, and those listed as Engine 2 use the Amazon Transcribe service.
We separate Engine 1 languages into quality tiers. For a description of each language quality tier in the following table, see the Language tiers section below.
Consult the table below to view languages and their supported features in Medallia Speech.
| Language | Engine | Feature support | Text Analytics |
|---|---|---|---|
|
Arabic (Gulf) / ar-AE | 2 |
| |
|
Chinese (Simplified) / zh-CN | 2 |
| |
|
Chinese (Traditional) / zh-TW | 2 |
| |
|
Czech / cs-CZ | 2 |
| |
|
Danish / da-DK | 2 |
| |
|
Dutch (Europe) / nl-NL | 1 (Tier 3) and 2 | For Engine 1 only:
| |
|
English (Australia) / en-AU |
1 (Tier 1) |
| |
|
English (Great Britain) / en-GB |
1 (Tier 1) |
| |
|
English (United States) / en-US |
1 (Tier 1) |
| |
|
French (Canada) / fr-CA |
1 (Tier 2) |
| |
|
French (France) / fr-FR |
1 (Tier 2) |
| |
|
German (Germany) / de-DE |
1 (Tier 2) |
| |
|
Greek / el-GR | 2 |
| |
|
Hebrew / he-IL | 2 |
| |
|
Hungarian / hu-HU | 2 |
| |
|
Italian (Italy) / it-IT |
1 (Tier 2) |
| |
|
Japanese / ja-JP | 2 |
| |
|
Korean / ko-KR | 2 |
| |
|
Norwegian / no-NO | 2 |
| |
|
Polish / pl-PL | 2 |
| |
|
Portuguese (Brazil) / pt-BR |
1 (Tier 1) |
| |
|
Portuguese / pt-PT | 2 |
| |
|
Romanian / ro-RO | 2 |
| |
|
Slovak / sk-SK | 2 |
| |
|
Spanish (Spain) / spa-ES |
1 (Tier 1) |
| |
|
Spanish (Americas) / spa-AMER |
1 (Tier 1) |
| |
|
Swedish / sv-SE | 2 |
| |
|
Turkish / tr-TR | 2 |
|
* Custom vocabulary (OOV) for Engine 2 has limited availability. It is configured differently than Engine 1 and must be enabled by a Medallia expert before use.
Available features
The following list defines the features available for the designated languages in the table below:
Acoustic emotion — Discover changes in customer's emotional expressions throughout an interaction.
Custom vocabulary (OOV) — Transcription accuracy can be improved by recognizing brand- and industry-specific terms.
Music detection — Detect when utterances should be classified as background music and not transcribed.
Numeric redaction — Numeric phrases are redacted from audio and transcripts.
Substitution — Correct errors in transcripts automatically using rules that find and replace errors with corrected text.
Voice anonymization — Audio can be modified to mask speakers' identities.
Language tiers
We separate Engine 1 languages into tiers to differentiate the datasets used for training language models. Higher tiers feature more robust models that consistently produce higher quality results.
Tier 1
Tier 1 languages are available to use with Speech. Tier 1 language models have been trained with a robust dataset and produce high quality transcription results. Audio and transcription analysis is recommended for output tuning prior to using Tier 1 languages in production environments.
Tier 2
Tier 2 languages are available to use with Speech. Tier 2 language models have been trained with a limited dataset, meaning transcription quality may not meet production requirements. Audio and transcription analysis is strongly recommended to evaluate transcription output prior to using Tier 2 languages in production environments.
Tier 3
Tier 3 language models are not generally available and are considered beta models. Tier 3 languages are not available for Speech without approval from Medallia Product and Engineering teams, and this approval process must include thorough audio and transcription analysis.
Text Analytics support
Ensure Text Analytics is turned on for the languages used for voice data by your company. On the Reporting > Text Processing > Global Text Analytics Settings screen in Medallia Setup, review the language processing settings for the languages used by your company, and make any changes as needed. For more information, see Text Analytics supported languages.
Amazon Transcribe languages
Speech uses Amazon Transcribe to support the languages listed as Engine 2 in the table above.
Some Speech features, including acoustic emotion, custom vocabulary, and redaction are not available when using Amazon Transcribe languages. Substitution is not enabled for Amazon Transcribe languages (Engine 2) by default. If you wish to enable substitution for these languages, contact a Medallia expert.
Amazon Transcribe's Language Identification feature can be used to identify languages supported with Amazon Transcribe and Medallia Text Analytics. One language is transcribed for any given call. Language identification is not available for Medallia's in-house transcription engine. For more information, see Connector settings.
Amazon Transcribe language models are developed and trained by Amazon. For more information, see Amazon Transcribe.
