This site contains only USER GUIDES. Looking for ADMIN content? Login here at help-auth.eleveo.com
Eleveo WFO

Speech Recognition Integration

Supported for CLOUD Deployments + on-premise DEPLOYMENTS + Hybrid Deployments

Overview

Eleveo offers a Speech Recognition package that is installed on a separate, dedicated, server. The solution is provided for both on-premise and cloud deployments. Feature availability may vary based on your installation.

General Limitations

Multiple language packs can be configured (based on what is supported by the given speech engine) but only a single speech engine can be configured.

  • The Eleveo solution does not support multiple engines in parallel.

Speech Recognition

Speech Recognition works with a limited number of languages. Speech Recognition is installed as an add-on to Quality Management and must be configured. This feature provides transcription services for all supported languages. The audio files generated in the contact center or back-office are sent via a dedicated API to a secondary system that processes the recording, analyzes the audio, detects emotion/sentiment, transcribes the audio, and tags the relevant section. View the transcription within the Conversation Explorer.

speech rec external speech eng.png
Cloud Installation: Graphical overview of how the Speech Recognition service is interconnected with other services

What is Supported - Based on Speech Recognition Service

Languages

Speech Recognition - Voci

Deprecated-EOL Jan 2027

Speech Recognition - Speechmatics

Supported Languages

Dialects supported

Supported Languages





English

  • North America

  • Australia

  • United Kingdom

  • Europe

  • Philippines

  • International

Automatic
Arabic
Bashkir
Basque
Belarusian
Bengali
Bulgarian
Cantonese
Catalan
Croatian
Czech
Danish
Dutch
English
Esperanto
Estonian
Finnish
French
Galician
German
Greek
Hebrew
Hindi
Hungarian
Indonesian
Interlingua
Irish
Italian
Japanese
Korean

Latvian
Lithuanian
Malay
Malay & English bilingual
Maltese
Mandarin
Mandarin & English bilingual
Marathi
Mongolian
Norwegian
Persian
Polish
Portuguese
Romanian
Russian
Slovakian
Slovenian
Spanish
Spanish & English bilingual
Swahili
Swedish
Tamil
Tamil & English bilingual
Thai
Turkish
Ukrainian
Urdu
Uyghur
Vietnamese
Welsh

French

  • Canada

  • France

  • Europe



Spanish

  • North America

  • Spain

  • Mexico

  • Argentina

  • Colombia

  • Panama

German


Italian


Portuguese

  • Brazil

Dutch


For up-to-date information regarding supported language packs. See the provider's documentation.

For up-to-date information regarding supported language packs. See the provider's documentation.

Medallia Documentation

Speechmatics Documentation

Additional Features - Installation Dependent

Feature

Speech Recognition - Voci

Speech Recognition - Speechmatics

Transcription

check mark

check mark

Phrase Spotting (on top of transcription)

check mark

check mark

Emotion/sentiment detection

check mark (English only)

Available on transcription utterance as well as participant level

check mark (English only)

Acoustic parameters 

Refer the list provided below.

check mark


check mark

Transcription redaction

check mark

check mark

Automated language identification

check mark
If automated language detection is enabled for your server, the system will automatically detect what language is used in the first twenty seconds of the recording and switch the language processor to the detected language. This means that if multiple languages are used in a conversation, the system will transcribe text according to the language detected at the beginning of the recording. If the speech recognition engine fails to detect the language accurately, it may produce transcriptions for the incorrect language. The system does not automatically detect and switch languages after the first twenty seconds, even if speakers switch between different languages.

Automatic recognition for the following language pairs is supported:

  • English / Spanish

  • English / French

check mark

If automated language detection is enabled for your server, the system will automatically detect what language is used. It is possible to configure the expected languages.

Transcription tuning

check mark Custom vocabulary is supported.

check mark Custom vocabulary is supported.

Supported formats

WAV only

WAV, MP3, MP4 (pvideo format only)

Availability

Cloud, Hybrid, On Prem

Cloud, Hybrid, On Prem

Reprocessing of Archived Media

minus

minus

Acoustic Parameters by Provider

The following list is provided as additional information. Data available may vary based on the quality of the recorded conversation.

  • Acoustic Parameters for participants are not available for mono recordings (such as MS Teams) and will not be available for reporting.

  • Acoustic metadata will be available in Eleveo Reporting dashboards and Insight Hub if all conversation segments are processed properly.

Acoustic Parameter

Speech Recognition - Voci

Speech Recognition - Speechmatics

General statistics – Aggregated for the entire conversation


Interruptions count – Number of interruptions

check mark

check mark

Total crosstalk duration (sec.) – Total time that the speakers were interrupting or speaking over each other

check mark

check mark

Total crosstalk ratio (%) – Ratio of time that the speakers were interrupting or speaking over each other

check mark

check mark

Silence count – Silence count includes all silences that are greater in length than 800 milliseconds. This means that the silence count may be 0. In contrast, Total silence duration might be greater than 0 as it combines all silence time, even short periods of silence.

check mark

check mark

Total silence duration (sec.) – How much time was silent (no audio)

check mark

check mark

Silence ratio (%) – Ratio of time that was silent relative to talk time

check mark

check mark

Talking count – Total count of utterances (i.e. phrases, sentences in the transcription)

check mark

check mark

Total talking duration (sec.) – Total time a participant was speaking

check mark

check mark

Talking ratio (%) – How much time (as a ratio) a participant was speaking

check mark

check mark

Speaker specific statistics


Gender (Male/Female) – If detected the system displays the gender of the speaker (this information is not displayed unless configured by an administrator)

check mark

minus

Total talking duration (sec.) – Total time the participant was speaking

check mark

check mark

Talking ratio (%) – How much time (as a ratio) the participant was speaking

check mark

check mark

Average speed (words/min.) – How fast the speaker was speaking. Average number of words per minute (rounded to 2 decimal places)

check mark

check mark

Interruptions count – Number of interruptions (times the speakers spoke over each other)

check mark

check mark

Total crosstalk duration (sec.) – Total time that the speaker was interrupting or speaking over the other

check mark

check mark

Total crosstalk ratio (%) – Ratio of time that the speaker was interrupting or speaking over the other

check mark

check mark

Average talk speed –  Average number of words spoken per minute

check mark

check mark

Agent talking ratio –  Ratio of the call, in percent, where the agent is speaking

check mark

check mark

Agent crosstalk ratio – Ratio of the call, in percent, where there is crosstalk

check mark

check mark

Agent number of interruptions – Number of times crosstalk is detected

check mark

check mark