Previous versions
V‑Cloud 1.9.4-2023.04.04
Updated V‑Cloud to version 7.4.2-1 of the ASR engine.
Bug fix
-
Updated V‑Cloud to version 7.4.2-1 of the ASR engine to address an issue with transcribing diarized u-law, A-law and non-16-bit PCM audio. For full details, review the V‑Blaze 7.4.2-1 release notes here.
V‑Cloud 1.9.4-2023.01.09 Release Notes
V‑Cloud 1.9.4-2023.01.09 updates to version 7.4.1-1 of the ASR engine and includes the following changes:
LID has been updated to version 2.1.0 and includes the following:
Performance and accuracy updates for English-Spanish identification.
Added language identification for English-French audio.
Define the probability of alternative language speech using the
lidpriorparameter. Learn more: Additional LID Options
Fixes to text processing modules:
Fixed an issue where setting
punctuatetofalseresulted in mixed letter cases due to certain substitutions. Disablingpunctuatenow results in lowercase transcripts unless substitutions contain uppercase letters.Fixed an issue where substitution rules containing a combination of non-ASCII characters and a backreference in the same word would cause utterances to drop from results.
V‑Cloud 1.9.3-2022.09.14 Release Notes
V‑Cloud 1.9.3-2022.09.14 updates to version 7.4 of the ASR engine and includes the following changes:
-
Added automatic audio resampling using the
resampletag. For more information on this parameter, refer to Adjusting for audio. -
Updated the
warningfield in the top level of JSON transcripts to include the highest priority warning if there are multiple warnings. Voci recommends logging thewarningfield for all ASR flows. -
Fixed a text processing issue where unrelated strings were pulled into earlier utterances when number merging was performed.
-
Improvements to the
noisesetting for thediarizeparameter.
V‑Cloud 1.9.0-2022.06.02 Release Notes
V‑Cloud 1.9.0-2022.06.02 includes the following changes:
-
Updates to the following language models for significantly increased accuracy and performance:
-
English (North America) language models:
-
French (Canada) language models:
-
Refer to Language model updates for more information on language model releases and updates.
V‑Cloud 1.8.1-2022.04.30 Release Notes
V‑Cloud 1.8.1-2022.04.30 includes the following changes:
Updates to the eng-us:callcenter model. The update includes minor revisions to the language pack, additional training data, and a small configuration adjustment.
V‑Cloud 1.8.0-2022.03.30 Release Notes
V‑Cloud 1.8.0-2022.03.30 includes the following changes:
-
Requests for deprecated eng1 models are now automatically remapped to use the improved eng-us models where available. Requests specifying any of the following eng1 models will default to the matching eng-us language model instead:
-
English (North America) Automotive Industry — eng-us:autodealership has deprecated eng1:autodealership.
-
English (North America) Call Center — eng-us:callcenter has deprecated eng1:callcenter.
-
English (North America) Financial — eng-us:financial has deprecated eng1:financial.
-
English (North America) Healthcare — eng-us:healthcare has deprecated eng1:healthcare.
-
English (North America) Insurance — eng-us:insurance has deprecated eng1:insurance.
-
English (North America) Large Vocabulary — eng-us:largevocab has deprecated eng1:largevocab.
-
-
V‑Cloud now includes the
uttmaxgapparameter. Refer to Utterance Controls for more information.
V‑Cloud 1.7.2-2022.03.02 Release Notes
V‑Cloud 1.7.2-2022.03.02 includes the following changes:
V‑Cloud has been updated to version 7.3.2-2 of the ASR engine.
V‑Cloud now includes the following language models:
English (North America) Language Models:
V‑Cloud now includes the
filetypeparameter, which can be used to manually specify theContent-Typeof your audio. Refer to Common tags for more information.
V‑Cloud 1.7.2-2022.01.25 Release Notes
-
V‑Cloud has been updated to version 7.3.2-1 of the ASR engine.
-
V‑Cloud includes updates to the following language models:
-
English (North America) language models:
-
eng-us:callcenter — Improved models for English (North America) Call Center domain including updates to the acoustic and language models for significantly increased accuracy and performance. The eng-us:callcenter model has deprecated eng1:callcenter, and is now the default when not otherwise specified.
-
eng1:callcenter — Updated language model for increased accuracy and a small decrease in performance. This update was performed to provide some improvement for those who are unable to use the eng-us:callcenter model. The model has deprecated the eng1:callcenter model and is recommended for all uses.
-
-
-
V‑Cloud now includes the following language models:
-
Portuguese (Brazil) language models:
-
V‑Cloud 1.6-2021.11.19 Release Notes
V‑Cloud 1.6-2021.11.19 includes the following updates and fixes:
V‑Cloud has been updated to version 7.3 of the ASR engine. Refer to V-Blaze Version 7.3.0-1 Release Notes for more information.
V‑Cloud now includes the following language models:
French (Canada) language models:
Spanish (Argentina) language models:
Spanish (Colombia) language models:
Spanish (Mexico) language models:
Spanish (Panama) language models:
Refer to Language models for more information on available language models.
V‑Cloud 1.6-2021.09.29 Release Notes
V‑Cloud 1.6-2021.09.29 includes the following updates and fixes:
-
Multiple improvements to the text processing modules to enhance presentation of words, numbers, and punctuation. This includes fixes for situations related to AM/PM, Q1-Q4, "O" as zero, words/events with spaces, and address-related ordinals.
-
Fixed an error with certain substitution patterns. Substitution patterns with left-hand-side unicode or slash-protected (from a previous rule) strings in multiple value sets/lists are now processed correctly.
-
Multiple language ID (LID) improvements and additions.
-
Changed
lidthresholdto now adjust the confidence level required for the system to select the alternative language. -
Optimized LID to automatically use different technologies based on language models, channels, per-stream vs per-utterance, and other characteristics.
Refer to lid for more information on these changes.
-
V‑Cloud 1.6-2021.05.20 Release Notes
V‑Cloud 1.6-2021.05.20 includes the following updates and fixes:
-
Made multiple improvements to number translation and web URL formation.
-
Changed number translation behavior to improve transcript readability. The translation is now more conservative by considering more context for various situations.
-
-
Made multiple improvements to redaction functionality:
-
Added the
scruboffsetoption when using redaction. Refer to Redaction for more information on this parameter. -
Fixed an issue from V-Cloud 1.6.2021.04.20, released on April 24, 2021, where number words concatenation was not working properly. This issue caused all single numbers (0 through 9) to be printed as a word instead of a numeral; for example, 1 was transcribed as "one." Numbers 10 and up were correctly processed. This issue affected audio and text scrubbing configurations that depend on single digits.
-
Fixed a rare truncated audio issue when using redaction.
-
-
Improved error handling:
-
Fixed an issue that could result in ASR errors not being reported correctly.
-
JSON output now only includes the
ended,model,nchannels, andaudiosecselements if the stream completed successfully. These elements do not display in JSON output if there was a problem processing audio.
-
V‑Cloud 1.6-2021.03.25 Release Notes
V‑Cloud 1.6-2021.03.25 includes the following updates and fixes:
V‑Cloud now returns 502 or 504 status codes when a URL audio source fails.
502 Bad Gateway returns when a service required for the request is inaccessible.
504 Gateway Timeout returns when a request is unable to complete within a reasonable amount of time.
Refer to Return codes for more information on V‑Cloud return codes.
Fixed an issue that caused URLs with a query component to break due to percent-encoding the URL for all URL audio sources before downloading.
V‑Cloud 1.6-2021.02.25 Release Notes
V‑Cloud 1.6-2021.02.25 includes the following updates and fixes:
-
V‑Cloud has been updated to version 7.2 of the ASR engine.
-
Language identification (LID) has been updated to a new engine that significantly increases LID accuracy.
-
V‑Cloud now supports basic HTTP authentication for URL audio sources. When a URL audio source requires authentication, prepend your access credentials to the hostname in the URL as shown in the following example:
curl -F token=your_token_here \ -F url=http://username:password@hostname.com/sample-audio-file.wav \ https://vcloud.vocitec.com/transcribe -
V‑Cloud now supports HTTP authorization request headers for URL audio sources. When a URL audio source requires authorization, include the authorization header, type, and credentials in your request as shown in the following example:
curl -F token=your_token_here \ -F url=http://hostname.com/sample-audio-file.wav \ -H 'Authorization: auth_type credentials' \ https://vcloud.vocitec.com/transcribe
V‑Cloud 1.6-2020.12.10 Release Notes
V‑Cloud 1.6-2020.12.10 includes the following updates:
V‑Cloud now includes
requestidas a top-level field in JSON transcripts.requestidis a unique identifier which is automatically generated in all transcripts for tracking purposes.V‑Cloud now includes the following language models:
English (United Kingdom) language models:
English (Europe) language models:
French (Canada) language models:
Numerous updates to the following language models:
English (Europe) language models:
French (Canada) language models:
V‑Cloud 1.6-2020.11.05 Release Notes
V‑Cloud 1.6-2020.11.05 includes the following updates and fixes:
Updated V‑Cloud to version 7.1 of the Automatic Speech Recognition (ASR) Engine.
Numerous updates to the following European and Mexican Spanish language models:
European Spanish language models:
Mexican Spanish language models:
Fixed an issue that prevented the use of the
Content-MD5HTTP header for data verification.Added user IP address logging to front
webserverlogs.Added
billingSecs,srcBkt, andresBktto frontwebserverreport.The default values for certain parameters can now be specified per individual token. The defaults of the following parameters are configurable:
modeloutputcallbackcallbackfmtcallbackurlkeepsrc
V‑Cloud 1.6-2020.10.07 (October 2020) Release Notes
V‑Cloud 1.6 includes several new features and parameters, including:
Language identification (LID) has been enhanced with new option parameters, JSON output elements, and other functionality improvements.
Added new optional parameters for
lidTable 1. New optional parameters for lidName
Values
Description
lidoffset= Ninteger
Delay start of LID until specified (N) seconds into audio. If there is not enough audio left after offset, this will process preceding utterances in reverse.
lidutttrue, false
Run LID on every utterance. The default is only once per stream or audio channel. This option is only available with V‑Blaze 7.1+.
Note: This option has a significant performance impact and should only be used when necessary.Added a new
lidparameter value for limiting utterance metadata.lid= language :notext- use this to not decode when a language is detected, just that a language was detected. If this option is specified, the JSON will not contain a model metadata element in each utterance. Ensure that the language is specified but the domain is not included. For example:lid=spa:notextis valid, butlid=eng1:callcenter:notextis not.Improved logic for LID decisions with low scores.
When LID scoring is below the decision threshold, the ASR engine will transcribe the audio with the language model specified by the model tag (or the default model for the ASR configuration if
modelis not explicitly provided). The results are indicated by alidinfo.langfinalelement in the JSON output.Made additions to JSON output.
lidinfo- now provided at utterance-level whenlidutt=truelanginfo- breakdown of language information that is added when there was more than one language detected.langfinal- added when the language specified in LID is below threshold and not the default language.
For more information on LID and using these parameters, refer to Receiving language identification information.
Added new debugging parameters and JSON elements to assist with improved warnings and logging when using substitutions.
New debugging parameters
Warning: These parameters are intended for debugging purposes only and should not be used in production.Table 2. New substitution debugging parameters Name
Values
Description
subst
true, false (default), none
The
substparameter can be used to enable or disable automatic system- and model-level substitutions.-
subst=true Enables system- and model-level substitutions
-
subst=false Disables system-level substitutions; model-level substitutions still apply
-
subst=none Disables both system- and model-level substitutions
substinfo
true, false (default)
Provides substitution details in JSON transcripts.
Set
substinfoto true to include a top-level JSON object that indicates the applied substitution rules and a number count for each rule.In addition to the top-level JSON object,
substinfoincludes another JSON object in the metadata that details each substitution's location, the substitution rule applied, and the substitution rule source.For more details on these and other parameters, refer to the Substitutions section of the V-Cloud API docs.
-
Added a new JSON output element:
nsubsshows a count of substitutions applied at both top-level and utterance levels. Whensubstinfo=true,nsubswill also includenumtranscounts within thesubstinfoarray. Top-levelnsubsdoes not includenumtranscounts. Thensubselement will not appear if no substitutions were applied.
For more information, refer to the V‑Cloud 1.6 Release Notes (July 2020).
Made quality improvements to eng2:largevocab and spa1:voicemail language models.
Hinting is now supported for eng1 version 7 models. Hinting for version 5 models is no longer supported.
For more information on hinting support, refer to the English page of the Language Models Reference.
Made minor enhancements and fixes to speech-to-text output processing (
textproc), including:The system now preserves timestamps on backrefs in substitutions instead of interpolating.
Fixed inadvertent uppercase of English cased backrefs (for example,
/\1/).Eliminated unexpected behavior of
pattern{min,max}whenmin=0.
Made minor improvements to Spanish time formatting.
Corrected the scope of emotion scoring to always score individual utterances.
Eliminated rare edge case decode failures.
V‑Cloud 1.6 Release Notes (July 2020)
V‑Cloud 1.6 includes several new features and parameters, including:
-
V‑Cloud now allows users to submit audio data through a URL. This method uses the
urlparameter as an alternative to thefileparameter. The provided URL must support HTTPGETand return aContent-Lengthheader when queried. V‑Cloud will verify that the data was properly received from the URL if either aContent-MD5orETagheader are provided in response to querying the URL. -
V‑Cloud now allows users to specify a callback format with the
callbackfmtparameter. There are three format options available:-
multiperforms a multi-partPOSTto the callback with two parts specifying the request ID and the resulting file content. -
singleperforms a standardPOSTthat only contains the result file content within thePOSTbody. -
putis identical tosingleexcept the callback is performed using an HTTPPUTinstead of aPOST.
-
-
V‑Cloud now allows users to send the results of erroneous jobs to a specific callback server with the
callbackerrorparameter. Ifcallbackerrorisn't specified, erroneous job results are sent to the URL defined in thecallbackparameter instead. -
V‑Cloud now allows users to enable or disable MD5 verification of audio data using the
filemd5parameter.filemd5has 3 options:-
trueis the default value and specifies that MD5 verification should be performed if possible. -
falsespecifies that MD5 verification should not be performed. -
User-specified 128-bit value - Specify a sequence of 32 hexadecimal digits to use for MD5 verification.
-
-
V‑Cloud now supports HTTP basic authentication for callbacks. When a callback server requires authentication information, prepend your access credentials to the hostname in the URL as shown in the following example:
https://username:password@hostname.com
