Yandex SpeechKit release notes: Speech recognition
- Release as of 18/03/2026
- Release as of 01/10/2025
- Release as of 19/09/2025
- Release as of 31/07/2025
- Release as of 24/04/2025
- Release as of 11/04/2025
- Release as of 05/03/2025
- Release as of 17/12/24
- Release as of 10/12/24
- Release as of 03/12/2024
- Release as of 31/10/24
- Release as of 09/08/24
- Release as of 26/06/24
- Release as of 03/06/24
- Release as of 23/04/24
- Release as of 09/04/24
- Release as of 27/03/24
- Release as of 28/02/24
- Release as of 27/02/24
- Release as of 12/01/24
- Release as of 12/01/24
- Release as of 29/12/23
- Release as of 22/11/23
- Release as of 10/11/23
- Release as of 06/09/23
- Release as of 15/08/23
- Release as of 20/07/23
- Release as of 07/07/23
- Release as of 13/06/23
- Release as of 07/06/23
- Release as of 25/05/23
- Release as of 17/05/23
- Release as of 14/04/23
- Release as of 16/03/23
- Release as of 07/03/23
- Release as of 08/02/23
- Release as of 20/12/22
- Release as of 20/10/22
- Release as of 05/10/22
- Release as of 20/09/22
- Release as of 29/06/22
- Release as of 07/06/22
- Release as of 25/04/22
- Release as of 19/04/22
- Release as of 14/03/22
- Release as of 02/03/22
- Release as of 17/02/22
- Release as of 03/02/22
- Release as of 26/01/22
SpeechKit provides updates by model and version.
For more information about the speech recognition methods, see About the technology.
Release as of 18/03/2026
Updates to general:rc:
- Improved Russian recognition quality in certain client scenarios.
- Enhanced classification of answerphones, genders, and negativity.
Release as of 01/10/2025
general:rc updates are now available in the general model.
Release as of 19/09/2025
Updates to general:rc:
- Improved the recognition quality for Russian and Uzbek in certain client scenarios.
- Enhanced the answerphone classification.
- Fixed the issues with word duplication and concatenation in recognition.
Release as of 31/07/2025
Added the option of querying generative text models for recognition. Learn more about this feature in Using LLMs to process recognition results.
Release as of 24/04/2025
Updates to general:rc:
- Improved the recognition quality for Uzbek.
- Improved the recognition quality for the medical vocabulary in Russian.
Release as of 11/04/2025
general:rc updates are now available in the general model.
Release as of 05/03/2025
Improved the recognition quality for Uzbek and Russian in general:rc.
Release as of 17/12/24
Improved the recognition quality for Uzbek and Kazakh in general:rc.
Release as of 10/12/24
Updates to general:rc effective as of December 3 are now available in general.
Release as of 03/12/2024
In general:rc, fixed and improved error messages when using unsupported recognition languages and audio formats.
Release as of 31/10/24
Improved the recognition quality for Uzbek and Turkish in general:rc.
Release as of 09/08/24
Updates to general:rc:
- Improved the recognition quality for Uzbek and Kazakh.
- You can now limit the recognition languages by specifying multiple values in the
language_restrictionfield.
Release as of 26/06/24
The general:rc updates of June 3 are now available in general.
Improved the recognition quality for Uzbek in general:rc.
Release as of 03/06/24
As requested by users, general:rc recognition quality was improved for abbreviations and medical terms in Russian.
Release as of 23/04/24
The general:rc updates of April 9 are now available in general.
Release as of 09/04/24
Changed the format of classifiers in general:rc. The formal_greeting, informal_greeting, formal_farewell, informal_farewell, insult, and profanity classifiers now return results as a probability of positives. The answerphone and negative classifiers now return only the probability of positives instead of the probability of belonging to two classes.
Release as of 27/03/24
All general:rc updates of February 28 are now available in general.
Updates to general:rc:
- Improved the recognition quality for Uzbek.
- Improved the speaker labeling in recognition results.
Release as of 28/02/24
Updates to general:rc:
- Improved the recognition quality for Uzbek.
- As requested by users, improved the recognition quality for medications, car models, and tobacco products in Russian.
Release as of 27/02/24
All general:rc updates are now available in general.
Release as of 12/01/24
Added support for the speaker labeling in recognition results in general:rc.
Release as of 12/01/24
Improved the recognition quality for Uzbek in general:rc.
Release as of 29/12/23
Updates to general:rc:
-
Fixed normalization errors for certain number representations (e.g., fifteen hundred ⟶ 1500).
-
Added support for the following classifiers:
gender. Returns probability values for themaleandfemaleclasses.negative. Returns probability values for thenegativeandnot_negativeclasses.answerphone. Returns probability values for theanswerphoneandnot_answerphoneclasses.
-
Added classifier positives for partial recognition results (
ON_PARTIALevent).
Release as of 22/11/23
All general:rc updates are now available in general.
Release as of 10/11/23
Updates to general:rc:
- Updated the speech recognition model for Russian.
- As requested by users, improved the recognition quality for the names of cities in Kazakhstan.
- Enhanced the quality of speech recognition result normalization for Kazakh.
- Fixed the internal server errors when working with small audio fragments.
Release as of 06/09/23
Updates to general:rc:
- Fixed the issue with English words appearing when using the recognition model in Russian.
- Improved the general recognition quality for Russian.
- As requested by users, improved the recognition quality for Russian.
- Improved the general recognition quality for Uzbek.
Audio classifiers added to general:rc in the August 15, 2023 release are now available in general.
Release as of 15/08/23
Added support for the audio classifiers in general:rc.
Release as of 20/07/23
Fixed the resampling and added the new conversation metrics in general.
Release as of 07/07/23
Updates to general:rc:
- Fixed the two-channel audio resampling bug in the API v3.
- Added the feature of calculating conversation metrics for speech analytics. It is set up using the
speech_analysisoption in theStreamingOptionsmessage.
Release as of 13/06/23
Fixed switching to English during Russian speech recognition in general:rc.
Release as of 07/06/23
Updates to general:rc:
- Improved the recognition quality for Uzbek, German, French, Dutch, Italian, Polish, and Hebrew.
- Added the number normalization for Uzbek.
- Added support for splitting text into phrases using
eou_updatein FullData mode.
Release as of 25/05/23
The May 17 release updates are now available in general.
Release as of 17/05/23
Updates to general:rc:
- Improved the general recognition quality for Russian.
- As requested by users, improved the recognition quality for Russian.
- Improved the recognition quality for Uzbek, German, French, Dutch, Italian, and Polish.
- Added support for the new recognition language: Hebrew (
he-HE).
Release as of 14/04/23
Improved the recognition quality for abbreviations in Russian based on client scenarios for the general:rc model.
Release as of 16/03/23
The March 7 release updates are now available in general.
Release as of 07/03/23
For general:rc:
- Improved the recognition quality for Uzbek.
- Added support for the number normalization when recognizing speech in English, German, French, Italian, Spanish, and Turkish. The number normalization is also available for Kazakh in test mode.
Release as of 08/02/23
- The first version of speech recognition for Uzbek is now available in the
general:rcmodel for all API versions. Under certain acoustic conditions, Uzbek may be recognized as Kazakh. The issue will be fixed in future model releases. - To query the
general:rcmodel in the API v3, you can now specify this value in themodelparameter.
Release as of 20/12/22
For general:rc:
- As requested by users, improved the recognition quality for medications, first and last names, and patronymics.
- Slightly improved the recognition quality for Kazakh and Turkish.
Release as of 20/10/22
For general:rc:
- Added recognition for Brazilian Portuguese; the language code is
pt-BR. - Improved the speech recognition quality for all languages in auto recognition mode.
- Slightly improved the recognition quality for Russian and Kazakh.
Release as of 05/10/22
The September 20 release updates are now available in general.
Release as of 20/09/22
For general:rc:
- Improved the recognition quality for Moscow districts and medications in Russian.
- Added the language classification in auto recognition mode.
The fixes are available for testing.
Release as of 29/06/22
- The multi-language model is now available in
general. - In
general:rcandgeneral, the multi-language model can accept hints on speech languages. - The June 7 release updates to
general:rcare now available ingeneralfor Russian.
Release as of 07/06/22
- Improved the punctuation and recognition of last names in
general:rc. - The April 25 release updates are now available in
general.
Release as of 25/04/22
Updates to general:rc:
- Improved the recognition of such words as gasification and regasification. in Russian.
- Added the service feedback when processing OGG-OPUS format. If a stream is not a valid audio in OPUS format, the service returns
Invalid_Argument.
Release as of 19/04/22
- Added the Turkish language to the multi-language speech recognition model.
- The new API version is available for the Yandex SpeechKit streaming recognition. The old interface will still be supported; however, all new features will only be available in the API v3.
Release as of 14/03/22
The March 2, 2022 general:rc version is now available by the general tag.
Release as of 02/03/22
The general model now offers improved recognition of names, addresses, and terms as well as punctuation placement in long sentences and texts with numbers.
Further updated the general:rc model based on the user data.
Release as of 17/02/22
Improved the quality of the Russian-language general:rc model in the following areas:
- Recognition of last and first names, patronymics, and addresses.
- Recognition of customer-specific terms. The model was enhanced with the data from the user request dated February 1, 2022, and corrected based on the user data from November 9, 2021.
- Punctuation in long sentences and texts with numbers.
Release as of 03/02/22
-
general:rcnow supports universal mode (the"auto"language). In this mode, the model can recognize speech in one of the following languages:- Russian
- Kazakh
- English
- German
- French
- Finnish
- Swedish
- Dutch
- Polish
- Portuguese
- Italian
- Spanish
-
New languages are also available under their own codes. The
general:rcmodel uses an indication as a hint for language recognition. If the language is indicated explicitly, the model will use it as a hint to improve the recognition quality. Currently, it only affects the recognition quality in Russian.
Known problems: In universal mode, the recognition quality may deteriorate in the case of continuous speech without pauses.
Release as of 26/01/22
-
The
generalandgeneral:rcrecognition models for Kazakh are now available in streaming and deferred recognition modes. -
general:rcnow supports a punctuator in streaming and deferred recognition modes. -
In deferred recognition mode, you can now work with MP3 format.