Package profile
SpeechRecognition
- Summary: Library for performing speech recognition, with support for several engines and APIs, online and offline.
- Author: "Anthony Zhang (Uberi)" <azhang9@gmail.com>
- License: BSD-3-Clause
- Homepage: https://github.com/Uberi/speech_recognition#readme
- Source: https://github.com/Uberi/speech_recognition#readme
- Number of releases: 75
- First release: 1.0.0 on 2014-04-23
- Latest release: 3.17.0 on 2026-06-17
- Latest release size: 31.3 MB (pure Python wheel)
Dependencies
SpeechRecognition has 23 dependencies, 20 of which optional.Dependent packages
| Package | Optional | Group |
|---|---|---|
| textract | false | |
| markitdown | true | all |
Similar packages
- gttsgTTS (Google Text-to-Speech), a Python library and CLI tool to interface with Google Translate text-to-speech API
- google-cloud-speechGoogle Cloud Speech API client library
- pyttsx3Text to Speech (TTS) library for Python 3. Works without internet connection or delay. Supports multiple TTS engines, including Sapi5, nsss, and espeak.
- webrtcvad-wheelsPython interface to the Google WebRTC Voice Activity Detector (VAD) [released with binary wheels!]
- audioreadMulti-library, cross-platform audio decoding.
- langdetectLanguage detection library ported from Google's language-detection.
- soundfileAn audio library based on libsndfile, CFFI and NumPy