Speech recognition module for Python, supporting several engines and APIs, online and offline.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.
DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on devices ranging from a Raspberry Pi 4 to high power GPU servers.
Neural building blocks for speaker diarization: speech activity detection, speaker change detection, overlapped speech detection, speaker embedding
Python Audio Analysis Library: Feature Extraction, Classification, Segmentation and Applications
Python interface to the WebRTC Voice Activity Detector