Packages
-
SpeechRecognition
Performs speech recognition and transcription…
permissive · active · 11.7M/mo
-
gTTS
gTTS converts text to speech using Google…
permissive · active · 5.3M/mo
-
lhotse
Lhotse prepares multimodal (speech, audio,…
permissive · active · 1.3M/mo
-
piper-tts
Piper TTS is a local neural text-to-speech…
copyleft · active · 891.3K/mo
-
funasr
FunASR is a speech recognition toolkit that…
permissive · active · 497.9K/mo
-
pocketsphinx
PocketSphinx provides Python bindings for…
permissive · active · 382.6K/mo
-
aic-sdk
Python bindings for ai-coustics audio…
permissive · active · 296.4K/mo
-
speechmatics-rt
Async Python client for real-time…
permissive · active · 265.7K/mo
-
onnx-asr
Automatic Speech Recognition using ONNX models…
permissive · active · 230.3K/mo
-
speechmatics-voice
Python SDK for building real-time voice…
permissive · active · 209.5K/mo
-
omnivoice
OmniVoice generates speech from text in over…
permissive · active · 206.4K/mo
-
fish-audio-sdk
Official Python client for the Fish Audio API,…
permissive · active · 192.1K/mo
-
coqui-tts
Coqui TTS synthesizes speech from text using…
copyleft · active · 183.7K/mo
-
sea-g2p
Converts text to phonemes for Vietnamese, Thai,…
permissive · active · 168.6K/mo
-
pvporcupine
Porcupine is a lightweight wake word detection…
permissive · active · 160.1K/mo
-
monotonic-alignment-search
Finds the most probable alignment between a…
permissive · aging · 131.0K/mo
-
TTS
TTS is a deep learning library for…
copyleft · dormant · 108.1K/mo
-
ko-speech-tools
Provides Korean language processing tools…
permissive · aging · 103.2K/mo
-
syncedlyrics
Fetches synchronized lyrics in LRC format for…
permissive · dormant · 101.4K/mo
-
deepfilternet
DeepFilterNet removes background noise from…
permissive · dormant · 78.6K/mo
-
kugelaudio
Official Python SDK for KugelAudio's…
permissive · active · 78.6K/mo
-
speechmatics-batch
Async Python client for submitting audio files…
permissive · active · 75.4K/mo