skillfed

Speech packages

22 packages

Packages

  • SpeechRecognition

    Performs speech recognition and transcription…

    permissive · active · 11.7M/mo

  • gTTS

    gTTS converts text to speech using Google…

    permissive · active · 5.3M/mo

  • lhotse

    Lhotse prepares multimodal (speech, audio,…

    permissive · active · 1.3M/mo

  • piper-tts

    Piper TTS is a local neural text-to-speech…

    copyleft · active · 891.3K/mo

  • funasr

    FunASR is a speech recognition toolkit that…

    permissive · active · 497.9K/mo

  • pocketsphinx

    PocketSphinx provides Python bindings for…

    permissive · active · 382.6K/mo

  • aic-sdk

    Python bindings for ai-coustics audio…

    permissive · active · 296.4K/mo

  • speechmatics-rt

    Async Python client for real-time…

    permissive · active · 265.7K/mo

  • onnx-asr

    Automatic Speech Recognition using ONNX models…

    permissive · active · 230.3K/mo

  • speechmatics-voice

    Python SDK for building real-time voice…

    permissive · active · 209.5K/mo

  • omnivoice

    OmniVoice generates speech from text in over…

    permissive · active · 206.4K/mo

  • fish-audio-sdk

    Official Python client for the Fish Audio API,…

    permissive · active · 192.1K/mo

  • coqui-tts

    Coqui TTS synthesizes speech from text using…

    copyleft · active · 183.7K/mo

  • sea-g2p

    Converts text to phonemes for Vietnamese, Thai,…

    permissive · active · 168.6K/mo

  • pvporcupine

    Porcupine is a lightweight wake word detection…

    permissive · active · 160.1K/mo

  • monotonic-alignment-search

    Finds the most probable alignment between a…

    permissive · aging · 131.0K/mo

  • TTS

    TTS is a deep learning library for…

    copyleft · dormant · 108.1K/mo

  • ko-speech-tools

    Provides Korean language processing tools…

    permissive · aging · 103.2K/mo

  • syncedlyrics

    Fetches synchronized lyrics in LRC format for…

    permissive · dormant · 101.4K/mo

  • deepfilternet

    DeepFilterNet removes background noise from…

    permissive · dormant · 78.6K/mo

  • kugelaudio

    Official Python SDK for KugelAudio's…

    permissive · active · 78.6K/mo

  • speechmatics-batch

    Async Python client for submitting audio files…

    permissive · active · 75.4K/mo