$npx skillfedfor your agent

kokoro

TTS

With conditionsPyPI LinguisticReleased Apr 2025622.8K downloads / mopermissive licensePure Python

Decision gist · record as of 2026-08-14

pure-Python wheel — kokoro-0.9.4-py3-none-any.whl
v0.9.4 · released 2025-04-05 · Python <3.13,>=3.10 · 6 runtime deps: huggingface-hub, loguru, misaki, numpy, torch, transformers

Yes, with conditions. Install kokoro if you need a lightweight, open-weight TTS model for prototyping or production use and can tolerate the aging maintenance status (last release 496 days ago). The Apache license is permissive, install friction is low, and the package has solid community adoption. However, verify that the preset voices and supported languages meet your needs, and be aware that torch and transformers are heavy dependencies. If you require active development or frequent updates, check the repository first.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • Requires espeak-ng system library (install via apt-get on Linux, .msi installer on Windows, or conda).
  • Python >=3.10,<3.13 only.
  • torch and transformers are heavy dependencies; initial model download from Hugging Face Hub occurs on first use.

License · maintenance · safety

permissive license (permissive) — Apache License 2.0 is permissive and allows commercial use, modification, and redistribution with minimal restrictions. You can deploy this package in production or personal projects without licensing concerns.

last release 2025-04-05 (496 days) · last repo commit 2025-08-06 · 8,421 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 622,780 downloads/mo, #5,708 on PyPI

Verify before relying

pip install kokoro>=0.9.4
from kokoro import KPipeline
pipeline = KPipeline(lang_code='a')
generator = pipeline('Hello world', voice='af_heart')
for gs, ps, audio in generator:
    print(gs, ps)  # graphemes, phonemes
    # audio is a numpy array at 24000 Hz sample rate
  • Whether voice cloning or custom voice loading is supported beyond the preset voices mentioned in examples.
  • Inference speed and memory footprint on CPU-only systems or edge devices.
  • Quality comparison to larger commercial TTS models in production scenarios.
  • Support status and roadmap given the aging maintenance signal (last release 496 days ago).
Same gist for agents: .md · .json

What it is and what it does

Kokoro is a Python wrapper around the Kokoro-82M text-to-speech model, a lightweight neural network with 82 million parameters designed to generate natural-sounding speech from text. It handles phoneme conversion, voice synthesis, and audio generation in a single pipeline, supporting multiple languages (American English, British English, Spanish, French, Hindi, Italian, Japanese, Brazilian Portuguese, and Mandarin Chinese) through language-specific codes and optional misaki extensions.

The library is built for both interactive use (Jupyter notebooks, Google Colab) and production deployment. It depends on torch for inference, transformers for model loading, huggingface-hub for downloading weights, and misaki for grapheme-to-phoneme conversion. Audio output is generated at 24000 Hz sample rate. The model weights are Apache-licensed, making it suitable for commercial and personal projects without licensing friction.

Use it for

  • Generate speech from long-form text in Jupyter notebooks or Colab for prototyping and testing voice synthesis.
  • Build a multilingual chatbot or voice assistant that speaks in multiple languages with preset voice profiles.
  • Create audiobook or podcast narration pipelines by splitting text and synthesizing each section with consistent voice.
  • Deploy a lightweight TTS service in resource-constrained environments where model size and inference speed matter.
  • Experiment with different voice profiles and languages without managing model weights or phoneme rules manually.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

With conditions

Yes, with conditions.

Install kokoro if you need a lightweight, open-weight TTS model for prototyping or production use and can tolerate the aging maintenance status (last release 496 days ago). The Apache license is permissive, install friction is low, and the package has solid community adoption. However, verify that the preset voices and supported languages meet your needs, and be aware that torch and transformers are heavy dependencies. If you require active development or frequent updates, check the repository first.

Install

kokoro on PyPI

Before you install

Low install friction with a pure-Python wheel. The package depends on torch, transformers, huggingface-hub, numpy, loguru, and misaki—all standard ML dependencies. Maintenance status is aging (last release 496 days ago), though the repository remains active with recent commits and substantial community interest (8421 stars).

Requires espeak-ng system library (install via apt-get on Linux, .msi installer on Windows, or conda). Python >=3.10,<3.13 only. torch and transformers are heavy dependencies; initial model download from Hugging Face Hub occurs on first use.

License in practice

Apache License 2.0 is permissive and allows commercial use, modification, and redistribution with minimal restrictions. You can deploy this package in production or personal projects without licensing concerns.

Quickstart

pip install kokoro>=0.9.4
from kokoro import KPipeline
pipeline = KPipeline(lang_code='a')
generator = pipeline('Hello world', voice='af_heart')
for gs, ps, audio in generator:
    print(gs, ps)  # graphemes, phonemes
    # audio is a numpy array at 24000 Hz sample rate

Verify before relying

  • Whether voice cloning or custom voice loading is supported beyond the preset voices mentioned in examples.
  • Inference speed and memory footprint on CPU-only systems or edge devices.
  • Quality comparison to larger commercial TTS models in production scenarios.
  • Support status and roadmap given the aging maintenance signal (last release 496 days ago).

Package facts

Licensepermissive license permissive
Python supportCapped below the current Python release <3.13,>=3.10
Install frictionLow. Pure-Python wheel
Runtime dependencies
6 packages
huggingface-hublogurumisakinumpytorchtransformers
MaintenanceAging 496 days since the last release
Last repo commit
First released
Downloads622,780 / month, #5,708 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
License :: OSI Approved :: Apache Software LicenseOperating System :: OS IndependentProgramming Language :: Python :: 3

Evidence: kokoro-0.9.4-py3-none-any.whl

Tags

Capabilities
text to speech inferenceTTS model deploymentmultilingual speech synthesislightweight neural TTSopen-weight voice generationkokoro TTS libraryspeech audio generation
Topics
text-to-speechneural-ttsmultilingual

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “TTS model deployment”

  • kokoroKokoro is an inference library for the Kokoro-82M text-to-speech…
  • kokoro-onnxConverts text to speech using ONNX Runtime, supporting multiple…
  • sileroSilero provides pre-trained text-to-speech models that convert text…

Give your agent the search over MCP, or paste the wish link into any chat.

More Linguistic packages

charset-normalizer Worth it
PyPI · Utilities · released Aug 2026

Detects and normalizes text encoding from unknown or ambiguous sources, supporting all IANA character sets that Python's core library provides codecs for, with the ability to register custom codecs.

permissive licensepure Python · 3.7+
1.7Bdownloads / mo
tiktoken Worth it
PyPI · Linguistic · released May 2026

tiktoken is a fast BPE tokenizer that converts text into token sequences compatible with OpenAI models, supporting multiple encoding schemes including o200k_base and model-specific encodings.

Install it if you work with OpenAI APIs or need to understand token boundaries in GPT-family models.

permissive licensecompiled wheel · 3.9+
233.0Mdownloads / mo
chardet Worth it
PyPI · Python Modules · released Aug 2026

Detects character encoding and language in byte sequences with high accuracy, supporting 99 encodings and returning confidence scores, language tags, and MIME types.

Install it if you need to detect character encoding or language in byte data; the rewrite makes it substantially faster and more accurate than its predecessors.

0BSDpure Python · 3.10+
199.0Mdownloads / mo
text-unidecode With conditions
PyPI · Python Modules · released Aug 2019

Converts Unicode text to ASCII by transliterating non-ASCII characters into their closest ASCII equivalents, with no runtime dependencies.

However, if transliteration quality or ongoing maintenance matters, consider unidecode instead despite its GPL-only license.

GPL-2.0-or-laterpure Pythonabandoned
89.0Mdownloads / mo
lark Worth it
PyPI · Python Modules · released Oct 2025

Lark is a parsing library that builds abstract syntax trees from context-free grammars, supporting multiple parsing algorithms (Earley, LALR(1), CYK) with automatic line and column tracking.

MITpure Python · 3.8+
79.7Mdownloads / mo
tree-sitter Worth it
PyPI · Linguistic · released Jun 2026

Python bindings to the tree-sitter parsing library, enabling incremental parsing and syntax tree analysis for source code.

MITcompiled wheel · 3.10+
79.0Mdownloads / mo

See also kokoro-onnx · misaki · chatterbox-tts · espeakng-loader · lhotse · coqui-tts · pocket-tts · piper-tts · omnivoice · mlx-audio