TTS
Deep learning for Text to Speech by Coqui.
Decision gist · record as of 2026-08-14
Yes, with conditions. Install if you need neural TTS synthesis and can accept the MPL-2.0 copyleft constraint and substantial dependency footprint. The library has recent commits and is in active use, but has not released since 2023-12-12. Avoid if you require active maintenance guarantees or proprietary licensing flexibility.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Python >=3.9.0, <3.12.
- Large model downloads on first use; GPU recommended for inference speed but CPU inference supported.
- Medium install friction due to 39 runtime dependencies including torch, torchaudio, scipy, and librosa.
License · maintenance · safety
MPL-2.0 (copyleft) — Licensed under MPL-2.0, a copyleft license requiring derivative works to be distributed under the same license and disclosing source code modifications. This affects proprietary deployments.
last release 2023-12-12 (976 days) · last repo commit 2024-08-16 · 45,899 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 108,143 downloads/mo, #12,574 on PyPI
Alternatives
Verify before relying
pip install TTS
from TTS.api import TTS
tts = TTS(model_name="tts_models/en/ljspeech/glow-tts", gpu=False)
tts.tts_to_file(text="Hello world", file_path="output.wav")- Whether dormant status indicates active maintenance or abandonment despite recent commits.
- Performance characteristics and inference latency for different model architectures.
- Memory requirements for different pretrained models during inference and training.
- Actual number of supported languages and models available in the pretrained collection.
What it is and what it does
TTS is a PyTorch-based library for neural text-to-speech synthesis that provides pretrained models and architectures including Tacotron2, Glow-TTS, VITS, XTTS, Bark, and Tortoise. It handles the full pipeline from text input to audio output, supporting both inference with released models and training custom models on new datasets.
The library depends on a substantial stack of 39 runtime packages: torch, torchaudio, scipy, librosa, scikit-learn, and language-specific tools including jieba, g2pkk, bangla, jamo, hangul-romanize, gruut, and pysbd. It includes vocoder models (MelGAN, HiFiGAN, ParallelWaveGAN) to convert spectrograms to waveforms, speaker encoders for multi-speaker synthesis, and utilities for dataset curation. Installation is straightforward via pip, though the dependency footprint and model downloads make it a medium-friction package.
Use it for
- Generate speech from text in applications requiring voice output, using pretrained models without training.
- Train custom TTS models on proprietary voice datasets for domain-specific or branded voice synthesis.
- Implement voice cloning by fine-tuning existing models with speaker-specific audio samples.
- Build multilingual speech synthesis pipelines with language-specific models and tools.
- Integrate TTS into web services or applications via the library's Flask-based server for synthesis.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, with conditions.
Install if you need neural TTS synthesis and can accept the MPL-2.0 copyleft constraint and substantial dependency footprint. The library has recent commits and is in active use, but has not released since 2023-12-12. Avoid if you require active maintenance guarantees or proprietary licensing flexibility.
Install
tts on PyPI
Before you install
Medium install friction due to 39 runtime dependencies including torch, torchaudio, scipy, and librosa. The project is dormant with last commit on 2024-08-16. Prebuilt wheels available for Python 3.9–3.11 on Linux x86_64 reduce friction for standard environments.
Requires Python >=3.9.0, <3.12. Large model downloads on first use; GPU recommended for inference speed but CPU inference supported.
License in practice
Licensed under MPL-2.0, a copyleft license requiring derivative works to be distributed under the same license and disclosing source code modifications. This affects proprietary deployments.
Quickstart
pip install TTS
from TTS.api import TTS
tts = TTS(model_name="tts_models/en/ljspeech/glow-tts", gpu=False)
tts.tts_to_file(text="Hello world", file_path="output.wav")
Verify before relying
- Whether dormant status indicates active maintenance or abandonment despite recent commits.
- Performance characteristics and inference latency for different model architectures.
- Memory requirements for different pretrained models during inference and training.
- Actual number of supported languages and models available in the pretrained collection.
Package facts
| License | MPL-2.0 copyleft |
| Python support | Capped below the current Python release >=3.9.0, <3.12 |
| Install friction | Medium. Platform-specific wheel |
| Runtime dependencies | 39 packagescythonscipytorchtorchaudiosoundfilelibrosascikit-learninflecttqdmanyasciipyyamlfsspecaiohttppackagingflaskpysbdumap-learnpandasmatplotlibtrainercoqpitjiebapypinyinhangul-romanizegruutjamonltkg2pkkbanglabnnumerizer |
| Maintenance | Dormant 976 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 108,143 / month, #12,574 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 3 - AlphaIntended Audience :: DevelopersIntended Audience :: Science/ResearchLicense :: OSI Approved :: Mozilla Public License 2.0 (MPL 2.0)Operating System :: POSIX :: LinuxProgramming Language :: PythonProgramming Language :: Python :: 3Programming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.9Topic :: MultimediaTopic :: Multimedia :: Sound/AudioTopic :: Multimedia :: Sound/Audio :: SpeechTopic :: Scientific/Engineering :: Artificial IntelligenceTopic :: Software DevelopmentTopic :: Software Development :: Libraries :: Python Modules |
Evidence: TTS-0.22.0-cp310-cp310-manylinux1_x86_64.whl; TTS-0.22.0-cp311-cp311-manylinux1_x86_64.whl; TTS-0.22.0-cp39-cp39-manylinux1_x86_64.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “text to speech synthesis”
- TTSTTS is a deep learning library for text-to-speech synthesis that…
- f5-ttsF5-TTS generates natural-sounding speech from text using…
- pyopenjtalkWraps OpenJTalk to provide Japanese text-to-speech synthesis,…
Give your agent the search over MCP, or paste the wish link into any chat.
More Software Development packages
Provides backported and experimental type hints for Python 3.9+, allowing use of newer typing features on older Python versions and enabling early experimentation with type system PEPs before they enter the standard library.
NumPy provides an N-dimensional array object and a comprehensive suite of mathematical, linear algebra, Fourier transform, and random number functions for scientific computing in Python.
FastAPI is a Python web framework for building REST APIs using type hints, with automatic request validation, serialization, and interactive API documentation.
Provides a way to document function parameters, class attributes, return types, and variables inline using Python's `Annotated` type hint syntax instead of traditional docstrings.
Typer builds command-line applications from Python functions using type hints, automatically generating help text, argument parsing, and shell completion.
Install it if you are building CLIs in Python.
Distlib provides low-level packaging utilities for building, distributing, and managing Python software—including metadata handling, version specifiers, wheel support, script installation, and dependency resolution.
See also coqui-tts · pocket-tts · speechbrain · chatterbox-tts · mlx-audio · piper-tts · pyttsx3 · kokoro-onnx · monotonic-alignment-search · voxcpm