phonemizer-fork
Simple text to phones converter for multiple languages
Decision gist · record as of 2026-08-14
Yes, if you need multilingual text-to-phoneme conversion. The package is actively maintained, has low install friction, and offers flexible backend selection for different use cases. The GPLv3+ license is a hard constraint only if you're building proprietary software; for research, open-source, or internal projects it poses no barrier. No known vulnerabilities.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Backend engines (espeak, festival, espeak-mbrola) may require separate system installation depending on which backend you select.
- Low install friction with a pure-Python wheel and only 5 runtime dependencies.
- Repository is active with recent commits and 1567 stars.
License · maintenance · safety
copyleft license (copyleft) — Licensed under GPLv3+, a copyleft license requiring that any derivative works or distributions remain under the same terms. Suitable for open-source projects but incompatible with proprietary software that cannot be released under GPL.
last release 2025-01-30 (561 days) · last repo commit 2026-08-04 · 1,567 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 966,328 downloads/mo, #4,619 on PyPI
Alternatives
Verify before relying
pip install phonemizer-fork
from phonemizer.phonemize import phonemize
phones = phonemize('hello', language='en-us', backend='espeak')- Whether backend engines are bundled or must be installed separately on the system
- Whether this fork is actively maintained or has diverged from the original phonemizer package
- Exact language and feature support for each backend beyond the comparison table
What it is and what it does
Phonemizer-fork is a Python library that converts written text into phonetic transcriptions—sequences of individual speech sounds (phones)—in multiple languages. It wraps four different phonemization backends (espeak, espeak-mbrola, festival, and segments), each with different capabilities: espeak supports IPA output for many languages at fast speed, espeak-mbrola uses SAMPA for a smaller set of languages but slower, festival handles American English with syllable-level tokenization, and segments uses user-defined grapheme-to-phoneme mappings. The library provides both a command-line tool and a Python API for straightforward text-to-phone conversion.
You would use this package when you need to convert text into phonetic form for speech synthesis, linguistic analysis, or training speech-related machine learning models. The choice of backend depends on your language, required phonetic alphabet (IPA vs. SAMPA vs. custom), and whether you need syllable-level or word-level tokenization. It handles punctuation preservation and stressed phones where supported, and integrates with joblib for parallel processing.
Use it for
- Prepare training data for speech systems by converting text to phonetic representations using espeak or festival
- Analyze pronunciation patterns across languages for linguistic research or language learning applications
- Build custom phonemization workflows for languages using the segments backend with user-provided mappings
- Extract syllable-level phonetic tokens for prosody analysis or rhythm-based speech processing in English
- Convert multilingual text to IPA for comparative phonetic studies
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need multilingual text-to-phoneme conversion.
The package is actively maintained, has low install friction, and offers flexible backend selection for different use cases. The GPLv3+ license is a hard constraint only if you're building proprietary software; for research, open-source, or internal projects it poses no barrier. No known vulnerabilities.
Install
phonemizer-fork on PyPI
Before you install
Low install friction with a pure-Python wheel and only 5 runtime dependencies. Repository is active with recent commits and 1567 stars.
Backend engines (espeak, festival, espeak-mbrola) may require separate system installation depending on which backend you select.
License in practice
Licensed under GPLv3+, a copyleft license requiring that any derivative works or distributions remain under the same terms. Suitable for open-source projects but incompatible with proprietary software that cannot be released under GPL.
Quickstart
pip install phonemizer-fork
from phonemizer.phonemize import phonemize
phones = phonemize('hello', language='en-us', backend='espeak')
Verify before relying
- Whether backend engines are bundled or must be installed separately on the system
- Whether this fork is actively maintained or has diverged from the original phonemizer package
- Exact language and feature support for each backend beyond the comparison table
Package facts
| License | copyleft license copyleft |
| Python support | Supports the current Python release >=3.8 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 5 packagesattrsdlinfojoblibsegmentstyping-extensions |
| Maintenance | Actively maintained 561 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 966,328 / month, #4,619 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | License :: OSI Approved :: GNU General Public License v3 or later (GPLv3+)Operating System :: OS IndependentProgramming Language :: Python :: 3 |
Evidence: phonemizer_fork-3.3.2-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “text to phonemes converter”
- phonemizer-forkConverts text to phonetic representations (phones) in multiple…
- sea-g2pConverts text to phonemes for Vietnamese, Thai, and Indonesian with…
- gruut-lang-enProvides English language data files for tokenization and IPA phoneme…
Give your agent the search over MCP, or paste the wish link into any chat.
More Linguistic packages
Detects and normalizes text encoding from unknown or ambiguous sources, supporting all IANA character sets that Python's core library provides codecs for, with the ability to register custom codecs.
tiktoken is a fast BPE tokenizer that converts text into token sequences compatible with OpenAI models, supporting multiple encoding schemes including o200k_base and model-specific encodings.
Install it if you work with OpenAI APIs or need to understand token boundaries in GPT-family models.
Detects character encoding and language in byte sequences with high accuracy, supporting 99 encodings and returning confidence scores, language tags, and MIME types.
Install it if you need to detect character encoding or language in byte data; the rewrite makes it substantially faster and more accurate than its predecessors.
Converts Unicode text to ASCII by transliterating non-ASCII characters into their closest ASCII equivalents, with no runtime dependencies.
However, if transliteration quality or ongoing maintenance matters, consider unidecode instead despite its GPL-only license.
Lark is a parsing library that builds abstract syntax trees from context-free grammars, supporting multiple parsing algorithms (Earley, LALR(1), CYK) with automatic line and column tracking.
Python bindings to the tree-sitter parsing library, enabling incremental parsing and syntax tree analysis for source code.
See also espeakng-loader · phonemizer · gruut-ipa · misaki · panphon · orthography2ipa · g2p-en · gruut · gruut-lang-en · epitran