$npx skillfedfor your agent

polyglot

Polyglot is a natural language pipeline that supports massive multilingual applications.

SkipPyPI Artificial IntelligenceReleased Jul 201691.2K downloads / moGPLv3Source build

Decision gist · record as of 2026-08-14

sdist only — polyglot-16.7.4.tar.gz · builds from source
v16.7.4 · released 2016-07-03

No—not for new projects. The package is dormant (last release 2016-07-03), classified as Beta, and carries high install friction due to external model dependencies. Maintenance status is unclear, and compatibility with modern Python versions is unverified. For current multilingual NLP work, consider actively maintained alternatives.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • Model data must be downloaded separately; the package itself has no runtime dependencies but requires external language models to function.
  • Installation friction is high; the package has been dormant since 2016-07-03 and carries no runtime dependencies, suggesting it may require manual model downloads or system-level setup to function.
  • Beta status and age raise questions about compatibility with current Python versions.

License · maintenance · safety

GPLv3 (copyleft) — GPLv3 is a copyleft license; any derivative work or modification must also be licensed under GPLv3, and distribution of the package in a proprietary application may require source disclosure or licensing negotiation.

last release 2016-07-03 (3694 days) · last repo commit 2023-11-10 · 2,362 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 91,223 downloads/mo, #13,531 on PyPI

Verify before relying

import polyglot
from polyglot.text import Text, Word

text = Text("Bonjour, Mesdames.")
print(text.language.code, text.language.name)
print(text.words)
print(text.pos_tags)
  • Whether model downloads work with current mirror or if Stony Brook DSL lab mirror is still active.
  • Python 3.4+ compatibility status; classifiers list Python 3.4 but package is dormant since 2016.
  • Whether high install friction is due to compiled dependencies or model data requirements.
  • Current maintenance status and whether the GitHub repository accepts issues or pull requests.
Same gist for agents: .md · .json

What it is and what it does

Polyglot is a multilingual natural language processing library that wraps language models and linguistic algorithms to perform common NLP tasks across many languages. It provides a unified interface for tokenization, language detection, named entity recognition, part-of-speech tagging, sentiment analysis, word embeddings, morphological analysis, and transliteration. The package supports between 16 and 196 languages depending on the task.

The library works by loading pre-trained models for each language and task, then exposing them through a simple Python API. You instantiate a Text object with a string, and then access properties like `.language`, `.words`, `.sentences`, `.pos_tags`, `.entities`, and `.polarity` to retrieve annotations. Individual words can be examined for embeddings, morphemes, and transliteration. However, the package has not been actively maintained since 2016-07-03, and installation requires downloading external model data, which may present compatibility or availability challenges.

Use it for

  • Detect the language of user-submitted text to route it to language-specific processing pipelines.
  • Extract named entities (people, locations, organizations) from multilingual documents for information extraction.
  • Perform sentiment analysis on social media posts or reviews in languages other than English.
  • Tokenize and tag parts of speech in non-English text for downstream linguistic analysis or machine learning.
  • Transliterate text between writing systems (e.g., English to Cyrillic) for cross-script applications.
  • Generate word embeddings for multilingual semantic similarity or clustering tasks.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

Skip

No—not for new projects.

The package is dormant (last release 2016-07-03), classified as Beta, and carries high install friction due to external model dependencies. Maintenance status is unclear, and compatibility with modern Python versions is unverified. For current multilingual NLP work, consider actively maintained alternatives.

Install

polyglot on PyPI

Before you install

Installation friction is high; the package has been dormant since 2016-07-03 and carries no runtime dependencies, suggesting it may require manual model downloads or system-level setup to function. Beta status and age raise questions about compatibility with current Python versions.

Model data must be downloaded separately; the package itself has no runtime dependencies but requires external language models to function.

License in practice

GPLv3 is a copyleft license; any derivative work or modification must also be licensed under GPLv3, and distribution of the package in a proprietary application may require source disclosure or licensing negotiation.

Quickstart

import polyglot
from polyglot.text import Text, Word

text = Text("Bonjour, Mesdames.")
print(text.language.code, text.language.name)
print(text.words)
print(text.pos_tags)

Verify before relying

  • Whether model downloads work with current mirror or if Stony Brook DSL lab mirror is still active.
  • Python 3.4+ compatibility status; classifiers list Python 3.4 but package is dormant since 2016.
  • Whether high install friction is due to compiled dependencies or model data requirements.
  • Current maintenance status and whether the GitHub repository accepts issues or pull requests.

Package facts

LicenseGPLv3 copyleft
Python supportNot specified
Install frictionHigh. Source build required
Runtime dependenciesNone
MaintenanceDormant 3,694 days since the last release
Last repo commit
First released
Downloads91,223 / month, #13,531 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
Development Status :: 4 - BetaEnvironment :: ConsoleIntended Audience :: EducationIntended Audience :: Science/ResearchLicense :: OSI Approved :: GNU General Public License v3 or later (GPLv3+)Natural Language :: AfrikaansNatural Language :: ArabicNatural Language :: BengaliNatural Language :: BosnianNatural Language :: BulgarianNatural Language :: CatalanNatural Language :: Chinese (Simplified)Natural Language :: Chinese (Traditional)Natural Language :: CroatianNatural Language :: CzechNatural Language :: DanishNatural Language :: DutchNatural Language :: EnglishNatural Language :: EsperantoNatural Language :: FinnishNatural Language :: FrenchNatural Language :: GalicianNatural Language :: GermanNatural Language :: GreekNatural Language :: HebrewNatural Language :: HindiNatural Language :: HungarianNatural Language :: IcelandicNatural Language :: IndonesianNatural Language :: ItalianNatural Language :: JapaneseNatural Language :: JavaneseNatural Language :: KoreanNatural Language :: LatinNatural Language :: LatvianNatural Language :: MacedonianNatural Language :: MalayNatural Language :: MarathiNatural Language :: NorwegianNatural Language :: PanjabiNatural Language :: PersianNatural Language :: PolishNatural Language :: PortugueseNatural Language :: Portuguese (Brazilian)Natural Language :: RomanianNatural Language :: RussianNatural Language :: SerbianNatural Language :: SlovakNatural Language :: SlovenianNatural Language :: SpanishNatural Language :: SwedishNatural Language :: TamilNatural Language :: TeluguNatural Language :: ThaiNatural Language :: TurkishNatural Language :: UkranianNatural Language :: UrduNatural Language :: VietnameseProgramming Language :: Python :: 2Programming Language :: Python :: 2.7Programming Language :: Python :: 3Programming Language :: Python :: 3.4Topic :: Scientific/Engineering :: Artificial IntelligenceTopic :: Text Processing :: Linguistic

Evidence: polyglot-16.7.4.tar.gz

Tags

Capabilities
multilingual NLP pipelinelanguage detection multiple languagesnamed entity recognition NLPpart of speech taggingsentiment analysis multilingualword embeddings translationmorphological analysis
Topics
multilingual-nlpdormant-unmaintained
PyPI keywords
polyglot

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “part of speech tagging”

  • polyglotPolyglot is a multilingual natural language processing pipeline that…
  • sherpa-onnxSherpa-onnx runs speech recognition, text-to-speech, speaker…
  • pymorphy2Morphological analyzer and inflection engine for Russian and…

Give your agent the search over MCP, or paste the wish link into any chat.

More Artificial Intelligence packages

litellm With conditions
PyPI · Artificial Intelligence · released Aug 2026

LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.

Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.

MITcompiled wheel
682.8Mdownloads / mo
huggingface-hub Worth it
PyPI · Artificial Intelligence · released Aug 2026

Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.

Install it if you work with Hugging Face Hub models or datasets.

Apache-2.0pure Python · 3.10.0+
442.4Mdownloads / mo
langchain Worth it
PyPI · Python Modules · released Aug 2026

LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.

MITpure Python
315.4Mdownloads / mo
hf-xet With conditions
PyPI · Artificial Intelligence · released Aug 2026

hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.

Apache-2.0compiled wheel · 3.8+
258.4Mdownloads / mo
tokenizers Worth it
PyPI · Artificial Intelligence · released Apr 2026

Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.

Apache-2.0compiled wheel · 3.10+
222.9Mdownloads / mo
transformers Worth it
PyPI · Artificial Intelligence · released Aug 2026

Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.

Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.

permissive licensepure Python · 3.10.0+
186.6Mdownloads / mo

See also konlpy · wn · nagisa · conllu · textblob · urduhack · spark-nlp · ginza · stanza · tokenizer