emot
Emoji and Emoticons detection package for Python
Decision gist · record as of 2026-08-14
Yes, if you need straightforward emoji and emoticon extraction with no external dependencies. However, exercise caution: the project is dormant (last release 2021-08-02), the GPL license terms are unclear, and Python version support is unspecified. Suitable for one-off scripts or legacy systems, but risky for new production code without verifying GPL compliance and testing against your target Python version.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Low friction install with no runtime dependencies.
- Maintenance is dormant—last release was 2021-08-02, though the repository remains active with 197 stars and a final commit in 2023-11-02.
License · maintenance · safety
GNU GENERAL PUBLIC LICENSE (unclear) — License treatment is unclear; the raw license states 'GNU GENERAL PUBLIC LICENSE' but no SPDX identifier is provided, making the exact terms and compatibility obligations ambiguous.
last release 2021-08-02 (1838 days) · last repo commit 2023-11-02 · 197 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 272,169 downloads/mo, #8,210 on PyPI
Alternatives
Verify before relying
import emot
emot_obj = emot.core.emot()
text = "I love python ☮ 🙂 ❤ :-) :-( :-)))"
result = emot_obj.emoji(text)
print(result) # {'value': ['☮', '🙂', '❤'], 'location': [[14, 15], [16, 17], [18, 19]], 'mean': [':peace_symbol:', ':slightly_smiling_face:', ':red_heart:'], 'flag': True}- Whether the GPL license is version 2, 3, or unspecified, and what that means for derivative works or commercial use.
- Current compatibility with modern Python versions and whether Python 3.X support extends to recent releases.
- Whether the multiprocessing bulk functions actually deliver measurable speedup on typical datasets.
What it is and what it does
Emot is a Python library that identifies emojis and emoticons within text strings and returns structured data about each match. It parses Unicode emoji characters and ASCII emoticon patterns (like :-) and :-() from input text, reporting the matched symbols, their byte positions, and their semantic meanings in a dictionary with a success flag.
The library generates detection patterns dynamically from an internal emoji database each time an object is instantiated, allowing customization by editing the database file. Version 3.0 added bulk processing functions that use multiprocessing to handle large text datasets, distributing work across available CPU cores (defaulting to half the system total) for potential performance gains on many strings at once.
Use it for
- Sentiment analysis pipelines that need to extract and interpret emojis as sentiment signals from social media text.
- Content moderation systems that identify and locate emojis or emoticons in user-generated text for filtering or flagging.
- Data preprocessing for NLP tasks where emoji presence or meaning should be captured as features or metadata.
- Bulk text processing of large datasets where extracting all emoji/emoticon occurrences and their positions is required.
- Text annotation or enrichment workflows that map emojis to their Unicode names for downstream analysis or display.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need straightforward emoji and emoticon extraction with no external dependencies.
However, exercise caution: the project is dormant (last release 2021-08-02), the GPL license terms are unclear, and Python version support is unspecified. Suitable for one-off scripts or legacy systems, but risky for new production code without verifying GPL compliance and testing against your target Python version.
Install
emot on PyPI
Before you install
Low friction install with no runtime dependencies. Maintenance is dormant—last release was 2021-08-02, though the repository remains active with 197 stars and a final commit in 2023-11-02.
License in practice
License treatment is unclear; the raw license states 'GNU GENERAL PUBLIC LICENSE' but no SPDX identifier is provided, making the exact terms and compatibility obligations ambiguous.
Quickstart
import emot
emot_obj = emot.core.emot()
text = "I love python ☮ 🙂 ❤ :-) :-( :-)))"
result = emot_obj.emoji(text)
print(result) # {'value': ['☮', '🙂', '❤'], 'location': [[14, 15], [16, 17], [18, 19]], 'mean': [':peace_symbol:', ':slightly_smiling_face:', ':red_heart:'], 'flag': True}
Verify before relying
- Whether the GPL license is version 2, 3, or unspecified, and what that means for derivative works or commercial use.
- Current compatibility with modern Python versions and whether Python 3.X support extends to recent releases.
- Whether the multiprocessing bulk functions actually deliver measurable speedup on typical datasets.
Package facts
| License | GNU GENERAL PUBLIC LICENSE unclear |
| Python support | Not specified |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | None |
| Maintenance | Dormant 1,838 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 272,169 / month, #8,210 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
Evidence: emot-3.1-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “emoji extraction from text”
- emotExtracts emojis and emoticons from text, returning their positions,…
- demojidemoji finds and removes emojis from text, mapping each emoji to its…
- emojiConverts between emoji characters and their text code names (e.g.,…
Give your agent the search over MCP, or paste the wish link into any chat.
More Linguistic packages
Detects and normalizes text encoding from unknown or ambiguous sources, supporting all IANA character sets that Python's core library provides codecs for, with the ability to register custom codecs.
tiktoken is a fast BPE tokenizer that converts text into token sequences compatible with OpenAI models, supporting multiple encoding schemes including o200k_base and model-specific encodings.
Install it if you work with OpenAI APIs or need to understand token boundaries in GPT-family models.
Detects character encoding and language in byte sequences with high accuracy, supporting 99 encodings and returning confidence scores, language tags, and MIME types.
Install it if you need to detect character encoding or language in byte data; the rewrite makes it substantially faster and more accurate than its predecessors.
Converts Unicode text to ASCII by transliterating non-ASCII characters into their closest ASCII equivalents, with no runtime dependencies.
However, if transliteration quality or ongoing maintenance matters, consider unidecode instead despite its GPL-only license.
Lark is a parsing library that builds abstract syntax trees from context-free grammars, supporting multiple parsing algorithms (Earley, LALR(1), CYK) with automatic line and column tracking.
Python bindings to the tree-sitter parsing library, enabling incremental parsing and syntax tree analysis for source code.
See also demoji · emoji · sphinxemoji · pytest-emoji · tensorflow-text · textract · ecoji · emoji-country-flag · parse