fuzzysearch
fuzzysearch is useful for finding approximate subsequence matches
Decision gist · record as of 2026-08-14
Yes, if you need to search for approximate substring matches in long text or binary data. The library is well-maintained, has no known vulnerabilities, and its single dependency and pure-Python fallback make installation reliable. It is not suitable if you need full-text indexing or string-pair similarity comparison—use it for ad-hoc fuzzy substring search only.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Medium install friction due to optional C and Cython extensions, but pure-Python fallbacks ensure installation always succeeds.
- Last release was 276 days ago; repository is active and not archived, though maintenance status is aging.
License · maintenance · safety
MIT (permissive) — MIT license is permissive, allowing commercial and private use with minimal restrictions—suitable for most projects.
last release 2025-11-11 (276 days) · last repo commit 2025-11-11 · 342 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 770,270 downloads/mo, #5,105 on PyPI
Alternatives
Verify before relying
pip install fuzzysearch
from fuzzysearch import find_near_matches
result = find_near_matches('PATTERN', '---PATERN---', max_l_dist=1)
print(result) # [Match(start=3, end=9, dist=1, matched="PATERN")]- Performance characteristics and speed improvements from C/Cython extensions versus pure-Python fallback are not quantified in the fact sheet.
- Whether the package handles very large files or sequences efficiently is not documented in the excerpt.
What it is and what it does
fuzzysearch is a specialized library for finding approximate substring matches within longer text or binary data, allowing for a configurable number of character insertions, deletions, substitutions, or a maximum Levenshtein distance. Unlike string-comparison libraries that measure overall similarity between two strings, fuzzysearch searches through a haystack for fuzzy matches to a needle pattern, making it suited for ad-hoc searching rather than indexed full-text retrieval.
The package provides two main functions: `find_near_matches()` for in-memory data and `find_near_matches_in_file()` for file-based searches. It automatically selects the fastest algorithm based on your matching parameters, includes optional C and Cython optimizations for performance, and falls back to pure-Python implementations if compilation fails. It has a single runtime dependency (attrs) and supports Python 3.8+ as well as PyPy.
Use it for
- Search for DNA or protein sequences with tolerance for mutations or sequencing errors in genomic data.
- Find typo-tolerant pattern matches in log files or large text corpora without building an index.
- Locate approximate keyword matches in user input or search queries with character-level flexibility.
- Detect near-duplicate or slightly corrupted records in data processing pipelines.
- Search binary data for patterns with known bit-level variations or corruption.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need to search for approximate substring matches in long text or binary data.
The library is well-maintained, has no known vulnerabilities, and its single dependency and pure-Python fallback make installation reliable. It is not suitable if you need full-text indexing or string-pair similarity comparison—use it for ad-hoc fuzzy substring search only.
Install
fuzzysearch on PyPI
Before you install
Medium install friction due to optional C and Cython extensions, but pure-Python fallbacks ensure installation always succeeds. Last release was 276 days ago; repository is active and not archived, though maintenance status is aging.
License in practice
MIT license is permissive, allowing commercial and private use with minimal restrictions—suitable for most projects.
Quickstart
pip install fuzzysearch
from fuzzysearch import find_near_matches
result = find_near_matches('PATTERN', '---PATERN---', max_l_dist=1)
print(result) # [Match(start=3, end=9, dist=1, matched="PATERN")]
Verify before relying
- Performance characteristics and speed improvements from C/Cython extensions versus pure-Python fallback are not quantified in the fact sheet.
- Whether the package handles very large files or sequences efficiently is not documented in the excerpt.
Package facts
| License | MIT permissive |
| Python support | Not specified |
| Install friction | Medium. Platform-specific wheel |
| Runtime dependencies | 1 packageattrs |
| Maintenance | Aging 276 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 770,270 / month, #5,105 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 4 - BetaIntended Audience :: DevelopersLicense :: OSI Approved :: MIT LicenseNatural Language :: EnglishProgramming Language :: Python :: 3Programming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.12Programming Language :: Python :: 3.13Programming Language :: Python :: 3.14Programming Language :: Python :: 3.8Programming Language :: Python :: 3.9Programming Language :: Python :: Implementation :: CPythonProgramming Language :: Python :: Implementation :: PyPyTopic :: Software Development :: Libraries :: Python Modules |
Evidence: fuzzysearch-0.8.1-cp310-cp310-macosx_10_9_x86_64.whl; fuzzysearch-0.8.1-cp310-cp310-macosx_11_0_arm64.whl; fuzzysearch-0.8.1-cp310-cp310-manylinux_2_5_i686.manylinux1_i686.manylinux_2_17_i686.manylinux2014_i686.whl; fuzzysearch-0.8.1-cp310-cp310-manylinux_2_5_x86_64.manylinux1_x86_64.manylinux_2_17_x86_64.manylinux2014_x86_64.whl; fuzzysearch-0.8.1-cp310-cp310-musllinux_1_2_i686.whl; fuzzysearch-0.8.1-cp310-cp310-musllinux_1_2_x86_64.whl; fuzzysearch-0.8.1-cp310-cp310-win32.whl; fuzzysearch-0.8.1-cp310-cp310-win_amd64.whl; fuzzysearch-0.8.1-cp311-cp311-macosx_10_9_x86_64.whl; fuzzysearch-0.8.1-cp311-cp311-macosx_11_0_arm64.whl; fuzzysearch-0.8.1-cp311-cp311-manylinux_2_5_i686.manylinux1_i686.manylinux_2_17_i686.manylinux2014_i686.whl; fuzzysearch-0.8.1-cp311-cp311-manylinux_2_5_x86_64.manylinux1_x86_64.manylinux_2_17_x86_64.manylinux2014_x86_64.whl; fuzzysearch-0.8.1-cp311-cp311-musllinux_1_2_i686.whl; fuzzysearch-0.8.1-cp311-cp311-musllinux_1_2_x86_64.whl; fuzzysearch-0.8.1-cp311-cp311-win32.whl; fuzzysearch-0.8.1-cp311-cp311-win_amd64.whl; fuzzysearch-0.8.1-cp312-cp312-macosx_10_9_x86_64.whl; fuzzysearch-0.8.1-cp312-cp312-macosx_11_0_arm64.whl; fuzzysearch-0.8.1-cp312-cp312-manylinux_2_5_i686.manylinux1_i686.manylinux_2_17_i686.manylinux2014_i686.whl; fuzzysearch-0.8.1-cp312-cp312-manylinux_2_5_x86_64.manylinux1_x86_64.manylinux_2_17_x86_64.manylinux2014_x86_64.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “approximate pattern matching”
- fuzzysearchFinds approximate substring matches in text or data with configurable…
- fastdtwfastdtw computes approximate Dynamic Time Warping (DTW) alignments…
- fuzzyfinderFuzzy finder that matches partial strings from a list and returns…
Give your agent the search over MCP, or paste the wish link into any chat.
More Python Modules packages
Converts domain names between Unicode and ASCII-compatible encoding (Punycode) according to IDNA 2008 and Unicode Technical Standard 46, with security validation and broader script coverage than the standard library.
Install it if you work with internationalized domain names, need to validate domains, or use HTTP clients that depend on it transitively.
Setuptools is a Python build backend and package management tool that handles building, distributing, and installing Python packages, including support for C/C++ extension modules.
PyYAML parses and emits YAML 1.1 data format, enabling serialization and deserialization of configuration files and Python objects to and from human-readable YAML text.
Pydantic validates Python data structures against type hints, coercing and checking input at runtime to ensure it matches a declared schema.
Provides reusable metadata objects for use with PEP-593 `typing.Annotated` to express common constraints like bounds, collection sizes, and predicates on types.
Install it if you use or build libraries that need to express type constraints in a standardized, inspectable way—or if you want to annotate your own types with…
Provides runtime tools to inspect and introspect Python type annotations, enabling programmatic examination of type hints at execution time.
See also fuzzyset2 · thefuzz · ngram · strsimpy · fuzzywuzzy · textdistance · python-Levenshtein · fuzzyfinder · RapidFuzz · tfidf-matcher