biotite
A comprehensive library for computational molecular biology
Decision gist · record as of 2026-08-14
Yes. Biotite is production-stable (Development Status 5), actively maintained, carries no known vulnerabilities, and offers a cohesive API for common bioinformatics tasks. Medium install friction is offset by prebuilt wheels and strong community adoption (top 5000 PyPI). Install if you work with sequences or structures; the unified interface and NumPy integration reduce boilerplate significantly.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Python 3.12 or later; network access needed for database fetching via entrez module.
- Medium install friction with prebuilt wheels for Python 3.12–3.14 across macOS, Linux, and Windows.
- Active maintenance (last commit 2026-08-11, release 53 days ago) and 969 GitHub stars indicate solid community support.
License · maintenance · safety
BSD-3-Clause (permissive) — BSD-3-Clause (permissive) allows commercial and private use with minimal restrictions; attribution and license text inclusion are required.
last release 2026-06-22 (53 days) · last repo commit 2026-08-11 · 969 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 3,319,784 downloads/mo, #2,662 on PyPI
Alternatives
Verify before relying
pip install biotite
import biotite.sequence.align as align
import biotite.database.entrez as entrez
file_name = entrez.fetch_single_file(
uids=["CAC34569", "ACL82594"], file_name="sequences.fasta",
db_name="protein", ret_type="fasta"
)
matrix = align.SubstitutionMatrix.std_protein_matrix()
alignments = align.align_optimal(seq1, seq2, matrix)- Whether matplotlib is bundled or must be installed separately for visualization features.
- Performance characteristics when handling large-scale sequence datasets or structure files.
- Compatibility with third-party bioinformatics tools beyond what the description mentions.
What it is and what it does
Biotite is a comprehensive Python library for computational molecular biology that unifies common bioinformatics workflows into a single API. It handles sequence and biomolecular structure data through file I/O (reading and writing popular formats), database queries (searching and fetching from biological databases), analysis and editing, visualization, and external tool integration. The library stores most data internally as NumPy ndarray objects, enabling fast C-accelerated computation, NumPy-like indexing syntax, and direct access to underlying arrays for custom analysis.
The package targets both small analysis scripts and larger bioinformatics software projects. It depends on numpy, requests, msgpack, networkx, and biotraj at runtime, with optional matplotlib support for plotting. The library is actively maintained, supports current Python versions (3.12+), and has no known security vulnerabilities.
Use it for
- Download protein sequences from NCBI Entrez and perform sequence alignment using substitution matrices.
- Parse and analyze biomolecular structure files (PDB, mmCIF) to identify structural features like disulfide bonds.
- Identify homologous sequence regions across a protein family using sequence search and alignment tools.
- Build custom bioinformatics pipelines by combining file parsing, analysis, and visualization in a single workflow.
- Visualize sequence alignments and protein structures for publication or interactive exploration.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes.
Biotite is production-stable (Development Status 5), actively maintained, carries no known vulnerabilities, and offers a cohesive API for common bioinformatics tasks. Medium install friction is offset by prebuilt wheels and strong community adoption (top 5000 PyPI). Install if you work with sequences or structures; the unified interface and NumPy integration reduce boilerplate significantly.
Install
biotite on PyPI
Before you install
Medium install friction with prebuilt wheels for Python 3.12–3.14 across macOS, Linux, and Windows. Active maintenance (last commit 2026-08-11, release 53 days ago) and 969 GitHub stars indicate solid community support.
Requires Python 3.12 or later; network access needed for database fetching via entrez module.
License in practice
BSD-3-Clause (permissive) allows commercial and private use with minimal restrictions; attribution and license text inclusion are required.
Quickstart
pip install biotite
import biotite.sequence.align as align
import biotite.database.entrez as entrez
file_name = entrez.fetch_single_file(
uids=["CAC34569", "ACL82594"], file_name="sequences.fasta",
db_name="protein", ret_type="fasta"
)
matrix = align.SubstitutionMatrix.std_protein_matrix()
alignments = align.align_optimal(seq1, seq2, matrix)
Verify before relying
- Whether matplotlib is bundled or must be installed separately for visualization features.
- Performance characteristics when handling large-scale sequence datasets or structure files.
- Compatibility with third-party bioinformatics tools beyond what the description mentions.
Package facts
| License | BSD-3-Clause permissive |
| Python support | Supports the current Python release >=3.12 |
| Install friction | Medium. Platform-specific wheel |
| Runtime dependencies | 6 packagesnumpybiotrajrequestsmsgpacknetworkxpackaging |
| Maintenance | Actively maintained 53 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 3,319,784 / month, #2,662 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 5 - Production/StableIntended Audience :: DevelopersIntended Audience :: Science/ResearchNatural Language :: EnglishOperating System :: MacOSOperating System :: Microsoft :: WindowsOperating System :: POSIX :: LinuxProgramming Language :: Python :: 3Programming Language :: Python :: Implementation :: CPythonTopic :: Scientific/Engineering :: Bio-Informatics |
Evidence: biotite-1.7.1-cp312-cp312-macosx_11_0_arm64.whl; biotite-1.7.1-cp312-cp312-manylinux_2_28_aarch64.whl; biotite-1.7.1-cp312-cp312-manylinux_2_28_x86_64.whl; biotite-1.7.1-cp312-cp312-win_amd64.whl; biotite-1.7.1-cp313-cp313-macosx_11_0_arm64.whl; biotite-1.7.1-cp313-cp313-manylinux_2_28_aarch64.whl; biotite-1.7.1-cp313-cp313-manylinux_2_28_x86_64.whl; biotite-1.7.1-cp313-cp313-win_amd64.whl; biotite-1.7.1-cp314-cp314-macosx_11_0_arm64.whl; biotite-1.7.1-cp314-cp314-manylinux_2_28_aarch64.whl; biotite-1.7.1-cp314-cp314-manylinux_2_28_x86_64.whl; biotite-1.7.1-cp314-cp314-win_amd64.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “protein sequence alignment”
- biotiteBiotite provides a unified Python library for computational molecular…
- logomakerLogomaker creates customized sequence logos—visual representations of…
- fair-esmProvides pre-trained transformer protein language models (ESM-2,…
Give your agent the search over MCP, or paste the wish link into any chat.
More Bio-Informatics packages
NetworkX provides data structures and algorithms for creating, analyzing, and manipulating graphs and networks, supporting everything from simple undirected graphs to complex directed and weighted networks.
Biopython provides Python tools for computational molecular biology, including sequence analysis, structure parsing, database access, and phylogenetic tree manipulation.
However, verify that the custom Biopython License Agreement aligns with your project's licensing requirements before committing to it in production or proprietary work.
Client library for the Firecrawl API that scrapes, crawls, and searches the web, returning clean Markdown or structured data; also indexes research papers from PubMed, bioRxiv, medRxiv, and arXiv.
A self-balancing interval tree data structure that stores and queries overlapping or enveloped ranges, supporting point lookups, range overlaps, and range envelopment queries.
Install it if you need to store and query overlapping or enveloped ranges; the self-balancing design and rich query interface make it significantly easier than…
Albumentations applies image transformations to training data, supporting classification, segmentation, object detection, and pose estimation with a unified API for images, masks, bounding boxes, and keypoints.
Install it if you need a unified, production-grade augmentation API for computer vision tasks.
PubChemPy is a Python wrapper around the PubChem REST API that lets you search for chemical compounds by name, substructure, or similarity, retrieve their properties, and convert between chemical file formats.
Install it if you need programmatic access to PubChem data.
See also bio · bx-python · deepbiop · fastpdb · logomaker · python-libsbml · scikit-bio · tmtools · pipebio