pymzml
high-throughput mzML parsing
Decision gist · record as of 2026-08-14
Yes. pymzml is actively maintained, has low install friction, carries no known vulnerabilities, and is the standard choice for mzML parsing in Python. Install it if you work with mass spectrometry data in proteomics or metabolomics; the core library is lightweight and optional features are available if needed.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Python 3.10 or higher; numpy must be installed before optional dependencies like pynumpress.
- Low friction install with only numpy and regex as core dependencies.
- Active maintenance with a recent release (52 days old) and ongoing repository activity.
License · maintenance · safety
The MIT license (permissive) — MIT license permits commercial and private use with minimal restrictions; you may use, modify, and distribute pymzml provided you include the license notice.
last release 2026-06-23 (52 days) · last repo commit 2026-07-24 · 194 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 166,281 downloads/mo, #10,501 on PyPI
Alternatives
Verify before relying
pip install pymzml
import pymzml
# Parse an mzML file and access spectra
run = pymzml.run.Reader('data.mzML')
for spectrum in run:
print(spectrum.ID, spectrum.mz, spectrum.i)- Whether the package's spectrum comparison and visualization functions work out-of-the-box or require optional dependencies.
- Performance characteristics for very large mzML files or highly compressed formats.
- Whether random access in compressed files requires specific file preparation or encoding.
What it is and what it does
pymzml is a Python parser for mzML, the standard mass spectrometry data format used in proteomics and metabolomics research. It provides fast, seekable access to spectra stored in mzML files—including support for compressed formats—and includes utilities for spectrum comparison and visualization. The core library depends only on numpy and regex, keeping installation simple, though optional extras add plotting and deconvolution capabilities.
The package is aimed at bioinformaticians and researchers who need to programmatically extract and analyze mass spectrometry data. It has been actively maintained since 2012 and currently requires Python 3.10 or higher. Installation is straightforward via pip, with optional feature flags for extended functionality like interactive plotting or pynumpress compression support.
Use it for
- Extract and iterate over spectra from mzML files in a proteomics workflow.
- Build custom mass spectrometry data analysis pipelines that need rapid spectrum access.
- Compare spectra or perform spectral matching in metabolomics studies.
- Integrate mzML parsing into bioinformatics tools for automated MS data processing.
- Visualize mass spectrometry data interactively during exploratory analysis.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes.
pymzml is actively maintained, has low install friction, carries no known vulnerabilities, and is the standard choice for mzML parsing in Python. Install it if you work with mass spectrometry data in proteomics or metabolomics; the core library is lightweight and optional features are available if needed.
Install
pymzml on PyPI
Before you install
Low friction install with only numpy and regex as core dependencies. Active maintenance with a recent release (52 days old) and ongoing repository activity. Requires Python 3.10 or higher.
Requires Python 3.10 or higher; numpy must be installed before optional dependencies like pynumpress.
License in practice
MIT license permits commercial and private use with minimal restrictions; you may use, modify, and distribute pymzml provided you include the license notice.
Quickstart
pip install pymzml
import pymzml
# Parse an mzML file and access spectra
run = pymzml.run.Reader('data.mzML')
for spectrum in run:
print(spectrum.ID, spectrum.mz, spectrum.i)
Verify before relying
- Whether the package's spectrum comparison and visualization functions work out-of-the-box or require optional dependencies.
- Performance characteristics for very large mzML files or highly compressed formats.
- Whether random access in compressed files requires specific file preparation or encoding.
Package facts
| License | The MIT license permissive |
| Python support | Supports the current Python release >=3.10.0 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 2 packagesnumpyregex |
| Maintenance | Actively maintained 52 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 166,281 / month, #10,501 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 4 - BetaEnvironment :: ConsoleIntended Audience :: DevelopersIntended Audience :: EducationIntended Audience :: Science/ResearchLicense :: OSI Approved :: MIT LicenseOperating System :: MacOS :: MacOS XOperating System :: Microsoft :: WindowsOperating System :: POSIXOperating System :: POSIX :: SunOS/SolarisOperating System :: UnixProgramming Language :: Python :: 3.5Topic :: Scientific/Engineering :: Bio-InformaticsTopic :: Scientific/Engineering :: ChemistryTopic :: Scientific/Engineering :: Medical Science Apps. |
Evidence: pymzml-2.6.1-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “mzml parser python”
- pymzmlpymzml parses mzML mass spectrometry data files in Python, providing…
- dparseParses Python dependency files (requirements.txt, conda.yml, tox.ini,…
- pyleriPyleri is a left-right parser generator that lets you define a…
Give your agent the search over MCP, or paste the wish link into any chat.
More Bio-Informatics packages
NetworkX provides data structures and algorithms for creating, analyzing, and manipulating graphs and networks, supporting everything from simple undirected graphs to complex directed and weighted networks.
Biopython provides Python tools for computational molecular biology, including sequence analysis, structure parsing, database access, and phylogenetic tree manipulation.
However, verify that the custom Biopython License Agreement aligns with your project's licensing requirements before committing to it in production or proprietary work.
Client library for the Firecrawl API that scrapes, crawls, and searches the web, returning clean Markdown or structured data; also indexes research papers from PubMed, bioRxiv, medRxiv, and arXiv.
A self-balancing interval tree data structure that stores and queries overlapping or enveloped ranges, supporting point lookups, range overlaps, and range envelopment queries.
Install it if you need to store and query overlapping or enveloped ranges; the self-balancing design and rich query interface make it significantly easier than…
Albumentations applies image transformations to training data, supporting classification, segmentation, object detection, and pose estimation with a unified API for images, masks, bounding boxes, and keypoints.
Install it if you need a unified, production-grade augmentation API for computer vision tasks.
PubChemPy is a Python wrapper around the PubChem REST API that lets you search for chemical compounds by name, substructure, or similarity, retrieve their properties, and convert between chemical file formats.
Install it if you need programmatic access to PubChem data.
See also pyteomics · alchemlyb · libhreels · mt2 · specutils · dnaio · peppy · edam-ontology