datapackage
Utilities to work with Data Packages as defined on specs.frictionlessdata.io
Decision gist · record as of 2026-08-14
Yes, if you are working within the Frictionless Data ecosystem or need to read/write standardized Data Package descriptors. The low install friction and permissive license make it accessible. However, dormant maintenance (885 days since last release) and outdated Python version classifiers (2.7–3.7) are concerns for new projects; evaluate the Frictionless Framework as a potential alternative if you need active support and modern Python compatibility.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- The package's Python version support (2.7–3.7 per classifiers) may not align with modern Python environments; verify compatibility with your target Python version.
- Low install friction with a pure-Python wheel distribution.
- Maintenance is dormant—last release was 885 days ago—though the repository remains active.
License · maintenance · safety
MIT (permissive) — MIT license permits commercial and private use with minimal restrictions, requiring only attribution and inclusion of the license text.
last release 2024-03-12 (885 days) · last repo commit 2024-03-12 · 192 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 129,064 downloads/mo, #11,685 on PyPI
Alternatives
Verify before relying
pip install datapackage
from datapackage import Package
package = Package('datapackage.json')
resource = package.get_resource('resource')
data = resource.read(keyed=True)- Whether the package works reliably with Python versions beyond 3.7 despite classifier declarations.
- Current maintenance status and whether the Frictionless Framework migration path is recommended for new projects.
- Performance characteristics when handling large datasets or complex schemas.
What it is and what it does
Datapackage-py is a library for working with Data Packages, a specification from Frictionless Data that standardizes how tabular and non-tabular datasets are packaged with metadata. It provides a Package class for managing data package descriptors, a Resource class for reading and iterating over individual datasets, and utilities for validating descriptors against profiles and inferring metadata from raw data files. The library integrates with Table Schema and handles CSV, JSON, and other tabular formats through its runtime dependencies including tableschema and dataflows-tabulator.
The package is designed for data practitioners who need to create reproducible, machine-readable dataset descriptions—common in open data initiatives, data publishing workflows, and data integration pipelines. It lets you load existing data packages from local or remote sources, inspect and modify their descriptors, validate them against the Frictionless specification, and save them as portable archives. However, the project is in dormant maintenance (last release 885 days ago), and the maintainers have released a successor framework; new projects should evaluate whether the Frictionless Framework is a better fit.
Use it for
- Create standardized metadata for CSV or tabular datasets to publish as open data with machine-readable schemas.
- Infer and validate data package descriptors from raw files to enforce data quality and consistency.
- Load and read tabular resources from a data package descriptor with automatic type casting and missing-value handling.
- Integrate tabular datasets into data pipelines by using Package and Resource classes to handle descriptor-driven data loading.
- Validate foreign key relationships and data integrity constraints defined in a data package descriptor.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you are working within the Frictionless Data ecosystem or need to read/write standardized Data Package descriptors.
The low install friction and permissive license make it accessible. However, dormant maintenance (885 days since last release) and outdated Python version classifiers (2.7–3.7) are concerns for new projects; evaluate the Frictionless Framework as a potential alternative if you need active support and modern Python compatibility.
Install
datapackage on PyPI
Before you install
Low install friction with a pure-Python wheel distribution. Maintenance is dormant—last release was 885 days ago—though the repository remains active. The package is in Beta status and supports Python 2.7 through 3.7 per classifiers, which may limit compatibility with modern Python versions.
The package's Python version support (2.7–3.7 per classifiers) may not align with modern Python environments; verify compatibility with your target Python version.
License in practice
MIT license permits commercial and private use with minimal restrictions, requiring only attribution and inclusion of the license text.
Quickstart
pip install datapackage
from datapackage import Package
package = Package('datapackage.json')
resource = package.get_resource('resource')
data = resource.read(keyed=True)
Verify before relying
- Whether the package works reliably with Python versions beyond 3.7 despite classifier declarations.
- Current maintenance status and whether the Frictionless Framework migration path is recommended for new projects.
- Performance characteristics when handling large datasets or complex schemas.
Package facts
| License | MIT permissive |
| Python support | Not specified |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 9 packagessixclickchardetrequestsjsonschemaunicodecsvjsonpointertableschemadataflows-tabulator |
| Maintenance | Dormant 885 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 129,064 / month, #11,685 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 4 - BetaIntended Audience :: DevelopersIntended Audience :: Information TechnologyLicense :: OSI Approved :: MIT LicenseProgramming Language :: Python :: 2.7Programming Language :: Python :: 3.4Programming Language :: Python :: 3.5Programming Language :: Python :: 3.6Programming Language :: Python :: 3.7Topic :: Utilities |
Evidence: datapackage-1.15.4-py2.py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “frictionless data utilities”
- datapackageProvides classes and utilities to create, load, validate, and…
- tableschema-to-templateConverts a Frictionless Table Schema into an Excel template with…
- tableschemaValidates, infers, and works with tabular data using the Table Schema…
Give your agent the search over MCP, or paste the wish link into any chat.
More Utilities packages
Converts domain names between Unicode and ASCII-compatible encoding (Punycode) according to IDNA 2008 and Unicode Technical Standard 46, with security validation and broader script coverage than the standard library.
Install it if you work with internationalized domain names, need to validate domains, or use HTTP clients that depend on it transitively.
Detects and normalizes text encoding from unknown or ambiguous sources, supporting all IANA character sets that Python's core library provides codecs for, with the ability to register custom codecs.
Setuptools is a Python build backend and package management tool that handles building, distributing, and installing Python packages, including support for C/C++ extension modules.
Pluggy provides a plugin system that lets you define hook specifications and register implementations to be called in sequence, enabling extensible Python applications without tight coupling.
Install it if you're building an extensible application or framework.
Pygments is a syntax highlighter that colorizes source code and text in over 500 languages and formats, outputting to HTML, LaTeX, RTF, SVG, images, or ANSI terminal sequences.
Install it if you need to display or transform source code.
Six provides utility functions to write Python code that runs on both Python 2.7 and Python 3.3+, smoothing over language differences between the two versions.
See also tableschema · frictionless · tabulator · feu · csvw · mltable · dataflows-tabulator · daff · tableschema-to-template · tablib