pandas-schema
A validation library for Pandas data frames using user-friendly schemas
Decision gist · record as of 2026-08-14
No. The package is abandoned (last commit 2023-03-24, no activity for over 1638 days) and will receive no bug fixes, security updates, or compatibility patches. While it has low install friction and a permissive license, the lack of maintenance makes it a liability for production use. Consider a maintained alternative for new projects.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Low install friction with three stable runtime dependencies (numpy, pandas, packaging).
- However, the package is archived and abandoned as of 2023-03-24, with no maintenance for over 1638 days.
- Use only if you can accept no future updates or bug fixes.
License · maintenance · safety
MIT (permissive) — MIT license is permissive and imposes no restrictions on use, modification, or distribution in your own projects.
last release 2022-02-18 (1638 days) · last repo commit 2023-03-24 · 192 stars · archived
0 known vulnerabilities (OSV.dev, 2026-08-14) · 109,565 downloads/mo, #12,510 on PyPI
Alternatives
Verify before relying
import pandas as pd
from pandas_schema import Column, Schema
from pandas_schema.validation import InRangeValidation, InListValidation
schema = Schema([
Column('Age', [InRangeValidation(0, 120)]),
Column('Sex', [InListValidation(['Male', 'Female'])])
])
errors = schema.validate(pd.read_csv('data.csv'))- Whether the package remains compatible with current pandas and numpy versions despite abandonment
- Whether validation performance scales acceptably for large datasets
- Active community forks or maintained alternatives that may have superseded this project
What it is and what it does
PandasSchema provides a declarative way to validate tabular data loaded into pandas DataFrames. You define a schema by specifying columns and attaching validation rules—such as range checks, pattern matching, whitespace detection, type coercion, and membership in allowed lists—then call validate() to get a list of all data quality errors found, including row and column location.
The package is built on top of pandas and numpy, making validation fast for CSV, TSV, and other tabular formats. It's designed for data pipelines where you need to catch malformed or out-of-spec input before processing. The repository is now archived and unmaintained; the last release was in February 2022.
Use it for
- Validate incoming CSV files against a known schema before loading into a database or data warehouse
- Check data quality in ETL pipelines by defining column constraints and running them on each batch
- Enforce data type and range requirements on user-uploaded spreadsheets in web applications
- Detect formatting issues like leading/trailing whitespace or invalid patterns in bulk data imports
- Build automated data quality reports that list all validation failures with row and column references
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
No.
The package is abandoned (last commit 2023-03-24, no activity for over 1638 days) and will receive no bug fixes, security updates, or compatibility patches. While it has low install friction and a permissive license, the lack of maintenance makes it a liability for production use. Consider a maintained alternative for new projects.
Install
pandas-schema on PyPI
Before you install
Low install friction with three stable runtime dependencies (numpy, pandas, packaging). However, the package is archived and abandoned as of 2023-03-24, with no maintenance for over 1638 days. Use only if you can accept no future updates or bug fixes.
License in practice
MIT license is permissive and imposes no restrictions on use, modification, or distribution in your own projects.
Quickstart
import pandas as pd
from pandas_schema import Column, Schema
from pandas_schema.validation import InRangeValidation, InListValidation
schema = Schema([
Column('Age', [InRangeValidation(0, 120)]),
Column('Sex', [InListValidation(['Male', 'Female'])])
])
errors = schema.validate(pd.read_csv('data.csv'))
Verify before relying
- Whether the package remains compatible with current pandas and numpy versions despite abandonment
- Whether validation performance scales acceptably for large datasets
- Active community forks or maintained alternatives that may have superseded this project
Package facts
| License | MIT permissive |
| Python support | Not specified |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 3 packagesnumpypandaspackaging |
| Maintenance | Abandoned 1,638 days since the last release |
| Last repo commit | repository archived |
| First released | |
| Downloads | 109,565 / month, #12,510 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 5 - Production/StableIntended Audience :: DevelopersLicense :: OSI Approved :: MIT LicenseProgramming Language :: Python :: 3 :: OnlyProgramming Language :: Python :: 3.5 |
Evidence: pandas_schema-0.3.6-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “pandas column validation”
- pandas-schemaValidates pandas DataFrames against user-defined schemas, checking…
- panderaPandera provides a flexible API for validating dataframe-like objects…
- gspread-formattingAdds complete cell formatting support to gspread, enabling you to…
Give your agent the search over MCP, or paste the wish link into any chat.
More Quality Assurance packages
Coverage.py measures which lines of Python code are executed during test runs, reporting coverage percentages and identifying untested code paths.
Install it if you want to measure test completeness or enforce coverage thresholds in your project.
Ruff is a Python linter and code formatter written in Rust that combines linting, formatting, and code fixing into a single tool, replacing Flake8, Black, isort, and related utilities.
Pexpect spawns and controls interactive console applications by sending input and matching output patterns, automating tasks that would otherwise require manual interaction.
Black reformats Python source code to a consistent style by parsing entire files and rewriting them according to an opinionated, deterministic set of rules, eliminating manual formatting decisions.
pytest-xdist distributes pytest tests across multiple CPU cores or machines to speed up test execution, with the simplest usage being `pytest -n auto` to spawn workers equal to available CPUs.
Install it if your test suite takes long enough that parallelization would save meaningful time.
Validates AWS CloudFormation templates in YAML or JSON format against resource provider schemas and best practices, checking property values and configuration correctness.
Install it if you work with CloudFormation templates.
See also schema · tdda · pycsvschema · tablib · pandera · csvw · tableschema · pystac-ext-table · quinn · pytest-schema