awkward-pandas
Awkward Array Pandas Extension
Decision gist · record as of 2026-08-14
Yes, if you need to work with nested or ragged data in pandas. The low install friction, permissive license, and lack of known vulnerabilities make it safe to try. However, the aging maintenance status (no release in over a year) means you should verify compatibility with your specific pandas and Python versions before relying on it in production.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires pandas and awkward as runtime dependencies; Python >= 3.8.
- Low install friction with a pure-Python wheel.
- Maintenance is aging—last release was over a year ago (2023-08-08) and the last commit is recent (2025-09-17), suggesting the project is maintained but not actively developed.
License · maintenance · safety
permissive license (permissive) — BSD License (permissive) allows use in commercial and private projects with minimal restrictions, requiring only license attribution.
last release 2023-08-08 (1102 days) · last repo commit 2025-09-17 · 52 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 363,766 downloads/mo, #7,216 on PyPI
Alternatives
Verify before relying
pip install awkward-pandas
import pandas as pd
import awkward_pandas
df = pd.DataFrame({'col': []})
df['col'] = df['col'].astype('awkward')- Whether the aging maintenance status (1102 days since release) affects compatibility with recent pandas versions.
- Performance characteristics when working with very large nested datasets compared to alternative approaches.
- Real-world usage patterns and adoption beyond the 52 GitHub stars.
What it is and what it does
awkward-pandas is a pandas extension that allows you to store Awkward Array objects as columns in DataFrames. Awkward Array is a library for working with nested, ragged, and complex data structures efficiently—the kind of data that doesn't fit neatly into a rectangular table. This extension bridges the two libraries, letting you use pandas' familiar API and tools while leveraging Awkward Array's capabilities for handling irregular data.
The package is in alpha status and has not seen a release since August 2023, though the repository remains active. It supports Python 3.8 through 3.11 and depends on both awkward and pandas. Use it when you need to work with nested or ragged data but want to stay within the pandas ecosystem rather than switching entirely to Awkward Array.
Use it for
- Store lists of varying lengths as DataFrame columns without flattening or padding the data.
- Work with hierarchical or nested JSON-like data structures directly in pandas operations.
- Combine ragged array data with tabular data in a single DataFrame for analysis.
- Leverage pandas' groupby and aggregation operations on data containing nested arrays.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need to work with nested or ragged data in pandas.
The low install friction, permissive license, and lack of known vulnerabilities make it safe to try. However, the aging maintenance status (no release in over a year) means you should verify compatibility with your specific pandas and Python versions before relying on it in production.
Install
awkward-pandas on PyPI
Before you install
Low install friction with a pure-Python wheel. Maintenance is aging—last release was over a year ago (2023-08-08) and the last commit is recent (2025-09-17), suggesting the project is maintained but not actively developed. Depends on awkward and pandas, both mature libraries.
Requires pandas and awkward as runtime dependencies; Python >= 3.8.
License in practice
BSD License (permissive) allows use in commercial and private projects with minimal restrictions, requiring only license attribution.
Quickstart
pip install awkward-pandas
import pandas as pd
import awkward_pandas
df = pd.DataFrame({'col': []})
df['col'] = df['col'].astype('awkward')
Verify before relying
- Whether the aging maintenance status (1102 days since release) affects compatibility with recent pandas versions.
- Performance characteristics when working with very large nested datasets compared to alternative approaches.
- Real-world usage patterns and adoption beyond the 52 GitHub stars.
Package facts
| License | permissive license permissive |
| Python support | Supports the current Python release >=3.8 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 2 packagesawkwardpandas |
| Maintenance | Aging 1,102 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 363,766 / month, #7,216 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 3 - AlphaLicense :: OSI Approved :: BSD LicenseProgramming Language :: Python :: 3Programming Language :: Python :: 3 :: OnlyProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.8Programming Language :: Python :: 3.9Topic :: Scientific/Engineering |
Evidence: awkward_pandas-2023.8.0-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “pandas awkward array extension”
- awkward-pandasExtends pandas DataFrames to store and manipulate Awkward Array…
- uprootUproot reads and writes ROOT files (the data format used in…
- dask-awkwardDask-awkward integrates Awkward Array with Dask to enable…
Give your agent the search over MCP, or paste the wish link into any chat.
More Scientific/Engineering packages
NumPy provides an N-dimensional array object and a comprehensive suite of mathematical, linear algebra, Fourier transform, and random number functions for scientific computing in Python.
pandas provides fast, flexible data structures (Series and DataFrame) for loading, cleaning, transforming, and analyzing labeled or relational data in Python.
scipy provides numerical algorithms for mathematics, science, and engineering—including optimization, integration, linear algebra, Fourier transforms, signal and image processing, and ODE solvers—built on numpy arrays.
scikit-learn provides a comprehensive Python library for supervised and unsupervised machine learning, including classification, regression, clustering, dimensionality reduction, and model evaluation tools built on NumPy and SciPy.
Install it if you need to train, evaluate, or deploy supervised or unsupervised learning models.
dill extends Python's pickle module to serialize and deserialize a much wider range of Python objects, including functions, lambdas, classes, and interpreter sessions, to byte streams for storage or network transmission.
Multiprocess is an enhanced fork of Python's standard multiprocessing library that uses dill for better serialization, allowing you to spawn processes with a threading-like API and share complex objects between them.
Install it if you use multiprocessing and encounter pickle serialization limits with lambdas or complex objects.
See also awkward · awkward0 · ipfn · newtools · dask-awkward · woodwork · pyjanitor · awkward-cpp · miceforest