pylance
python wrapper for Lance columnar format
Decision gist · record as of 2026-08-14
Yes, if you work with large analytical datasets and want a columnar storage format integrated with the pyarrow ecosystem. The permissive Apache license, active maintenance, and zero known vulnerabilities support adoption. Medium install friction (compiled wheels, multiple dependencies) is typical for data science packages. Be aware that Alpha status means the API may change; suitable for production use only if you can tolerate potential breaking changes in future releases.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Python 3.10 or later; pyarrow and numpy must be installed (typically handled by pip automatically).
- Medium install friction due to compiled wheels for multiple platforms (arm64, x86_64, manylinux variants).
- Active maintenance with a release 7 days ago.
License · maintenance · safety
permissive license (permissive) — Licensed under Apache License 2.0 (permissive), allowing commercial and private use with minimal restrictions—suitable for most production and research contexts.
last release 2026-08-07 (7 days)
0 known vulnerabilities (OSV.dev, 2026-08-14) · 4,059,100 downloads/mo, #2,387 on PyPI
Alternatives
Verify before relying
pip install pylance
import lance
# Create a Lance dataset from a PyArrow table or pandas DataFrame
data = lance.write_table(table, uri="./my_dataset.lance")- Specific performance characteristics (latency, throughput) compared to other columnar formats.
- Whether the package supports streaming or incremental writes beyond the basic write_table API.
- Current state of documentation and examples beyond the contribution guide reference.
- Actual download volume and adoption metrics beyond position in the top PyPI packages list.
What it is and what it does
Pylance is a Python SDK that wraps the Lance columnar data format, a format designed for efficient storage and retrieval of large analytical datasets. It sits on top of Apache Arrow, inheriting Arrow's type system and columnar layout while providing a Python-friendly interface for reading and writing data. The package targets data science and machine learning workflows where columnar storage offers advantages in compression, query performance, and memory efficiency.
The package is in active development (Alpha status) and requires Python 3.10 or later with compiled wheels available for common platforms. Its main dependencies—pyarrow, numpy, and lance-namespace—are typical for data science stacks. With zero known vulnerabilities and a permissive license, it presents a low security risk, though as an Alpha project it may still undergo API changes.
Use it for
- Store large machine learning training datasets in a columnar format optimized for batch access and feature extraction.
- Build data pipelines that serialize analytical tables to disk with better compression than row-oriented formats.
- Integrate with pyarrow-based workflows to leverage Arrow's ecosystem while using Lance's storage optimizations.
- Archive and query time-series or event data where columnar layout reduces I/O and memory overhead.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you work with large analytical datasets and want a columnar storage format integrated with the pyarrow ecosystem.
The permissive Apache license, active maintenance, and zero known vulnerabilities support adoption. Medium install friction (compiled wheels, multiple dependencies) is typical for data science packages. Be aware that Alpha status means the API may change; suitable for production use only if you can tolerate potential breaking changes in future releases.
Install
pylance on PyPI
Before you install
Medium install friction due to compiled wheels for multiple platforms (arm64, x86_64, manylinux variants). Active maintenance with a release 7 days ago. Requires Python 3.10 or later and depends on pyarrow and numpy, which are themselves substantial compiled packages.
Requires Python 3.10 or later; pyarrow and numpy must be installed (typically handled by pip automatically).
License in practice
Licensed under Apache License 2.0 (permissive), allowing commercial and private use with minimal restrictions—suitable for most production and research contexts.
Quickstart
pip install pylance
import lance
# Create a Lance dataset from a PyArrow table or pandas DataFrame
data = lance.write_table(table, uri="./my_dataset.lance")
Verify before relying
- Specific performance characteristics (latency, throughput) compared to other columnar formats.
- Whether the package supports streaming or incremental writes beyond the basic write_table API.
- Current state of documentation and examples beyond the contribution guide reference.
- Actual download volume and adoption metrics beyond position in the top PyPI packages list.
Package facts
| License | permissive license permissive |
| Python support | Supports the current Python release >=3.10 |
| Install friction | Medium. Platform-specific wheel |
| Runtime dependencies | 3 packagespyarrownumpylance-namespace |
| Maintenance | Actively maintained 7 days since the last release |
| First released | |
| Downloads | 4,059,100 / month, #2,387 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 3 - AlphaEnvironment :: ConsoleIntended Audience :: Science/ResearchLicense :: OSI Approved :: Apache Software LicenseOperating System :: OS IndependentProgramming Language :: PythonProgramming Language :: Python :: 3Programming Language :: Python :: 3 :: OnlyProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.12Programming Language :: Python :: 3.13Programming Language :: Python :: 3.14Programming Language :: RustTopic :: Scientific/Engineering |
Evidence: pylance-10.0.0-cp310-abi3-macosx_11_0_arm64.whl; pylance-10.0.0-cp310-abi3-manylinux_2_17_aarch64.manylinux2014_aarch64.whl; pylance-10.0.0-cp310-abi3-manylinux_2_17_x86_64.manylinux2014_x86_64.whl; pylance-10.0.0-cp310-abi3-manylinux_2_28_aarch64.whl; pylance-10.0.0-cp310-abi3-manylinux_2_28_x86_64.whl; pylance-10.0.0-cp310-abi3-win_amd64.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “lance columnar format python”
- pylancePylance is a Python wrapper for the Lance columnar data format,…
- lance-namespace-urllib3-clientAuto-generated REST client for the Lance Namespace specification,…
- lance-namespaceProvides an abstract interface and plugin registry for Lance…
Give your agent the search over MCP, or paste the wish link into any chat.
More Scientific/Engineering packages
NumPy provides an N-dimensional array object and a comprehensive suite of mathematical, linear algebra, Fourier transform, and random number functions for scientific computing in Python.
pandas provides fast, flexible data structures (Series and DataFrame) for loading, cleaning, transforming, and analyzing labeled or relational data in Python.
scipy provides numerical algorithms for mathematics, science, and engineering—including optimization, integration, linear algebra, Fourier transforms, signal and image processing, and ODE solvers—built on numpy arrays.
scikit-learn provides a comprehensive Python library for supervised and unsupervised machine learning, including classification, regression, clustering, dimensionality reduction, and model evaluation tools built on NumPy and SciPy.
Install it if you need to train, evaluate, or deploy supervised or unsupervised learning models.
dill extends Python's pickle module to serialize and deserialize a much wider range of Python objects, including functions, lambdas, classes, and interpreter sessions, to byte streams for storage or network transmission.
Multiprocess is an enhanced fork of Python's standard multiprocessing library that uses dill for better serialization, allowing you to spawn processes with a threading-like API and share complex objects between them.
Install it if you use multiprocessing and encounter pickle serialization limits with lambdas or complex objects.
See also lance-context · lancedb · pyarrow · vortex-data · anima-python · feather-format · arrow-odbc · polars · lance-namespace-urllib3-client · geoarrow-pyarrow