Subcategories
Packages
Generates time-based UUID formats (versions 6, 7, and 8) optimized for use as database keys, extending Python's standard UUID class with RFC 9562 compliant functions.
Install it if you generate UUIDs for database keys and want better index locality than random UUIDs provide.
Gremlin-Python is a Python client for Apache TinkerPop that lets you write graph traversals to query and manipulate property graphs on remote TinkerPop-enabled servers using a functional, chainable syntax.
JayDeBeApi bridges Python to Java JDBC database drivers, providing a DB-API v2.0-compliant interface for accessing databases from cPython (via JPype1) or Jython.
However, dormant maintenance (no updates since 2020) and the requirement for a properly configured Java environment and JPype1 are significant constraints.
A native TCP driver for ClickHouse that executes queries and returns results, supporting both a direct client interface and Python DB API 2.0 specification.
aiomysql provides async/await access to MySQL databases by wrapping PyMySQL with asyncio support, letting you query MySQL non-blockingly from async Python code.
Install it if you need non-blocking MySQL access in an asyncio application; the single dependency (PyMySQL) and pure-Python wheel keep setup simple.
Apache Avro is a data serialization system that encodes structured data into a compact binary format for efficient storage and transmission over networks.
Install it if you need compact binary encoding, schema-driven serialization, or cross-language data interchange.
fastparquet reads and writes Apache Parquet files in Python, offering a native implementation that integrates with pandas, numpy, and other data processing libraries.
PickleShare provides a dictionary-like datastore where multiple processes can read and write concurrently by storing each key-value pair in a separate file within a directory.
However, the dormant maintenance status (no releases since 2018) means you should verify compatibility with your Python version and filesystem before relying on it…
Official Python driver for connecting to and querying Neo4j graph databases using the Bolt protocol, with support for both synchronous and asynchronous operations.
Install it if you need to connect to a Neo4j database; it is the standard choice for that task.
Provides a Python client for reading, writing, and querying data in Azure Tables (a NoSQL storage service accessible via HTTP/HTTPS), supporting both Azure Storage and Cosmos DB accounts.
Install it if you need to read, write, or query data in Azure Storage or Cosmos DB tables from Python.
Provides programmatic management of Azure Cosmos DB resources—databases, containers, clusters, and accounts—via the Azure control plane.
Provides type stubs for PyMySQL to enable static type checking of code using the PyMySQL database driver.
Chroma is a vector database and search infrastructure that stores, indexes, and queries document embeddings with optional metadata filtering and full-text search capabilities.
However, 28 runtime dependencies create medium install friction, and two known vulnerabilities (GHSA-f4j7-r4q5-qw2c, PYSEC-2026-311) require review before production…
Generates marshmallow schemas from SQLAlchemy models to serialize and deserialize database objects to and from JSON or other formats.
Install it if you're building APIs or services that need to serialize SQLAlchemy models.
Runs a Redis-compatible in-process database written in Rust, providing an async Python interface as a drop-in replacement for redis.asyncio.Redis without needing an external server.
Generates Microsoft Excel 97–2003 XLS spreadsheet files from Python code with no external dependencies.
impyla is a Python DB API 2.0-compliant client for querying HiveServer2 implementations like Impala and Hive, supporting Kerberos, LDAP, SSL, and JWT authentication.
Install it if you need to query Impala or Hive from Python.
Provides type annotations and IDE autocompletion for boto3 DynamoDB operations, enabling static type checking with mypy, pyright, and other tools.
Install it if you use boto3 DynamoDB and want IDE autocompletion or static type checking.
Motor provides a non-blocking MongoDB driver for asyncio and Tornado applications, presenting a coroutine-based API for concurrent database access without blocking the event loop.
However, it is scheduled for deprecation on May 14th, 2026, with MongoDB strongly recommending migration to the PyMongo Async driver.
Provides an abstract interface and plugin registry for Lance namespace implementations, enabling connection to different storage backends and custom namespace providers.
Install it if you are building or using Lance applications that need pluggable namespace support or custom storage backends.
Provides Python client library for Azure Synapse Artifacts, enabling programmatic access to create, manage, and interact with Synapse workspace artifacts including pipelines, datasets, and linked services.
Install it if you need programmatic control over Synapse workspace artifacts; skip it if you only use the Azure Portal or CLI for artifact management.
Ingest data into Azure Kusto clusters from files and blob storage using Python, with support for multiple data formats and queued ingestion.
Install it if you need to load data into Azure Kusto clusters from Python; it's the standard choice for that task.
dbt-snowflake is the Snowflake adapter for dbt, enabling data transformation and analytics workflows in Snowflake using dbt's SQL and YAML-based modeling framework.
Install it if you are using Snowflake and want to adopt dbt's SQL-based transformation and testing framework.
Rtree wraps libspatialindex to provide spatial indexing for Python, enabling nearest-neighbor search, intersection queries, and multi-dimensional spatial data structures.
dbt-databricks is a dbt adapter that enables data transformation workflows on Databricks using SQL and Python, with native support for Delta tables and Unity Catalog.
Install it if you are building dbt workflows on Databricks.
Provides a Python interface to a Rust-based SQL parser for extracting data lineage information compatible with the OpenLineage standard.
However, verify the license terms first (currently unclear in metadata) and confirm it integrates with your broader lineage platform—it is a library component, not a…
Provides a high-level Python API for building and executing Elasticsearch queries using a DSL instead of raw JSON, plus optional object-relational mapping for documents.
Provides asyncio-compatible caching with pluggable backends (memory, Redis, memcached) and a uniform interface for get, set, delete, and other standard cache operations.
Provides type hints for psycopg2, enabling static type checkers like mypy and pyright to validate code that uses the PostgreSQL adapter.
Provides Python bindings to manage Azure Redis Cache resources—create, update, delete, and configure Redis instances through the Azure management API.
Install it if you need to manage Azure Redis resources programmatically; it is the standard tool for that task.
Python client library for querying and writing data to InfluxDB 2.x using Flux language and Line Protocol, with support for Pandas DataFrames and reactive streams.
Install it if you are using InfluxDB 2.x or InfluxDB Cloud; if you are on InfluxDB 1.x or 3.x, use the version-specific client instead.
Pinecone Python SDK provides a client for creating and managing vector database indexes, upserting and querying vectors, and running inference operations against the Pinecone vector database service.
Provides programmatic management of Azure SQL resources—databases, servers, elastic pools, and related infrastructure—through the Azure SDK for Python.
Install it if you need to programmatically manage Azure SQL resources; skip it if you only use the Azure Portal or CLI for SQL operations.
Provides IAM-based authentication and encrypted connections to Google Cloud SQL instances from Python, handling TLS encryption and identity verification without requiring SSL certificates.
Install it if you are running Python on Google Cloud and need Cloud SQL connectivity.
Mongomock provides an in-memory mock of MongoDB collections for testing Python code that uses pymongo, without requiring a running MongoDB instance.
Install it if you test pymongo-dependent code.
Provides a Postgres-backed checkpoint saver for LangGraph, enabling durable state persistence for long-running workflows and agents.
Provides programmatic management of Azure relational database services (MySQL and PostgreSQL) through the Azure SDK, enabling creation, configuration, and administration of database servers via Python.
dbt-postgres is a dbt adapter that connects dbt to PostgreSQL databases, enabling data transformation workflows using SQL and dbt's templating and testing framework.
Install it if you are building or maintaining data transformation pipelines in Postgres and want version control, testing, and documentation for your SQL logic.
TinyDB is a lightweight, pure-Python document database that stores JSON-like dictionaries to local files, with no external dependencies or server required.
Provides probabilistic data structures (MinHash, HyperLogLog, and related indexes) for fast similarity estimation and cardinality counting on large datasets with minimal memory overhead.