Packages
Provides pandas extension data types for SQL systems like BigQuery and Spanner, enabling native representation of database-specific types within pandas DataFrames.
Install it if you regularly work with SQL data in pandas and need type fidelity.
Command-line client for interacting with Douban (豆瓣) social features including groups, people, photos, and status updates via a wrapper around Douban's web API.
Reads and writes dBase III, FoxPro, Visual FoxPro, and Clipper .dbf database files, converting between legacy DBF formats and Python data types.
However, maintenance is aging (last release 346 days ago), so verify compatibility with your specific Python version and DBF variant before committing to production use.
Reads DBF database files (dBase, Visual FoxPro, FoxBase+) and returns records as native Python data types, supporting both streaming and in-memory access.
DiscoverX automates bulk administration tasks across Lakehouse assets by executing SQL templates or Python functions concurrently against multiple Delta tables matching selection patterns.
However, verify the license terms and test compatibility with your current Databricks and Python versions before production use, and understand that support is…
Retrieves profile and metadata information about Databricks workspace resources through a Python SDK, primarily for use by security analysis tooling.
However, the aging maintenance status and unclear license terms warrant verification before production use; confirm with Databricks that the package is appropriate…
Tempo provides time series operations on Spark DataFrames, including AS OF joins, rolling statistics, lagged feature generation, and Delta Lake optimization for time-partitioned data.
Install it if you work with time-indexed data; skip it if you do not need time series transformations.
Generates synthetic data at scale within Databricks using Spark, supporting repeatable, predictable datasets for testing, benchmarking, and demos across all Spark SQL primitive types.
However, verify the 'Databricks License' terms for your use case before installing, particularly if you plan to use it outside Databricks or in proprietary applications.
DBOS adds durable workflows and queues to Python applications by checkpointing execution state in Postgres, allowing programs to automatically recover from failures without external orchestration infrastructure.
Python client for querying and managing datasets in the Data Bookkeeping Service 3 (DBS3), a CMS experiment event data catalog that tracks datasets, processing history, files, and runs via REST API.
Provides a Python client wrapper around pycurl for making HTTP requests, built on pycurl and certifi for SSL certificate handling.
However, the dormant maintenance status means you should verify that pycurl itself meets your needs and that you're comfortable relying on a package that is not…
dbstream is a meta package designed to provide unified connection and access to multiple database systems through a single interface.
A command-line interface for running dbt commands against dbt Cloud development environments, enabling data transformation workflows from your local terminal.
However, the high install friction and unclear license require upfront investigation.
Provides base adapter protocols and shared functionality that database adapters use to integrate with dbt-core, handling connections, dialect translation, relation caching, and core interface management.
Parses dbt artifact JSON files (catalog, manifest, run-results, sources) into typed Python objects using pydantic, supporting multiple schema versions across dbt 0.19 to 1.12.
dbt-athena is a dbt adapter that enables data transformation workflows to run against Amazon Athena, supporting table and incremental models, snapshots, and Python models on Hive and Iceberg table formats.
Install it if you use dbt and need to transform data in Athena.
A dbt adapter that enables data transformation workflows on AWS Athena, supporting table and incremental models, snapshots, Python models, and integration with Iceberg and Hive table formats.
Install it if you're building dbt workflows on AWS Athena; the low dependency footprint and permissive license make adoption straightforward.
dbt-autofix scans dbt projects for deprecated configurations and automatically updates them to align with current best practices and dbt Fusion compatibility.
dbt-bigquery is a dbt adapter that enables data transformation workflows in Google BigQuery, allowing analysts and engineers to organize, cleanse, and prepare raw warehouse data using dbt's SQL and YAML-based practices.
Install it if you use BigQuery and want to adopt dbt's SQL-based transformation practices; skip it if you prefer procedural data pipelines or are not yet using BigQuery.
dbt-bouncer validates dbt projects against configurable conventions and naming rules, running checks on dbt artifacts to enforce project-wide standards.
Install it if your team uses dbt and wants to standardize project structure and naming without manual code review.
A dbt adapter that enables data transformation and testing workflows on ClickHouse, supporting table, view, incremental, and materialized view materializations with ClickHouse-specific configurations.
Install it if you use ClickHouse and want to adopt dbt for data transformation.
Extracts column-level lineage from dbt projects and generates an interactive HTML dashboard showing how data flows through your transformations.
Provides shared utilities and common code used by dbt-core and dbt adapter implementations to avoid duplication across the dbt ecosystem.
Provides utility functions for running Django and Flask applications in AWS ECS via AWS Copilot, including configuration helpers for networking, databases, Celery workers, and error tracking.
However, verify the license terms first, as they are not clearly documented in the package metadata.
dbt-core transforms raw data in a warehouse using SQL select statements organized as models, handling compilation, testing, and dependency management across data transformation projects.
Install it if you need to build SQL-based data transformations with testing and documentation; you will also need to install a warehouse-specific adapter…
Experimental parser for dbt Core v2.0 (alpha), a ground-up Rust rewrite of the data transformation framework that parses, compiles, and executes dbt projects faster than v1.
No, not for production use.
Provides a lightweight Python interface to dbt-core (v1.8+) for in-memory SQL compilation, macro evaluation, and project management without leaving Python.
However, maintenance is aging (215 days since release), so verify that the version supports your dbt-core release and that data quality features meet your stability…
Measures documentation and test coverage of dbt projects by analyzing manifest and catalog files, reporting per-model coverage percentages and totals via CLI.
Install it if you want to measure and enforce documentation and test coverage in your dbt projects; the zero-config design and low friction make it a straightforward…
dbt-databricks is a dbt adapter that enables data transformation workflows on Databricks using SQL and Python, with native support for Delta tables and Unity Catalog.
Install it if you are building dbt workflows on Databricks.
dbt-dremio is a dbt adapter that enables data transformation workflows in Dremio Cloud and Dremio Software (versions 22.0 and later) using dbt's ELT practices.
Install it if you use Dremio (Cloud or Software 22.0+) and want to adopt dbt for data transformation.
dbt-duckdb connects dbt (a SQL/Python transformation framework) to DuckDB, an embedded OLAP database that can read and write CSV, JSON, and Parquet files directly without loading them first.
dbt-exasol is a dbt adapter that enables data transformation workflows on Exasol warehouses using dbt's SQL and Python capabilities.
However, be aware of platform limitations: Python models, materialized views, and native zero-copy clones are not supported.
Extracts metadata (refs, sources, configs) from Jinja templates in dbt model files without executing Python, using tree-sitter parsing and static type checking.
A dbt adapter that connects dbt-core to Microsoft Fabric Synapse Data Warehouse, enabling data transformation and modeling workflows on Fabric.
Install only if you have a Fabric warehouse and dbt-core 1.4 or newer; it is not useful as a standalone tool.
A dbt adapter that connects dbt to Apache Spark in Microsoft Fabric via Livy endpoints, enabling SQL-based data transformation workflows on Fabric Lakehouses with or without schema support.
Install it if you use dbt with Microsoft Fabric Spark and need to transform data in Lakehouses.
Upgrades dbt package dependencies to versions compatible with dbt Fusion by parsing packages.yml/dependencies.yml files, checking compatibility against a local cache of package hub data, and rewriting version strings in place.
The package is narrowly scoped to a specific dbt workflow; install only if you actually need Fusion compatibility upgrades.
dbt-glue is a dbt adapter that enables running dbt transformations against AWS Glue's Spark engine using the Glue Interactive Sessions API.
However, verify that your use case aligns with Spark's capabilities and Glue's cost model, and confirm production-readiness expectations given the Beta status.
dbt-loom is a dbt Core plugin that fetches public model definitions from upstream dbt project artifacts and injects them into your dbt project, enabling multi-project deployments to share models across repositories.
However, verify the license terms before adopting in proprietary contexts, as the license metadata is missing from the package.
Exposes dbt project metadata and operations through a Model Context Protocol server, enabling AI agents to query models, metrics, lineage, and execute dbt commands via standardized tools.
Syncs dbt project metadata (table relationships, descriptions, semantic types) to Metabase and extracts Metabase questions and dashboards back into dbt as exposures.
Install it if you use both tools and want a single source of truth for metadata.