copulas
Create tabular synthetic data using copulas-based modeling.
Decision gist · record as of 2026-08-14
Yes, with conditions. The package is actively maintained, has low install friction, and solves a real need for synthetic tabular data generation. However, the BUSL-1.1 license restricts commercial use until a future date—verify the license terms against your use case before committing. The pre-alpha development status suggests the API may change; for production use, pin the version and monitor releases.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Installation is straightforward with low friction; the package has four common runtime dependencies (numpy, pandas, plotly, scipy) and maintains active development with a recent release.
License · maintenance · safety
BUSL-1.1 (unclear) — The package uses BUSL-1.1 (Business Source License), which restricts commercial use until a future date and may require review before adoption in commercial projects.
last release 2026-02-05 (190 days) · last repo commit 2026-08-10 · 650 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 201,903 downloads/mo, #9,656 on PyPI
Alternatives
Verify before relying
pip install copulas
from copulas.datasets import sample_trivariate_xyz
from copulas.multivariate import GaussianMultivariate
real_data = sample_trivariate_xyz()
copula = GaussianMultivariate()
copula.fit(real_data)
synthetic_data = copula.sample(len(real_data))- Whether BUSL-1.1 license restrictions apply to your intended use case and timeline
- Performance characteristics and scalability limits for large datasets
- Specific copula types supported beyond Gaussian, Vine, and Archimedian mentioned in description
What it is and what it does
Copulas is a Python library for learning multivariate distributions from numerical data and generating synthetic data that follows the same statistical patterns. You provide a table of real data, fit a copula model (choosing from options like Gaussian Copula, Vine Copulas, or Archimedian Copulas), and then sample new synthetic records that preserve the correlations and marginal distributions of the original. The library integrates with numpy, pandas, scipy for computation and plotly for visualization, allowing you to compare real and synthetic data side-by-side in 1D, 2D, and 3D plots.
The package is part of the Synthetic Data Vault Project and targets developers and data engineers building synthetic data pipelines. It exposes the learned model parameters for inspection and tuning, making it suitable for both exploratory work and production use where you need control over the generation process. The codebase is actively maintained, supports Python 3.9 through 3.14, and has no known security vulnerabilities.
Use it for
- Generate privacy-preserving synthetic datasets for testing and development without exposing real customer or sensitive data
- Create balanced training datasets for machine learning by sampling from learned multivariate distributions
- Augment small datasets by learning their statistical structure and generating additional synthetic records
- Validate data pipelines and analytics by producing synthetic data with known statistical properties
- Compare real versus synthetic data distributions visually to verify model quality before deployment
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, with conditions.
The package is actively maintained, has low install friction, and solves a real need for synthetic tabular data generation. However, the BUSL-1.1 license restricts commercial use until a future date—verify the license terms against your use case before committing. The pre-alpha development status suggests the API may change; for production use, pin the version and monitor releases.
Install
copulas on PyPI
Before you install
Installation is straightforward with low friction; the package has four common runtime dependencies (numpy, pandas, plotly, scipy) and maintains active development with a recent release.
License in practice
The package uses BUSL-1.1 (Business Source License), which restricts commercial use until a future date and may require review before adoption in commercial projects.
Quickstart
pip install copulas
from copulas.datasets import sample_trivariate_xyz
from copulas.multivariate import GaussianMultivariate
real_data = sample_trivariate_xyz()
copula = GaussianMultivariate()
copula.fit(real_data)
synthetic_data = copula.sample(len(real_data))
Verify before relying
- Whether BUSL-1.1 license restrictions apply to your intended use case and timeline
- Performance characteristics and scalability limits for large datasets
- Specific copula types supported beyond Gaussian, Vine, and Archimedian mentioned in description
Package facts
| License | BUSL-1.1 unclear |
| Python support | Supports the current Python release <3.15,>=3.9 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 4 packagesnumpypandasplotlyscipy |
| Maintenance | Actively maintained 190 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 201,903 / month, #9,656 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 2 - Pre-AlphaIntended Audience :: DevelopersNatural Language :: EnglishProgramming Language :: Python :: 3Programming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.12Programming Language :: Python :: 3.13Programming Language :: Python :: 3.14Programming Language :: Python :: 3.9Topic :: Scientific/Engineering :: Artificial Intelligence |
Evidence: copulas-0.14.1-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “copula modeling”
- copulasCopulas models multivariate statistical distributions and generates…
- fprime-fppFPP is a modeling language and compiler for the F Prime flight…
- pvlibSimulates photovoltaic energy system performance and provides…
Give your agent the search over MCP, or paste the wish link into any chat.
More Artificial Intelligence packages
LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.
Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.
Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.
Install it if you work with Hugging Face Hub models or datasets.
LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.
hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.
Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.
Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.
Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.
See also sdv · deepecho · ctgan · hyppo · sdmetrics · data-designer · prince · statsmodels · corner · rdt