Packages
k-diffusion is a PyTorch library implementing diffusion-based generative models from Karras et al. (2022), with improved sampling algorithms, transformer-based model support, and compatibility wrappers for other diffusion frameworks.
K-means clustering with enforced minimum and maximum cluster sizes, using a minimum cost flow algorithm to solve the constrained assignment step.
Not recommended if you need vanilla k-means performance on large datasets or if cluster size constraints are not a hard requirement.
k5test sets up isolated Kerberos 5 test environments and provides test case classes and decorators to run Python unit tests within those environments without affecting system Kerberos configuration.
However, note that maintenance is dormant (last release 877 days ago); if you need active support or compatibility with very recent Kerberos versions, evaluate…
Provides Python type models for Kubernetes resources, generated from OpenAPI specifications, enabling type-safe interaction with Kubernetes APIs.
Not recommended if you need active maintenance or support for the latest Kubernetes features.
A Python client library for creating and managing isolated sandbox environments on Kubernetes clusters, supporting multiple connection modes from local development to production cloud deployments.
Install it if you are building agents or orchestration tools on Kubernetes and need reliable, scalable sandbox lifecycle management.
A pure-Python client library for Apache Kafka that provides high-level producer and consumer APIs designed to mirror the official Java client, with support for consumer groups, message serialization, and compression.
A command-line tool and Python client for managing Apache Kafka Connect clusters and connectors via the Kafka Connect REST API, supporting connector creation, configuration, status inspection, and lifecycle operations.
However, the aging maintenance status warrants caution for production use—verify that the package still covers your Kafka Connect version's API surface and consider…
Pure-Python Apache Kafka client library providing high-level consumer, producer, and admin APIs, plus CLI tools for interactive Kafka operations without external compiled dependencies.
A Python client library for Apache Kafka that provides high-level producer and consumer APIs for publishing and consuming messages from Kafka brokers, with support for consumer groups, compression, and serialization.
However, the repository is archived and abandoned as of 2024-10-28, with the description itself recommending use of the GitHub repository due to release issues.
Integrates Kafka with Confluent Schema Registry to manage topic creation, schema publishing, and message serialization in a single workflow.
However, be aware that the package has not been updated since mid-2022; if you depend on recent Kafka or Schema Registry features, or if you need active maintenance…
Command-line interface to interact with Kaggle—download competition data, manage datasets and models, submit to competitions, and browse forums.
Kaggle Environments provides a framework for creating and running multi-agent game simulations, where agents compete or cooperate in configurable environments like Connect X and Tic Tac Toe.
However, verify the license terms first since they are not clearly stated in the package metadata.
kagglehub provides Python access to Kaggle datasets, models, and notebook outputs with integrated caching and native support for the Kaggle notebook environment.
Install it if you regularly work with Kaggle datasets or models and want to automate downloads in Python code.
Kagglesdk provides Python bindings to Kaggle's external-facing APIs, allowing programmatic access to Kaggle's services through automatically-generated client code.
A Rust-compiled workflow orchestration engine that executes the same API as the open-source kailash library, optimized for performance and memory efficiency.
Kaitai Struct provides a Python runtime library for parsing binary data structures described in a declarative YAML-based format, enabling you to read and unpack binary file formats and network packets without writing custom parsing code.
Extracts Kaldi-compatible filterbank (fbank) audio features from waveforms in real-time without external dependencies, supporting online processing across multiple architectures and operating systems.
However, if your use case is offline batch processing or you already have Kaldi or another audio library integrated, the aging maintenance status (309 days since last…
Reads and writes Kaldi binary archives, alignment files, and neural network training examples in Python, providing sequential and random-access interfaces to Kaldi's data formats.
Computes edit distance, alignment, and word error rate (WER) between sequences using Kaldi's original algorithms, with support for compound word matching and statistical confidence intervals.
Python wrapper for OpenFst and Kaldi extensions, providing finite-state transducer operations like fstdeterminizestar without requiring Kaldi installation.
Kaldiio reads and writes Kaldi archive (ark) and script (scp) files in pure Python, handling matrices, vectors, and audio data with support for binary, text, and compressed formats.
However, verify the license status before use in proprietary contexts, and note that maintenance is aging—critical bug fixes may be slow.
Kaleido generates static images (PNG, SVG, PDF, etc.) from figures, primarily for use with Plotly.py. It requires Chrome to be installed on the system.
Install it if you need to export visualizations to static formats.
Sends event and metric data to Kanaries' data infrastructure for tracking and analysis via HTTP requests with automatic retry logic.
However, the package is dormant (latest release 2024-05-13), so verify that the Kanaries endpoint is still operational and that you are comfortable with no active…
A Python client library that wraps the Kanboard JSON-RPC API, allowing you to programmatically manage projects, tasks, and other Kanboard resources from Python code.
Install it if you need to automate Kanboard workflows or integrate Kanboard into a larger Python application.
Manages and distributes plugins for exteraGram and AyuGram Telegram clients through a centralized plugin store.
Converts between Japanese kanji number representations and integers, supporting numbers up to 10^72 - 1, with configurable output styles and daiji (formal kanji) variants.
Kantoku runs and watches multiple processes and sockets, providing process lifecycle management and monitoring through a library or command-line interface.
However, the aging maintenance status warrants checking whether the project meets your long-term support expectations before adopting it for new critical systems.
Kappa is a command-line tool that automates the deployment, configuration, and testing of AWS Lambda functions, handling IAM policy creation, code packaging, and CloudWatch log retrieval.
Provides fast encryption and decryption functions for JSON and other data, with a 4-byte header prepended to encrypted output.
However, verify the underlying algorithm meets your security requirements and audit the 4-byte header behavior before using in production.
Kazoo provides a higher-level Python client API for Apache ZooKeeper, simplifying coordination, configuration management, and distributed locking operations.
Install it if you're building on ZooKeeper for coordination; skip it only if you don't need ZooKeeper or prefer a different coordination backend.
Python client for the Keboola Storage API that provides methods to interact with buckets, tables, and workspaces—exporting table data to files, creating tables, and listing bucket contents.
However, note that maintenance is aging and the client does not yet cover the entire API surface, so verify that your required endpoints are implemented before…
kcidb-io validates and manipulates Linux Kernel CI reports in JSON format, providing schema-based validation and report creation for kernel testing data.
kcli is a command-line provisioning and management tool for virtual machines across multiple virtualization providers (libvirt, KubeVirt, oVirt, OpenStack, VMware vSphere, AWS, Azure, GCP, IBM Cloud, Hcloud) using YAML plan files and cloud images.
However, verify Python version compatibility and system-level libvirt dependencies before installing, and clarify the license treatment if you need legal certainty…
Kconfiglib is a Python implementation of the Kconfig configuration language used in Linux kernel builds and embedded systems, providing both a library API and command-line tools for parsing, generating, and interactively editing Kconfig files.
Constructs, modifies, and searches kd-trees—spatial data structures for organizing points in multi-dimensional space to enable efficient nearest-neighbor queries.
However, the last release was 2017-10-19 and maintenance is aging; for production use at scale or with modern Python versions, verify compatibility and consider…
Wraps the Keboola Common Interface to simplify Docker component development, handling configuration loading, I/O mapping, manifest processing, and logging within the Keboola Connection environment.
Install it if you are building Keboola components or transformations in Python and want to avoid writing Common Interface boilerplate.
Records, sanitizes, and validates HTTP interactions for Keboola component testing by capturing real API responses as JSON cassettes and redacting secrets before storage or component consumption.
Kedro is a Python framework for building production-ready data engineering and data science pipelines with automatic dependency resolution, data catalog management, and built-in support for reproducibility and modularity.
Install it if you need structure and best practices for data engineering or data science projects beyond notebooks.
Kedro-Datasets provides data connectors for Kedro's DataCatalog, implementing AbstractDataset for formats like CSV, Excel, Parquet, JSON, SQL, and Spark DataFrames across local, network, and cloud storage.
Kedro-Telemetry is a plugin that collects anonymized usage analytics from Kedro projects and sends them to Heap Analytics to help the Kedro team understand how the framework is used.