Subcategories
Packages
Panel Material UI brings Material Design components to Panel dashboards, offering styled buttons, sliders, cards, dialogs, and other widgets with built-in theming and dark mode support.
Install it if you want Material Design styling in Panel without custom CSS work.
Loads and runs fastText language identification models with a minimal, dependency-free predict interface.
Provides a C-based API for annotating events, code ranges, and resources in applications to enable capture and visualization through NVIDIA's Visual Profiler.
Provides a Python DSL for writing high-performance CUDA kernels using CUTLASS and CuTe concepts, targeting NVIDIA Tensor Cores on Ampere, Hopper, and Blackwell architectures.
However, verify the unclear license terms for your use case, and note that the package is in beta—expect potential API changes before summer 2026.
Provides a portable intermediate representation (TileIR) for CUDA kernels that abstracts away language-specific details, enabling kernel compilation and optimization across different contexts.
Provides NVIDIA's collective communication library (NCCL) runtime for GPU-accelerated all-reduce, all-gather, reduce, broadcast, and reduce-scatter operations optimized for CUDA 11.
Qiskit is an open-source SDK for building and executing quantum circuits, operators, and primitives on quantum computers and simulators.
Install it if you're learning quantum computing, prototyping algorithms, or need to integrate with quantum hardware ecosystems.
Generates n-dimensional alpha shapes—concave or convex bounding polygons around point sets—with support for static, dynamic, or automatically-solved alpha parameters.
Install it if you need adaptive bounding shapes; skip it only if you require only convex hulls or if your use case is strictly 2D and you prefer a lighter dependency…
Awkward Array provides NumPy-like operations on nested, variable-sized data structures—lists, records, mixed types, and missing values—with compiled performance and dynamic typing.
Install it if your workflow involves JSON-like structures, ragged arrays, or hierarchical data that NumPy alone cannot handle efficiently.
nvdisasm disassembles NVIDIA CUDA cubin files into human-readable CUDA assembly code, extracting and presenting the compiled GPU kernel information.
hvPlot provides a high-level plotting API that wraps HoloViews, Bokeh, and other visualization backends, letting you create interactive plots from Pandas, Polars, XArray, Dask, and other data sources using a familiar `.plot()` syntax.
TF-Keras is a pure-TensorFlow implementation of Keras that provides a high-level deep learning API for building and training neural networks.
The nightly build is suitable for developers and researchers who can tolerate frequent changes and want early access to new features, but not for production systems…
Provides predefined Kubeflow Pipelines components for Google Cloud Vertex AI that handle model training, deployment, and data processing tasks without writing custom pipeline logic.
Install only if you have a GCP project with Vertex API enabled and authenticated credentials; it is not useful without Google Cloud infrastructure.
Supervision provides utilities for loading, annotating, and processing computer vision datasets and model outputs—detection, segmentation, and classification results from any model framework.
Install it if you work with detection, segmentation, or classification models and want to avoid reinventing dataset utilities.
Provides custom trait types for scientific computing, extending the traitlets validation framework with domain-specific attribute types.
However, the aging maintenance status (296 days since last release) means you should verify that the trait types you need are already implemented and stable; expect…
cvxopt is a Python library for solving convex optimization problems, providing algorithms and data structures for linear, quadratic, and semidefinite programming.
httpstan provides an HTTP REST interface to the Stan C++ library for Bayesian inference, allowing clients to compile Stan models and draw samples via HTTP requests rather than direct C++ calls.
Provides the NVIDIA CUDA nvcc compiler for building CUDA applications, packaged as a Python distribution for easy installation across Linux and Windows platforms.
However, verify that the proprietary license terms fit your use case, and confirm that CUDA 12.9.86 matches your GPU and toolkit requirements before committing to…
pysam reads, manipulates, and writes genomic data files (SAM/BAM/CRAM/VCF/BCF/BED/GFF/GTF/FASTA/FASTQ) and provides access to samtools and bcftools command-line functionality through a Python wrapper around HTSlib.
GraphFrames Python wrapper provides graph processing and analysis on Apache Spark, enabling operations like centrality metrics, motif finding, community detection, and traversals on distributed graph data.
The main gotcha is the external JVM dependency—you cannot use this package standalone; it requires a working Spark environment.
Fetches macroeconomic and factor data from remote sources like FRED, Fama/French, World Bank, and OECD, returning it as pandas DataFrames.
Install it if you need programmatic access to FRED, Fama/French, World Bank, or similar sources; skip it if you only need ad-hoc manual downloads or work with…
A Hatchling metadata hook that automatically generates an `all` extra combining all optional dependencies, deduplicating and sorting them alphabetically.
hmmlearn implements unsupervised learning and inference algorithms for Hidden Markov Models with a scikit-learn compatible API.
Silero VAD detects speech activity in audio files and streams, identifying when voice is present and returning timestamps of speech segments.
The main gotcha is ensuring an audio backend (FFmpeg, sox, or soundfile) is available on your deployment target—verify that before committing to it in a containerized…
scikit-video reads, writes, and analyzes video files using FFmpeg, providing Python functions for frame extraction and video quality metrics including SSIM, PSNR, NIQE, and BRISQUE.
The main gotcha is the hard requirement for FFmpeg on the system PATH and Python >= 3.10; verify your environment supports both before committing.
Canmatrix parses, manipulates, and converts CAN bus database files between multiple formats (.dbc, .dbf, .kcd, .arxml, .yaml, .xls, .sym, .xml, .ldf, .odx) and exports to formats including .json, .lua, and .scapy.
Install it if you work with automotive CAN protocols, need to translate between vendor database formats, or want to programmatically inspect CAN message definitions.
Annoy searches for approximate nearest neighbors in high-dimensional vector spaces using a C++ library with Python bindings, and stores indexes as memory-mapped files that multiple processes can share.
TensorFlow Probability provides probabilistic modeling, statistical inference, and Bayesian machine learning tools integrated with TensorFlow, including distributions, variational inference, MCMC sampling, and neural network layers with uncertainty quantification.
Provides compiled C++ kernels and extensions that accelerate operations on nested, variable-sized data structures with performance comparable to NumPy.
Adds XRootD storage support to fsspec, enabling file access to XRootD servers through fsspec's unified interface using the 'root' protocol.
sktime provides a unified interface for time series machine learning tasks including forecasting, classification, clustering, anomaly detection, and regression, with scikit-learn compatible tools for model building and validation.
Spark NLP provides distributed natural language processing on Apache Spark, offering pretrained pipelines and models for tokenization, named entity recognition, sentiment analysis, machine translation, and embeddings across multiple languages.
Install only if you already have Apache Spark 3.0+ and Java 8 or 11 in your environment; it is not suitable for lightweight single-machine NLP work.
Uproot reads and writes ROOT files (the data format used in high-energy physics) directly in Python using NumPy, without requiring the C++ ROOT library.
Install it if you work with ROOT files or need to integrate them into Python data pipelines.
Provides a Python interface for writing high-performance CUDA kernels using CuTe DSL abstractions, targeting NVIDIA Tensor Cores on Ampere, Hopper, and Blackwell architectures without requiring deep C++ expertise.
However, the public beta status and unclear license terms warrant caution for production use—verify licensing and test stability for your workload before committing.
Mlxtend provides ensemble methods, feature selection, visualization utilities, and frequent pattern mining algorithms for machine learning workflows.
Install it if you need these specific capabilities.
Serializes and deserializes NumPy arrays and Python complex types using the msgpack binary format, preserving numerical data types during encoding and decoding.
However, verify compatibility with your Python version (last release was 2022-06-09) and confirm that read-only deserialized arrays and the object-dtype pickle…
Provides unit-aware measurement objects for Python that support conversion between different units and arithmetic operations across multiple measurement types including distance, weight, temperature, energy, speed, volume, time, and area.
Stanza is a Python NLP library that runs accurate natural language processing tools on 60+ languages, including tokenization, part-of-speech tagging, dependency parsing, and named entity recognition, with optional access to Java Stanford CoreNLP.
Install it if you need dependency parsing, NER, or POS tagging across many languages or in biomedical domains; skip it only if you need real-time performance on…
scikit-network provides graph algorithms and analysis tools for Python, representing graphs as sparse matrices and offering a scikit-learn-inspired API for machine learning on network data.
plyfile reads and writes ASCII and binary PLY (Polygon File Format) files, commonly used for 3D mesh and point cloud data.