mrmr-selection
minimum-Redundancy-Maximum-Relevance algorithm for feature selection
Decision gist · record as of 2026-08-14
Yes, if you need minimal-optimal feature selection and accept dormant maintenance. The algorithm is well-established and used in production systems, dependencies are stable, and no known vulnerabilities exist. Install friction is low. However, the last release was 2023-06-30 with no recent commits, so expect no active support for new dependency versions or bug fixes. GPL v3.0 licensing is a hard constraint for proprietary projects. Best suited for open-source or internal ML workflows where feature selection is a one-time or infrequent task.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Low install friction with a pure Python wheel and common data science dependencies.
- Maintenance is dormant—last release was 2023-06-30 and no commits since 2024-11-19—so expect no active bug fixes or updates, though the core algorithm is stable.
License · maintenance · safety
GNU General Public License v3.0 (unclear) — Licensed under GNU General Public License v3.0, which requires derivative works and distributions to also be open-source under GPL v3.0. This is a strong copyleft license; proprietary or closed-source projects may face legal constraints.
last release 2023-06-30 (1141 days) · last repo commit 2024-11-19 · 631 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 73,500 downloads/mo, #14,988 on PyPI
Alternatives
Verify before relying
pip install mrmr_selection
import pandas as pd
from mrmr import mrmr_classif
X = pd.DataFrame([[1, 2], [3, 4]])
y = pd.Series([0, 1])
selected_features = mrmr_classif(X=X, y=y, K=1)- Whether the package actively maintains compatibility with recent versions of its dependencies (pandas, numpy, scipy, etc.)
- Current status of Spark and BigQuery module support and any known limitations
- Whether Python version support is documented elsewhere (requires_python is empty in metadata)
- Performance characteristics and scalability limits on large datasets
What it is and what it does
mrmr_selection implements a minimal-optimal feature selection algorithm designed to find the smallest subset of features that retain predictive power for a machine learning task. Unlike all-relevant methods that identify every feature with some relationship to the target, mRMR prioritizes efficiency by selecting only the most informative features while minimizing redundancy among them.
The package provides separate modules for Pandas, Polars, Spark, and BigQuery, each exposing mrmr_classif (for categorical targets) and mrmr_regression (for numeric targets) functions. It depends on pandas, numpy, scipy, joblib, category-encoders, jinja2, tqdm, and polars. The algorithm returns a ranked list of the top K selected features, allowing further filtering if needed.
Use it for
- Reduce dataset dimensionality in production ML pipelines where frequent, automated feature selection is needed without manual tuning.
- Identify the most predictive features in high-dimensional datasets to lower memory and computation costs in model training and inference.
- Improve model interpretability by selecting a minimal set of features that explain predictions while maintaining accuracy.
- Perform feature selection on large-scale data stored in Spark or BigQuery without loading entire datasets into memory.
- Benchmark feature importance across classification and regression tasks to guide domain expert review of model inputs.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need minimal-optimal feature selection and accept dormant maintenance.
The algorithm is well-established and used in production systems, dependencies are stable, and no known vulnerabilities exist. Install friction is low. However, the last release was 2023-06-30 with no recent commits, so expect no active support for new dependency versions or bug fixes. GPL v3.0 licensing is a hard constraint for proprietary projects. Best suited for open-source or internal ML workflows where feature selection is a one-time or infrequent task.
Install
mrmr-selection on PyPI
Before you install
Low install friction with a pure Python wheel and common data science dependencies. Maintenance is dormant—last release was 2023-06-30 and no commits since 2024-11-19—so expect no active bug fixes or updates, though the core algorithm is stable.
License in practice
Licensed under GNU General Public License v3.0, which requires derivative works and distributions to also be open-source under GPL v3.0. This is a strong copyleft license; proprietary or closed-source projects may face legal constraints.
Quickstart
pip install mrmr_selection
import pandas as pd
from mrmr import mrmr_classif
X = pd.DataFrame([[1, 2], [3, 4]])
y = pd.Series([0, 1])
selected_features = mrmr_classif(X=X, y=y, K=1)
Verify before relying
- Whether the package actively maintains compatibility with recent versions of its dependencies (pandas, numpy, scipy, etc.)
- Current status of Spark and BigQuery module support and any known limitations
- Whether Python version support is documented elsewhere (requires_python is empty in metadata)
- Performance characteristics and scalability limits on large datasets
Package facts
| License | GNU General Public License v3.0 unclear |
| Python support | Not specified |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 9 packagescategory-encodersjinja2tqdmjoblibpandasnumpyscikit-learnscipypolars |
| Maintenance | Dormant 1,141 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 73,500 / month, #14,988 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
Evidence: mrmr_selection-0.2.8-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “feature selection algorithm”
- mrmr-selectionImplements the mRMR (minimum Redundancy - Maximum Relevance) feature…
- BorutaBoruta performs all-relevant feature selection by identifying all…
- autogluon.tabularAutomates machine learning model training and prediction on tabular…
Give your agent the search over MCP, or paste the wish link into any chat.
More Artificial Intelligence packages
LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.
Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.
Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.
Install it if you work with Hugging Face Hub models or datasets.
LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.
hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.
Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.
Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.
Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.
See also Boruta · repartipy · percentify · rank-bm25 · polars-ds · feature-engine · sagemaker-feature-store-pyspark-3.1 · k-means-constrained · lttb · azureml-train-automl