padelpy
A Python wrapper for PaDEL-Descriptor software
Decision gist · record as of 2026-08-14
Yes, if you need PaDEL-Descriptor's specific descriptor set and have Java 8+ available. The package is actively maintained, has no Python dependencies, and integrates cleanly into cheminformatics workflows. Install it if you're already committed to PaDEL's descriptor types; if you're exploring alternatives, consider RDKit or Mordred first, as they have different descriptor engines and dependency profiles.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Java Runtime Environment (JRE) 8+ on PATH; verify with `java -version`.
- Python 3.10 or later required.
- Low install friction: pure Python wheel with no runtime dependencies beyond the standard library.
License · maintenance · safety
MIT (permissive) — MIT license is permissive, allowing commercial and private use with minimal restrictions. You may use, modify, and distribute this package freely as long as you include the license notice.
last release 2026-07-22 (23 days) · last repo commit 2026-07-25 · 233 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 92,235 downloads/mo, #13,465 on PyPI
Alternatives
Verify before relying
pip install padelpy
from padelpy import from_smiles
# Compute descriptors for a molecule
descriptors = from_smiles("CCC")
# Compute descriptors and fingerprints
desc_fp = from_smiles("CCC", fingerprints=True)- Performance characteristics and typical runtime for descriptor calculation on large molecule batches
- Specific descriptor and fingerprint types available beyond PubChem fingerprints mentioned
- Memory requirements for processing large SDF or MDL files
What it is and what it does
PaDELPy is a thin Python wrapper around the PaDEL-Descriptor Java engine, a well-established tool for computing molecular descriptors and fingerprints. It exposes PaDEL's command-line interface through Python functions, allowing you to calculate chemical features from SMILES strings, MDL MolFiles, and SDF files without needing to invoke the CLI directly.
The package bundles the PaDEL JAR and dependencies, so no separate download is needed. It provides high-level helpers (`from_smiles`, `from_mdl`, `from_sdf`) that return descriptor values as dictionaries or lists of dictionaries, plus a lower-level `padeldescriptor` function for direct CLI control. It has no Python dependencies beyond the standard library, but requires a Java 8+ runtime on your system PATH.
Use it for
- Calculate molecular descriptors for machine learning pipelines in drug discovery or materials science
- Batch-process SMILES strings or structure files to extract chemical fingerprints for similarity searches
- Generate both 2D and 3D descriptors with optional tautomer standardization and aromaticity detection
- Export descriptor results to CSV for downstream analysis in data science workflows
- Parallelize descriptor computation across multiple molecules using the threads parameter
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need PaDEL-Descriptor's specific descriptor set and have Java 8+ available.
The package is actively maintained, has no Python dependencies, and integrates cleanly into cheminformatics workflows. Install it if you're already committed to PaDEL's descriptor types; if you're exploring alternatives, consider RDKit or Mordred first, as they have different descriptor engines and dependency profiles.
Install
padelpy on PyPI
Before you install
Low install friction: pure Python wheel with no runtime dependencies beyond the standard library. Wheels are large (20+ MB) because they bundle the PaDEL JAR and libraries, but installation is straightforward. Repository is active with recent releases.
Requires Java Runtime Environment (JRE) 8+ on PATH; verify with `java -version`. Python 3.10 or later required.
License in practice
MIT license is permissive, allowing commercial and private use with minimal restrictions. You may use, modify, and distribute this package freely as long as you include the license notice.
Quickstart
pip install padelpy
from padelpy import from_smiles
# Compute descriptors for a molecule
descriptors = from_smiles("CCC")
# Compute descriptors and fingerprints
desc_fp = from_smiles("CCC", fingerprints=True)
Verify before relying
- Performance characteristics and typical runtime for descriptor calculation on large molecule batches
- Specific descriptor and fingerprint types available beyond PubChem fingerprints mentioned
- Memory requirements for processing large SDF or MDL files
Package facts
| License | MIT permissive |
| Python support | Supports the current Python release >=3.10 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | None |
| Maintenance | Actively maintained 23 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 92,235 / month, #13,465 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | License :: OSI Approved :: MIT LicenseOperating System :: OS IndependentProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.12Programming Language :: Python :: 3.13 |
Evidence: padelpy-0.1.17-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “chemical fingerprints python”
- padelpyPaDELPy wraps the PaDEL-Descriptor Java engine to compute molecular…
- aimsim-coreProvides core molecular featurization and similarity comparison…
- mhfpEncodes molecular structures as MinHash fingerprints (MHFP6) for fast…
Give your agent the search over MCP, or paste the wish link into any chat.
More Scientific/Engineering packages
NumPy provides an N-dimensional array object and a comprehensive suite of mathematical, linear algebra, Fourier transform, and random number functions for scientific computing in Python.
pandas provides fast, flexible data structures (Series and DataFrame) for loading, cleaning, transforming, and analyzing labeled or relational data in Python.
scipy provides numerical algorithms for mathematics, science, and engineering—including optimization, integration, linear algebra, Fourier transforms, signal and image processing, and ODE solvers—built on numpy arrays.
scikit-learn provides a comprehensive Python library for supervised and unsupervised machine learning, including classification, regression, clustering, dimensionality reduction, and model evaluation tools built on NumPy and SciPy.
Install it if you need to train, evaluate, or deploy supervised or unsupervised learning models.
dill extends Python's pickle module to serialize and deserialize a much wider range of Python objects, including functions, lambdas, classes, and interpreter sessions, to byte streams for storage or network transmission.
Multiprocess is an enhanced fork of Python's standard multiprocessing library that uses dill for better serialization, allowing you to spawn processes with a threading-like API and share complex objects between them.
Install it if you use multiprocessing and encounter pickle serialization limits with lambdas or complex objects.
See also dscribe · mordredcommunity · aimsim-core · datamol · PubChemPy · epam-indigo · mhfp · py2opsin · prolif · selfies