Subcategories
Packages
Jupytext converts Jupyter notebooks to and from plain-text formats (Python scripts, Markdown, or other text files) while preserving notebook structure, enabling version control and IDE editing of notebook content.
Install it if you work with Jupyter and Git together, or need to treat notebooks as code.
Provides complete external type annotations for lxml, enabling static type checking and IDE support for code using lxml's XML and HTML processing functionality.
Install it if you use lxml and rely on type checking or IDE support.
Converts XML documents into Python objects, allowing you to work with XML data as native Python instances rather than parsing raw XML text.
Converts marked-up text between formats including plain text, XHTML, RTF, and PDF, preserving basic structure like paragraphs, headings, lists, and simple tables.
Converts standard Markdown to Slack's mrkdwn format, handling headings, text formatting, lists, tables, code blocks, links, and other common Markdown elements for use in Slack messages.
Install it if you need to programmatically format Markdown for Slack; skip it only if you have no Slack integration needs.
Mdformat is a CommonMark-compliant command-line Markdown formatter and Python library that enforces consistent style across Markdown files.
Install it if you want a lightweight, pure-Python formatter; consider its plugin ecosystem if you need support for non-CommonMark dialects.
oyaml is a drop-in replacement for PyYAML that preserves dictionary ordering when dumping and loading YAML, eliminating scrambled key order.
Converts terminal output with ANSI color codes to HTML or LaTeX, preserving colors and formatting for display in browsers or documents.
Install it if you need to convert ANSI-colored terminal output to HTML or LaTeX.
Converts Python dictionaries and other native data types into valid XML strings, handling nested structures, type attributes, and custom root elements.
Renders Handlebars templates for composing LLM prompts, with compile-time validation against Pydantic models to catch typos and missing fields before rendering.
Pybtex reads bibliography data from BibTeX, BibTeXML, or YAML files and generates formatted bibliographies in LaTeX, HTML, markdown, or plain text, supporting both BibTeX style files and Python-based custom styles.
Install it if you need to process BibTeX files programmatically, generate bibliographies in multiple formats, or integrate bibliography handling into a Python workflow.
Provides a Python codec to convert between LaTeX-encoded text and Unicode, suitable for processing short text fragments like BibTeX entries or paragraphs rather than full documents.
Converts Markdown files to Confluence Storage Format and publishes them to Confluence wiki via REST API, handling text formatting, images, code blocks, diagrams, and tables.
Provides the sgmllib SGML parser from Python 2.7 for use with feedparser, enabling parsing of SGML-based feed formats in modern Python versions.
mdutils generates Markdown files programmatically from Python code, letting you format and structure text output with headers, tables, lists, links, and styling as your code runs.
Jinja2 extension that adds template tags for working with dates and times, including current time retrieval with timezone support and relative time offsets.
No—not for new projects.
Enables Jinja2 templating within YAML files by preprocessing templates before YAML parsing and postprocessing after rendering, allowing you to use template logic while keeping the file valid YAML.
Install only if you have an existing codebase already depending on it and cannot migrate.
Mdformat-gfm extends mdformat to format GitHub Flavored Markdown, adding support for tables, task lists, strikethroughs, autolinks, and disallowed raw HTML syntax.
Install it if your workflow involves GitHub markdown; skip it if you only use CommonMark.
Converts Python docstrings from reStructuredText and Google format to Markdown on the fly, with extensibility via entry points for custom converters.
A Python-Markdown extension that fixes list rendering by allowing custom indentation for nested lists and preventing unwanted paragraph tags and line breaks within list items.
Parses BibTeX files into Python data structures, extracting bibliographic metadata from standard citation formats.
However, the high install friction is a concern: verify that installation succeeds in your environment before committing to it.
Parses Python docstrings in ReST, Google, Numpydoc, and Epydoc formats, extracting structured metadata like parameters, return types, and exceptions.
Install it if you need programmatic access to docstring metadata; verify against the CHANGELOG.md that the fork's fixes and additions matter for your use case.
SaxonC-HE is a Python wrapper for Saxon, an XML processor that runs XSLT 3.0 transformations, XQuery 3.1 queries, XPath 3.1 expressions, and XML Schema validation.
Python Liquid is a template engine that renders Liquid templates—a safe, customer-facing template language used by Shopify—from strings or files with variable substitution and control flow.
A command-line tool that renders Jinja2 templates using data from INI, YAML, JSON, or environment variables, with support for custom filters, tests, and plugins.
Install it if you need Jinja2 templating outside a Python application.
Converts between ASCIIMath, LaTeX, and MathML mathematical notation formats, with bidirectional translation support and a command-line interface.
However, do not rely on it for new features or bug fixes—treat it as a snapshot tool.
Implements JSONPath, a query language for extracting data from JSON documents using XPath-like syntax to navigate and filter nested structures.
No—not recommended for new projects.
Sphinx extension that parses and converts Jupyter Notebooks (.ipynb files) directly into Sphinx documentation, built on the MyST markdown parser.
Scrapling is a web scraping and crawling framework that handles single requests to full-scale crawls, with built-in anti-bot bypass, adaptive element relocation, proxy rotation, and concurrent spider support.
Install it if you need to scrape protected or dynamic websites at scale; skip it if you only need simple static HTML parsing.
Parses MediaWiki wikicode into a queryable syntax tree, allowing you to extract, analyze, and modify templates, links, and other wiki markup elements programmatically.
Pyandoc wraps Pandoc to convert between document formats by reading from and writing to a Document object's format properties.
Generates formatted markdown tables from lists of dictionaries, with support for padding, alignment, multiline text, and float rounding.
Converts markdown documents to plain text by stripping formatting and rendering only the text content.
Provides a Python-packaged distribution of TinyXML-2, a lightweight C++ XML parser, via the cmeel build system for cross-platform use.
EmPy is a templating system that embeds Python code directly into text documents using customizable markup (default `@`), processing the combined source to produce output with Python expressions, statements, and control structures evaluated inline.
PyLaTeX generates LaTeX documents and code snippets from Python, letting you programmatically create PDFs and typeset content without writing raw LaTeX by hand.
Encodes and decodes TOON (Token-Oriented Object Notation), a compact data format designed to represent structured data in 30-60% fewer tokens than JSON, optimized for transmission to Large Language Models.
imgkit wraps the wkhtmltoimage command-line tool to convert HTML (from URLs, files, or strings) into image files using the Webkit rendering engine.
However, do not use it in production without understanding that no upstream maintenance is available—if wkhtmltoimage itself breaks or you encounter bugs in imgkit,…
Converts XML documents into Python objects with dot-notation access to elements and attributes, handling special characters by substituting them with underscores.
However, avoid it for production systems, security-sensitive applications, or projects requiring ongoing maintenance—consider lxml or xml.etree.ElementTree instead.
Converts markdown text to PDF files, with support for tables, images, hyperlinks, table of contents, custom CSS styling, and optional diagram rendering via plugins.