$npx skillfedfor your agent

markitdown-no-magika

Utility tool for converting various files to Markdown

With conditionsPyPI MarkupReleased Jun 2025126.9K downloads / moMITPure Python

Decision gist · record as of 2026-08-14

pure-Python wheel — markitdown_no_magika-0.1.2-py3-none-any.whl
v0.1.2 · released 2025-06-30 · Python >=3.10 · 5 runtime deps: beautifulsoup4, charset-normalizer, defusedxml, markdownify, requests

Yes, if you need to convert multiple file formats to Markdown and can tolerate Beta-stage API stability. The low install friction, active maintenance, and permissive license make it a reasonable choice for indexing and text analysis tasks. Avoid for production systems where API stability is critical until the package reaches a stable release.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • Requires Python 3.10 or later; some file formats may require additional system dependencies not listed in the package metadata.
  • Low friction install with five runtime dependencies.
  • Actively maintained with recent commits; however, the package is in Beta status and has been available for only a few months, so production use should account for potential API changes.

License · maintenance · safety

MIT (permissive) — MIT license permits commercial and private use with minimal restrictions; you must include a copy of the license and copyright notice in distributions.

last release 2025-06-30 (410 days) · last repo commit 2026-07-29 · 173,795 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 126,858 downloads/mo, #11,760 on PyPI

Verify before relying

pip install markitdown-no-magika

from markitdown import MarkItDown

md = MarkItDown()
result = md.convert("test.xlsx")
print(result.text_content)
  • Which file formats are supported beyond PDF, Excel, Word, HTML, and images—complete list not provided in fact sheet.
  • Whether the package handles large files efficiently or has memory/performance constraints.
  • What 'all' extras install and whether they are necessary for common use cases.
Same gist for agents: .md · .json

What it is and what it does

MarkItDown is a Python utility that converts documents and files from various formats into Markdown. It works both as a standalone command-line tool and as a Python library, making it useful for preparing content for indexing, text analysis, or integration into documentation pipelines. The package depends on beautifulsoup4 for HTML parsing, markdownify for format conversion, requests for fetching remote content, and defusedxml for secure XML handling.

The package is actively maintained and currently in Beta status (version 0.1.2), with support for Python 3.10 through 3.13. It has a permissive MIT license and low installation friction. The repository shows strong community interest, though as a young project it may still see breaking changes.

Use it for

  • Convert PDF documents to Markdown for indexing in search systems or knowledge bases.
  • Extract text from Excel spreadsheets and Word documents as structured Markdown.
  • Batch-process HTML files or web content into Markdown for archival or analysis.
  • Prepare various file formats for input to text analysis pipelines.
  • Automate document conversion workflows in CI/CD or data processing scripts.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

With conditions

Yes, if you need to convert multiple file formats to Markdown and can tolerate Beta-stage API stability.

The low install friction, active maintenance, and permissive license make it a reasonable choice for indexing and text analysis tasks. Avoid for production systems where API stability is critical until the package reaches a stable release.

Install

markitdown-no-magika on PyPI

Before you install

Low friction install with five runtime dependencies. Actively maintained with recent commits; however, the package is in Beta status and has been available for only a few months, so production use should account for potential API changes.

Requires Python 3.10 or later; some file formats may require additional system dependencies not listed in the package metadata.

License in practice

MIT license permits commercial and private use with minimal restrictions; you must include a copy of the license and copyright notice in distributions.

Quickstart

pip install markitdown-no-magika

from markitdown import MarkItDown

md = MarkItDown()
result = md.convert("test.xlsx")
print(result.text_content)

Verify before relying

  • Which file formats are supported beyond PDF, Excel, Word, HTML, and images—complete list not provided in fact sheet.
  • Whether the package handles large files efficiently or has memory/performance constraints.
  • What 'all' extras install and whether they are necessary for common use cases.

Package facts

LicenseMIT permissive
Python supportSupports the current Python release >=3.10
Install frictionLow. Pure-Python wheel
Runtime dependencies
5 packages
beautifulsoup4charset-normalizerdefusedxmlmarkdownifyrequests
MaintenanceActively maintained 410 days since the last release
Last repo commit
First released
Downloads126,858 / month, #11,760 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
Development Status :: 4 - BetaProgramming Language :: PythonProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.12Programming Language :: Python :: 3.13Programming Language :: Python :: Implementation :: CPythonProgramming Language :: Python :: Implementation :: PyPy

Evidence: markitdown_no_magika-0.1.2-py3-none-any.whl

Tags

Capabilities
convert files to markdownpdf to markdown converterdocument format conversionmarkdown extraction toolfile to markdown pythonbatch document conversiontext extraction and markdown
Topics
document-conversionmarkdown-generationcli-tool

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “markdown extraction tool”

  • markitdown-no-magikaConverts various file formats (PDF, Excel, Word, HTML, images, and…
  • strip-markdownConverts markdown text to plain text, removing all formatting, links,…
  • notion-to-md-pyConverts Notion pages to Markdown files via the Notion API,…

Give your agent the search over MCP, or paste the wish link into any chat.

More Markup packages

PyYAML Worth it
PyPI · Python Modules · released Sep 2025

PyYAML parses and emits YAML 1.1 data format, enabling serialization and deserialization of configuration files and Python objects to and from human-readable YAML text.

MITcompiled wheel · 3.8+
1.2Bdownloads / mo
markdown-it-py Worth it
PyPI · Python Modules · released May 2026

A Python markdown parser that converts markdown to HTML following the CommonMark specification, with support for plugins and custom syntax rules.

Install it if you need reliable markdown-to-HTML conversion.

MITpure Python · 3.10+
613.8Mdownloads / mo
beautifulsoup4 Worth it
PyPI · Python Modules · released Jun 2026

Beautiful Soup parses HTML and XML documents into a navigable tree, providing Pythonic methods to search, iterate, and modify the parsed content.

Install it if you need to parse or extract data from markup documents.

MITpure Python · 3.7.0+
432.1Mdownloads / mo
et-xmlfile With conditions
PyPI · Markup · released Oct 2024

et_xmlfile writes large XML files with minimal memory overhead by serializing elements to disk as they are created, rather than holding the entire tree in memory.

Install it if incremental XML writing fits your use case; skip it if your XML documents are small or you already use lxml.

MITpure Python · 3.8+dormant
343.3Mdownloads / mo
tomlkit Worth it
PyPI · Markup · released Jul 2026

Parses and edits TOML files while preserving formatting, comments, and structure, then serializes them back with layout intact.

Install it if you're building tools that touch TOML files and user readability of the source matters.

MITpure Python · 3.9+
338.0Mdownloads / mo
docstring-parser Worth it
PyPI · Python Modules · released Apr 2026

Parses Python docstrings in ReST, Google, Numpydoc, and Epydoc formats, extracting structured information like descriptions, parameters, return types, and exceptions.

Install it if you need to programmatically read and extract structured data from Python docstrings.

MITpure Python · 3.8+
274.9Mdownloads / mo

See also markitdown · markitdown-mcp · datalab-python-sdk · landingai-ade · md2pdf · markdown-pdf · strip-markdown · marker-pdf · pypandoc · mdit-plain