html-table-parser-python3
A small and simple HTML table parser not requiring any external dependency.
Decision gist · record as of 2026-08-14
Yes, if you need simple HTML table extraction without external dependencies and can accept an abandoned codebase. The package is stable and works on current Python versions, but it will not receive updates. Consider copying the class directly into your code if you want to avoid a dependency on an unmaintained package. The AGPL-3.0-or-later license requires careful review for proprietary use.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Low install friction with no runtime dependencies.
- However, the package is abandoned—last commit was 2022-12-12 and no updates have been released since the latest release on 2022-12-06.
- It remains functional for current Python versions but will not receive bug fixes or security updates.
License · maintenance · safety
AGPL-3.0-or-later (agpl) — Licensed under AGPL-3.0-or-later, which requires derivative works and modifications to be released under the same license and made available to users. This is a strong copyleft license; using it in proprietary software requires careful licensing review.
last release 2022-12-06 (1347 days) · last repo commit 2022-12-12 · 86 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 78,556 downloads/mo, #14,430 on PyPI
Alternatives
Verify before relying
from html_table_parser import TableHTMLParser
parser = TableHTMLParser()
parser.feed(html_string)
tables = parser.tables- Whether the package handles malformed or edge-case HTML robustly given its abandoned status
- Performance characteristics on large HTML documents or tables with many rows
What it is and what it does
html-table-parser-python3 is a lightweight HTML table parser that requires only Python's standard library. It extracts tables from HTML markup and returns them as nested lists—tables contain rows, rows contain cells as strings—with all HTML tags stripped and text content joined. The package includes both a programmatic API and a command-line tool for converting HTML tables to CSV.
The parser is intentionally minimal; the author notes that copying the single class from parse.py into your own code is a viable alternative to installing the package. It supports Python 3.4 through 3.11 and has no external dependencies, making it suitable for simple table extraction tasks where you want to avoid heavyweight parsing libraries.
Use it for
- Extract tabular data from archived or static HTML pages for data analysis or migration
- Convert HTML tables to CSV format using the included command-line tool
- Parse tables from web scraping tasks when you want to avoid external dependencies
- Quick prototyping of table extraction logic without adding library dependencies to your project
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need simple HTML table extraction without external dependencies and can accept an abandoned codebase.
The package is stable and works on current Python versions, but it will not receive updates. Consider copying the class directly into your code if you want to avoid a dependency on an unmaintained package. The AGPL-3.0-or-later license requires careful review for proprietary use.
Install
html-table-parser-python3 on PyPI
Before you install
Low install friction with no runtime dependencies. However, the package is abandoned—last commit was 2022-12-12 and no updates have been released since the latest release on 2022-12-06. It remains functional for current Python versions but will not receive bug fixes or security updates.
License in practice
Licensed under AGPL-3.0-or-later, which requires derivative works and modifications to be released under the same license and made available to users. This is a strong copyleft license; using it in proprietary software requires careful licensing review.
Quickstart
from html_table_parser import TableHTMLParser
parser = TableHTMLParser()
parser.feed(html_string)
tables = parser.tables
Verify before relying
- Whether the package handles malformed or edge-case HTML robustly given its abandoned status
- Performance characteristics on large HTML documents or tables with many rows
Package facts
| License | AGPL-3.0-or-later agpl |
| Python support | Supports the current Python release >=3,<4 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | None |
| Maintenance | Abandoned 1,347 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 78,556 / month, #14,430 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | License :: OSI Approved :: GNU Affero General Public License v3 or later (AGPLv3+)Programming Language :: Python :: 3Programming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.4Programming Language :: Python :: 3.5Programming Language :: Python :: 3.6Programming Language :: Python :: 3.7Programming Language :: Python :: 3.8Programming Language :: Python :: 3.9 |
Evidence: html_table_parser_python3-0.3.1-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “parse HTML tables”
- html-table-parser-python3Parses HTML tables into nested lists of rows and cells without…
- html-to-jsonConverts HTML documents and HTML tables into JSON structures,…
- sphinx-markdown-tablesAdds markdown table support to Sphinx documentation by extending…
Give your agent the search over MCP, or paste the wish link into any chat.
More HTML packages
MarkupSafe provides a text object that escapes special characters so untrusted strings can be safely embedded in HTML and XML without injection attacks.
Jinja2 is a templating engine that renders dynamic content by combining templates with Python-like syntax and data, supporting template inheritance, macros, autoescaping, and sandboxed execution.
Beautiful Soup parses HTML and XML documents into a navigable tree, providing Pythonic methods to search, iterate, and modify the parsed content.
Install it if you need to parse or extract data from markup documents.
lxml provides Python bindings to libxml2 and libxslt, enabling parsing, validation, and transformation of XML and HTML documents through an ElementTree-compatible API with support for XPath, XSLT, and schema validation.
Install it if you need robust XML/HTML parsing, validation, or transformation; avoid it only if you must stay pure-Python and can accept slower performance.
Docutils converts plaintext documentation in reStructuredText format into multiple output formats including HTML, XML, and LaTeX using a modular processing system.
Converts Markdown text to HTML using a Python implementation of John Gruber's Markdown specification, with support for extensions.
Install it if you need to parse Markdown in Python.
See also numbers-parser · texttable · camelot-py · aspose-cells-python · sttable · html-to-json · img2table · aspose-cells · query-string · tabulator