$npx skillfedfor your agent

requests-html

HTML Parsing for Humans.

With conditionsPyPI WWW/HTTPReleased Feb 2019488.4K downloads / moMITPure Python

Decision gist · record as of 2026-08-14

pure-Python wheel — requests_html-0.10.0-py3-none-any.whl
v0.10.0 · released 2019-02-17 · Python >=3.6.0 · 7 runtime deps: requests, pyquery, fake-useragent, parse, bs4, w3lib, pyppeteer

Yes, if you need simple HTML parsing and scraping for a one-off project or legacy codebase already using requests. No, if you're starting a new scraping project—the dormant maintenance (last release 2019, no active development) means you won't get bug fixes, security updates, or compatibility improvements for modern websites. Consider actively maintained alternatives like BeautifulSoup with requests, or Scrapy for larger projects.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • pyppeteer (for JavaScript rendering) may require a Chromium binary; verify system dependencies before relying on JavaScript support.
  • Low install friction with a pure-Python wheel.
  • However, the package is dormant—last release was 2019-02-17 and the last commit 2024-06-19, so it receives no active maintenance.

License · maintenance · safety

MIT (permissive) — MIT license (permissive) means you can use, modify, and distribute this package freely in commercial and private projects with minimal restrictions.

last release 2019-02-17 (2735 days) · last repo commit 2024-06-19 · 327 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 488,419 downloads/mo, #6,381 on PyPI

Verify before relying

from requests_html import HTMLSession

session = HTMLSession()
r = session.get('https://example.com')
links = r.html.links
  • Whether pyppeteer's Chromium dependency is automatically provisioned or requires manual setup on all platforms.
  • Current compatibility with modern websites that use heavy JavaScript frameworks or anti-scraping measures.
  • Whether dormant status means security patches or bug fixes are unlikely if issues arise.
Same gist for agents: .md · .json

What it is and what it does

requests-html is a web scraping and HTML parsing library that wraps the requests library with parsing capabilities. It lets you fetch web pages and extract data using CSS selectors (jQuery-style), XPath, or direct link enumeration. The library includes mocked user-agent headers, automatic redirect following, connection pooling, and cookie persistence—all the conveniences of requests plus parsing.

The package supports both synchronous and asynchronous workflows, and includes JavaScript rendering via pyppeteer for pages that require client-side execution. It's designed for developers who want to scrape or parse HTML without learning a heavy framework, though its dormant maintenance status (last release 2019-02-17) means it receives no active updates or security patches.

Use it for

  • Extract all links from a web page for crawling or analysis without manually parsing HTML.
  • Scrape structured data (product names, prices, metadata) from websites using CSS or XPath selectors.
  • Render and scrape JavaScript-heavy pages that require a browser engine to populate content.
  • Build async scrapers that fetch multiple pages concurrently to speed up data collection.
  • Parse HTML responses in a requests-based workflow without switching to a separate parsing library.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

With conditions

Yes, if you need simple HTML parsing and scraping for a one-off project or legacy codebase already using requests.

No, if you're starting a new scraping project—the dormant maintenance (last release 2019, no active development) means you won't get bug fixes, security updates, or compatibility improvements for modern websites. Consider actively maintained alternatives like BeautifulSoup with requests, or Scrapy for larger projects.

Install

requests-html on PyPI

Before you install

Low install friction with a pure-Python wheel. However, the package is dormant—last release was 2019-02-17 and the last commit 2024-06-19, so it receives no active maintenance. The pyppeteer dependency for JavaScript support adds a transitive browser automation layer that may require system dependencies.

pyppeteer (for JavaScript rendering) may require a Chromium binary; verify system dependencies before relying on JavaScript support.

License in practice

MIT license (permissive) means you can use, modify, and distribute this package freely in commercial and private projects with minimal restrictions.

Quickstart

from requests_html import HTMLSession

session = HTMLSession()
r = session.get('https://example.com')
links = r.html.links

Verify before relying

  • Whether pyppeteer's Chromium dependency is automatically provisioned or requires manual setup on all platforms.
  • Current compatibility with modern websites that use heavy JavaScript frameworks or anti-scraping measures.
  • Whether dormant status means security patches or bug fixes are unlikely if issues arise.

Package facts

LicenseMIT permissive
Python supportSupports the current Python release >=3.6.0
Install frictionLow. Pure-Python wheel
Runtime dependencies
7 packages
requestspyqueryfake-useragentparsebs4w3libpyppeteer
MaintenanceDormant 2,735 days since the last release
Last repo commit
First released
Downloads488,419 / month, #6,381 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
License :: OSI Approved :: MIT LicenseProgramming Language :: PythonProgramming Language :: Python :: 3.6Programming Language :: Python :: Implementation :: CPythonProgramming Language :: Python :: Implementation :: PyPy

Evidence: requests_html-0.10.0-py3-none-any.whl

Tags

Capabilities
web scraping html parsingcss selector xpath extractionjavascript rendering htmlasync web scrapinghtml parsing requestsweb page link extractionbrowser automation scraping
Topics
web-scrapinghtml-parsingdormant

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “web scraping html parsing”

  • requests-htmlParses and scrapes HTML from web pages using a simple, intuitive API…
  • BeautifulSoupBeautiful Soup parses HTML and XML documents into a navigable tree,…
  • beautifulsoup4Beautiful Soup parses HTML and XML documents into a navigable tree,…

Give your agent the search over MCP, or paste the wish link into any chat.

More WWW/HTTP packages

urllib3 Worth it
PyPI · Libraries · released May 2026

urllib3 is an HTTP client library that provides thread-safe connection pooling, SSL/TLS verification, multipart file uploads, request retries, compression support, and proxy handling for Python applications.

MITpure Python · 3.10+
1.8Bdownloads / mo
requests Worth it
PyPI · Libraries · released May 2026

Requests is a Python HTTP library that simplifies sending HTTP/1.1 requests with automatic handling of headers, authentication, cookies, and response parsing.

Apache-2.0pure Python · 3.10+
1.8Bdownloads / mo
h11 With conditions
PyPI · WWW/HTTP · released Apr 2025

h11 is a pure-Python HTTP/1.1 protocol implementation that handles parsing and serializing HTTP messages without any built-in I/O, letting you integrate it with any network layer you choose.

MITpure Python · 3.8+aging
894.9Mdownloads / mo
httpx Worth it
PyPI · WWW/HTTP · released Dec 2024

HTTPX is a fully featured HTTP client library for Python that provides both sync and async APIs, with support for HTTP/1.1 and HTTP/2, plus an integrated command-line client.

Install it if you are building new projects or modernizing existing ones that rely on HTTP.

BSD-3-Clausepure Python · 3.8+
797.0Mdownloads / mo
httpcore With conditions
PyPI · WWW/HTTP · released Apr 2025

A minimal low-level HTTP client library that sends HTTP requests with thread-safe and task-safe connection pooling, supporting HTTP/1.1, HTTP/2, proxies, and both sync and async interfaces.

BSD-3-Clausepure Python · 3.8+aging
783.6Mdownloads / mo
aiohttp Worth it
PyPI · WWW/HTTP · released Jul 2026

aiohttp is an async HTTP client and server framework built on asyncio, supporting both WebSockets and middleware-based routing for building concurrent web applications.

Install it if you need async HTTP client or server capabilities in asyncio-based applications.

permissive licensecompiled wheel · 3.10+
643.6Mdownloads / mo

See also requestium · cssselect · itemloaders · parsel · turbohtml · scrapling · cssselect2 · Crawl4AI · selectolax · readable-content