scrapfly-sdk
Scrapfly SDK for Scrapfly
Decision gist · record as of 2026-08-14
Yes, if you need cloud-based web scraping with anti-bot features and have a Scrapfly account. The SDK is actively maintained, has low install friction, carries a permissive license, and integrates well with LLM frameworks. Install only if you plan to use the Scrapfly service; it is a client library, not a standalone scraper.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires a valid Scrapfly API key from https://www.scrapfly.io/ to make actual requests.
- Low install friction with a pure Python wheel and six common runtime dependencies.
- Active maintenance with a recent release (61 days ago) and ongoing repository activity.
License · maintenance · safety
BSD (permissive) — BSD permissive license allows commercial and private use with minimal restrictions.
last release 2026-06-14 (61 days) · last repo commit 2026-08-05 · 62 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 208,755 downloads/mo, #9,526 on PyPI
Alternatives
Verify before relying
pip install scrapfly-sdk
from scrapfly_sdk import ScrapeConfig, ScrapflyClient
client = ScrapflyClient(api_key="your_api_key")
config = ScrapeConfig(url="https://example.com")
result = client.scrape(config)- Whether the optional extras (seepdup, concurrency, scrapy, webhook-server) add meaningful value for typical use cases.
- Performance characteristics and rate limits of the Scrapfly service itself.
- Pricing model and free tier limitations for API usage.
What it is and what it does
Scrapfly SDK is a Python client library for the Scrapfly cloud web scraping service. It wraps three main API endpoints—Web Scraping, Extraction, and Screenshot—allowing developers to scrape web pages, extract structured data, and capture screenshots programmatically. The SDK handles authentication, request configuration, and response parsing, with built-in support for advanced features like JavaScript rendering, anti-bot bypass (ASP), proxy pool selection, and custom JavaScript execution.
The package integrates with LlamaIndex and LangChain for RAG (Retrieval-Augmented Generation) workflows, enabling developers to scrape web content and feed it directly into LLM training pipelines. It depends on decorator, requests, python-dateutil, loguru, urllib3, and backoff for HTTP handling, logging, and retry logic. Optional extras add asyncio/threading support, Scrapy integration, and a Flask-based webhook server for event handling.
Use it for
- Scrape JavaScript-heavy websites by enabling cloud-based headless browser rendering and anti-bot bypass.
- Extract structured data from web pages and feed it into LLM pipelines via LlamaIndex or LangChain for RAG systems.
- Rotate through proxy pools and geographic locations to bypass IP-based blocking and geo-restrictions.
- Capture full-page screenshots of websites for visual monitoring or archival purposes.
- Build automated data collection pipelines with retry logic and error handling for unreliable or blocking websites.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need cloud-based web scraping with anti-bot features and have a Scrapfly account.
The SDK is actively maintained, has low install friction, carries a permissive license, and integrates well with LLM frameworks. Install only if you plan to use the Scrapfly service; it is a client library, not a standalone scraper.
Install
scrapfly-sdk on PyPI
Before you install
Low install friction with a pure Python wheel and six common runtime dependencies. Active maintenance with a recent release (61 days ago) and ongoing repository activity.
Requires a valid Scrapfly API key from https://www.scrapfly.io/ to make actual requests.
License in practice
BSD permissive license allows commercial and private use with minimal restrictions.
Quickstart
pip install scrapfly-sdk
from scrapfly_sdk import ScrapeConfig, ScrapflyClient
client = ScrapflyClient(api_key="your_api_key")
config = ScrapeConfig(url="https://example.com")
result = client.scrape(config)
Verify before relying
- Whether the optional extras (seepdup, concurrency, scrapy, webhook-server) add meaningful value for typical use cases.
- Performance characteristics and rate limits of the Scrapfly service itself.
- Pricing model and free tier limitations for API usage.
Package facts
| License | BSD permissive |
| Python support | Supports the current Python release >=3.6 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 6 packagesdecoratorrequestspython-dateutilloguruurllib3backoff |
| Maintenance | Actively maintained 61 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 208,755 / month, #9,526 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Intended Audience :: DevelopersProgramming Language :: PythonProgramming Language :: Python :: 3Programming Language :: Python :: 3 :: OnlyProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.6Programming Language :: Python :: 3.7Programming Language :: Python :: 3.8Programming Language :: Python :: 3.9Topic :: InternetTopic :: Software Development :: Libraries :: Python Modules |
Evidence: scrapfly_sdk-0.11.1-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “web scraping sdk”
- scrapfly-sdkPython SDK for the Scrapfly web scraping service, providing access to…
- scrapegraph-pyClient SDK for the ScrapeGraphAI managed API, enabling web scraping,…
- scrapingbeePython SDK wrapper for ScrapingBee's web scraping API, providing…
Give your agent the search over MCP, or paste the wish link into any chat.
More Python Modules packages
Converts domain names between Unicode and ASCII-compatible encoding (Punycode) according to IDNA 2008 and Unicode Technical Standard 46, with security validation and broader script coverage than the standard library.
Install it if you work with internationalized domain names, need to validate domains, or use HTTP clients that depend on it transitively.
Setuptools is a Python build backend and package management tool that handles building, distributing, and installing Python packages, including support for C/C++ extension modules.
PyYAML parses and emits YAML 1.1 data format, enabling serialization and deserialization of configuration files and Python objects to and from human-readable YAML text.
Pydantic validates Python data structures against type hints, coercing and checking input at runtime to ensure it matches a declared schema.
Provides reusable metadata objects for use with PEP-593 `typing.Annotated` to express common constraints like bounds, collection sizes, and predicates on types.
Install it if you use or build libraries that need to express type constraints in a standardized, inspectable way—or if you want to annotate your own types with…
Provides runtime tools to inspect and introspect Python type annotations, enabling programmatic examination of type hints at execution time.
See also scrapingbee · scrapinghub · linkup-sdk · scrapegraph-py · zenrows · scrapy-zyte-api · scrapling · spider-client · firecrawl · oxylabs