{"categories":[{"label":"Python Modules","url":"https://skillfed.io/packages/category/software-development-libraries-python-modules/16"},{"label":"Internet","url":"https://skillfed.io/packages/category/internet/4"}],"enrichment":{"capability":"Python SDK for the Scrapfly web scraping service, providing access to web scraping, extraction, and screenshot APIs with support for JavaScript rendering, proxy rotation, and anti-bot bypass.","skillfed_tags":["web-scraping","llm-integration","cloud-api"],"use_cases":["Scrape JavaScript-heavy websites by enabling cloud-based headless browser rendering and anti-bot bypass.","Extract structured data from web pages and feed it into LLM pipelines via LlamaIndex or LangChain for RAG systems.","Rotate through proxy pools and geographic locations to bypass IP-based blocking and geo-restrictions.","Capture full-page screenshots of websites for visual monitoring or archival purposes.","Build automated data collection pipelines with retry logic and error handling for unreliable or blocking websites."],"what_it_does":"Scrapfly SDK is a Python client library for the Scrapfly cloud web scraping service. It wraps three main API endpoints\u2014Web Scraping, Extraction, and Screenshot\u2014allowing developers to scrape web pages, extract structured data, and capture screenshots programmatically. The SDK handles authentication, request configuration, and response parsing, with built-in support for advanced features like JavaScript rendering, anti-bot bypass (ASP), proxy pool selection, and custom JavaScript execution.\n\nThe package integrates with LlamaIndex and LangChain for RAG (Retrieval-Augmented Generation) workflows, enabling developers to scrape web content and feed it directly into LLM training pipelines. It depends on decorator, requests, python-dateutil, loguru, urllib3, and backoff for HTTP handling, logging, and retry logic. Optional extras add asyncio/threading support, Scrapy integration, and a Flask-based webhook server for event handling.","worth_installing":"Yes, if you need cloud-based web scraping with anti-bot features and have a Scrapfly account. The SDK is actively maintained, has low install friction, carries a permissive license, and integrates well with LLM frameworks. Install only if you plan to use the Scrapfly service; it is a client library, not a standalone scraper."},"id":"scrapfly-sdk","links":{"html":"https://skillfed.io/packages/scrapfly-sdk","md":"https://skillfed.io/packages/scrapfly-sdk.md","pypi":"https://pypi.org/project/scrapfly-sdk/"},"maintenance":{"status":"active"},"meta":{"latest_release":"2026-06-14","license_spdx":null,"license_treatment":"permissive","name":"scrapfly-sdk","python_support":"supports_current","summary":"Scrapfly SDK for Scrapfly"},"popularity":{"monthly_downloads":208755,"position":9526,"tier":"top_15000"},"security":{"n_vulnerabilities":0},"version":"0.11.1"}
