unstructured-client
Python Client SDK for Unstructured API
Decision gist · record as of 2026-08-14
Yes, if you need to extract structured data from unstructured documents and are willing to use a cloud API. The SDK is actively maintained, has low install friction, carries no known vulnerabilities, and uses a permissive license. Requires Python 3.11+. Verify that the free tier and API pricing fit your volume and use case before committing.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Python 3.11 or later.
- API key or credentials for the Unstructured Platform are needed to authenticate requests.
- Low friction installation with a pure-Python wheel.
License · maintenance · safety
MIT (permissive) — MIT license is permissive; you can use this package freely in commercial and private projects with minimal restrictions.
last release 2026-08-05 (9 days) · last repo commit 2026-08-05 · 118 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 8,893,676 downloads/mo, #1,581 on PyPI
Alternatives
Verify before relying
pip install unstructured-client
from unstructured_client import UnstructuredClient
with UnstructuredClient() as client:
res = client.destinations.create_connection_check_destinations(
request={"destination_id": "cb9e35c1-0b04-4d98-83fa-fa6241323f96"}
)- Whether the free tier (1000 pages per day for 14 days) is sufficient for typical use cases.
- Performance characteristics when processing large documents or high-volume batches.
- Supported file formats beyond PDF (mentioned in dependencies but not detailed in excerpt).
What it is and what it does
unstructured-client is a Python HTTP client for the Unstructured Platform API, a cloud service that extracts structured data from unstructured documents. It wraps the Platform's Partition and Workflow endpoints, allowing you to send documents to the API and retrieve parsed, structured output. The SDK handles authentication, retries, error handling, and file uploads.
You use it by instantiating an UnstructuredClient, then calling methods on its resource objects (e.g., destinations, partitions) to submit documents and retrieve results. The SDK depends on httpx for HTTP transport, pydantic for data validation, and pypdf/pypdfium2 for local PDF handling. It supports custom retry strategies and detailed error reporting via UnstructuredClientError and its subclasses.
Use it for
- Extract structured data from PDFs, images, or other documents by sending them to the Unstructured Platform API.
- Build document processing pipelines that parse invoices, contracts, or forms and return JSON-structured output.
- Integrate document parsing into a larger ETL workflow using the Workflow Endpoint.
- Handle file uploads and manage API responses with built-in retry logic and error handling.
- Prototype document extraction without maintaining local parsing infrastructure.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need to extract structured data from unstructured documents and are willing to use a cloud API.
The SDK is actively maintained, has low install friction, carries no known vulnerabilities, and uses a permissive license. Requires Python 3.11+. Verify that the free tier and API pricing fit your volume and use case before committing.
Install
unstructured-client on PyPI
Before you install
Low friction installation with a pure-Python wheel. Active maintenance with a release 9 days ago. Requires Python 3.11 or later. Seven runtime dependencies are all standard HTTP and data-handling libraries (httpx, pydantic, pypdf, pypdfium2, requests-toolbelt, aiofiles, httpcore).
Requires Python 3.11 or later. API key or credentials for the Unstructured Platform are needed to authenticate requests.
License in practice
MIT license is permissive; you can use this package freely in commercial and private projects with minimal restrictions.
Quickstart
pip install unstructured-client
from unstructured_client import UnstructuredClient
with UnstructuredClient() as client:
res = client.destinations.create_connection_check_destinations(
request={"destination_id": "cb9e35c1-0b04-4d98-83fa-fa6241323f96"}
)
Verify before relying
- Whether the free tier (1000 pages per day for 14 days) is sufficient for typical use cases.
- Performance characteristics when processing large documents or high-volume batches.
- Supported file formats beyond PDF (mentioned in dependencies but not detailed in excerpt).
Package facts
| License | MIT permissive |
| Python support | Supports the current Python release >=3.11 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 7 packagesaiofileshttpcorehttpxpydanticpypdfpypdfium2requests-toolbelt |
| Maintenance | Actively maintained 9 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 8,893,676 / month, #1,581 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
Evidence: unstructured_client-0.46.1-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “unstructured data extraction sdk”
- unstructured-clientHTTP client SDK for the Unstructured Platform API that processes…
- azure-ai-contentunderstandingExtracts structured content from documents, video, audio, and images…
- llama-cloud-servicesSDK client for LlamaCloud services: document parsing (LlamaParse),…
Give your agent the search over MCP, or paste the wish link into any chat.
More WWW/HTTP packages
urllib3 is an HTTP client library that provides thread-safe connection pooling, SSL/TLS verification, multipart file uploads, request retries, compression support, and proxy handling for Python applications.
Requests is a Python HTTP library that simplifies sending HTTP/1.1 requests with automatic handling of headers, authentication, cookies, and response parsing.
h11 is a pure-Python HTTP/1.1 protocol implementation that handles parsing and serializing HTTP messages without any built-in I/O, letting you integrate it with any network layer you choose.
HTTPX is a fully featured HTTP client library for Python that provides both sync and async APIs, with support for HTTP/1.1 and HTTP/2, plus an integrated command-line client.
Install it if you are building new projects or modernizing existing ones that rely on HTTP.
A minimal low-level HTTP client library that sends HTTP requests with thread-safe and task-safe connection pooling, supporting HTTP/1.1, HTTP/2, proxies, and both sync and async interfaces.
aiohttp is an async HTTP client and server framework built on asyncio, supporting both WebSockets and middleware-based routing for building concurrent web applications.
Install it if you need async HTTP client or server capabilities in asyncio-based applications.
See also unstructured · unstructured-inference · langchain-unstructured · unstructured-ingest · fathom-python · llama-cloud · google-cloud-documentai · cohere-compass-sdk · Unified-python-sdk