llama-cloud-services
Tailored SDK clients for LlamaCloud services.
Decision gist · record as of 2026-08-14
No for new projects—migrate to llama-cloud>=1.0 instead. Yes-with-conditions for existing codebases: only if you are already using this package and plan to migrate before May 1, 2026. The deprecation status and active migration path make this a short-term dependency; do not adopt it for new work.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires a valid LlamaCloud API key from https://cloud.llamaindex.ai/ (or EU region at https://cloud.eu.llamaindex.ai/)
- Low install friction with a pure-Python wheel.
- Maintenance status is aging: the package is deprecated as of the fact sheet, with active support ending May 1, 2026; users are directed to migrate to llama-cloud>=1.0.
License · maintenance · safety
MIT (permissive) — MIT license permits commercial and private use with minimal restrictions, making it suitable for most projects—though the deprecation status should factor into adoption decisions.
last release 2026-02-13 (182 days)
0 known vulnerabilities (OSV.dev, 2026-08-14) · 93,936,970 downloads/mo, #363 on PyPI
Alternatives
Verify before relying
pip install llama-cloud-services
from llama_cloud_services import LlamaParse
parser = LlamaParse(api_key="YOUR_API_KEY")
# Use parser to process documents- Whether the deprecated package will continue to receive security patches until May 1, 2026, or only critical fixes
- Performance and feature parity between this version and the recommended llama-cloud>=1.0 migration target
- Whether existing code using this package will break when support ends in May 2026
What it is and what it does
llama-cloud-services is a Python SDK for interacting with LlamaCloud, a suite of GenAI-native document processing services. It provides three main components: LlamaParse for parsing complex documents, LlamaExtract for transforming data into structured JSON, and LlamaCloud Index for automated document ingestion and retrieval. The package wraps remote API calls to these services, requiring an API key and internet connectivity.
The package is built on llama-index-core and standard utilities (click, pydantic, tenacity, python-dotenv), with low install friction. However, it is deprecated and will be maintained only until May 1, 2026; the maintainers recommend migrating to llama-cloud>=1.0 for new projects. Existing users should plan a migration path.
Use it for
- Parse complex PDF, image, and document formats into structured text for RAG pipelines
- Extract and transform unstructured document data into validated JSON schemas for downstream processing
- Build automated document ingestion workflows with LlamaCloud Index for retrieval-augmented generation
- Process documents in EU-compliant infrastructure using the EU SaaS endpoint
- Integrate document parsing into agent workflows that require structured data extraction
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
No for new projects—migrate to llama-cloud>=1.0 instead.
Yes-with-conditions for existing codebases: only if you are already using this package and plan to migrate before May 1, 2026. The deprecation status and active migration path make this a short-term dependency; do not adopt it for new work.
Install
llama-cloud-services on PyPI
Before you install
Low install friction with a pure-Python wheel. Maintenance status is aging: the package is deprecated as of the fact sheet, with active support ending May 1, 2026; users are directed to migrate to llama-cloud>=1.0.
Requires a valid LlamaCloud API key from https://cloud.llamaindex.ai/ (or EU region at https://cloud.eu.llamaindex.ai/)
License in practice
MIT license permits commercial and private use with minimal restrictions, making it suitable for most projects—though the deprecation status should factor into adoption decisions.
Quickstart
pip install llama-cloud-services
from llama_cloud_services import LlamaParse
parser = LlamaParse(api_key="YOUR_API_KEY")
# Use parser to process documents
Verify before relying
- Whether the deprecated package will continue to receive security patches until May 1, 2026, or only critical fixes
- Performance and feature parity between this version and the recommended llama-cloud>=1.0 migration target
- Whether existing code using this package will break when support ends in May 2026
Package facts
| License | MIT permissive |
| Python support | Supports the current Python release <4.0,>=3.9 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 9 packagesclickeval-type-backportllama-cloudllama-index-corepackagingplatformdirspydanticpython-dotenvtenacity |
| Maintenance | Aging 182 days since the last release |
| First released | |
| Downloads | 93,936,970 / month, #363 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
Evidence: llama_cloud_services-0.6.94-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “llm document ingestion”
- llama-cloud-servicesSDK client for LlamaCloud services: document parsing (LlamaParse),…
- h2ogptePython client for querying and managing documents in H2OGPTe, an…
- unstructuredIngests and pre-processes unstructured documents (PDFs, HTML, Word,…
Give your agent the search over MCP, or paste the wish link into any chat.
More Artificial Intelligence packages
LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.
Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.
Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.
Install it if you work with Hugging Face Hub models or datasets.
LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.
hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.
Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.
Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.
Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.
See also llama-cloud · llama-index-indices-managed-llama-cloud · llama-parse · llama-index-readers-llama-parse · llama-index · llama-index-cli · h2ogpte · llama-index-core · llama-index-retrievers-bm25 · llama-index-embeddings-langchain