neo4j-graphrag
Python package to allow easy integration to Neo4j's GraphRAG features
Decision gist · record as of 2026-08-14
Yes, if you have a Neo4j instance and need to build RAG applications grounded in structured knowledge graphs. The package is actively maintained, has no known vulnerabilities, and offers a permissive license. Install with caution if you rely on spaCy-based NLP features on Python 3.14 (currently unsupported upstream); otherwise, low friction and well-suited for production use.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires a running Neo4j instance (local or remote).
- APOC core library must be installed in Neo4j for knowledge graph construction.
- At least one LLM provider extra must be installed (e.g., openai, anthropic).
License · maintenance · safety
Apache License, Version 2.0 (permissive) — Apache License 2.0 (permissive): you can use, modify, and distribute this package freely in commercial and open-source projects, provided you include a copy of the license and state material changes.
last release 2026-06-24 (51 days) · last repo commit 2026-08-13 · 1,254 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 440,023 downloads/mo, #6,649 on PyPI
Alternatives
Verify before relying
pip install 'neo4j-graphrag[openai]'
from neo4j import GraphDatabase
from neo4j_graphrag.embeddings import OpenAIEmbeddings
from neo4j_graphrag.experimental.pipeline.kg_builder import SimpleKGPipeline
from neo4j_graphrag.llm import OpenAILLM
driver = GraphDatabase.driver('neo4j://localhost:7687', auth=('neo4j', 'password'))
embedder = OpenAIEmbeddings(model='text-embedding-3-large')
llm = OpenAILLM(model_name='gpt-4')
kg_builder = SimpleKGPipeline(llm=llm, driver=driver, embedder=embedder, schema={...})
await kg_builder.run_async(text='your text here')- Whether the package's experimental KG builder features are production-ready or intended for development/testing only
- Performance characteristics when working with large knowledge graphs or high-volume retrieval queries
- Whether APOC library installation in Neo4j is required for all features or only specific ones
What it is and what it does
Neo4j GraphRAG is a first-party Python library from Neo4j that bridges language models and graph databases to build retrieval-augmented generation systems. It provides two main workflows: constructing knowledge graphs from text or PDFs using LLM-driven entity and relationship extraction, and retrieving relevant graph data to augment LLM prompts for question-answering and reasoning tasks.
The package wraps Neo4j's graph database with high-level abstractions—SimpleKGPipeline for streamlined knowledge graph building, Pipeline for advanced customization, and multiple retriever strategies (vector search, text-to-Cypher, hybrid traversal). It integrates with LLM providers (OpenAI, Anthropic, Cohere, Bedrock, etc.) and optional embeddings backends (sentence-transformers, Weaviate, Pinecone, Qdrant). Core dependencies include pydantic for schema validation, tenacity for retry logic, and utilities like pypdf and json-repair for data handling.
Use it for
- Build a knowledge graph from unstructured documents, then use it to answer domain-specific questions with LLM-grounded retrieval.
- Implement hybrid search combining vector similarity with graph traversal to find contextually relevant entities and their relationships.
- Extract structured facts (entities, relationships, patterns) from PDFs or text using LLM-guided pipelines without manual annotation.
- Create a text-to-Cypher retriever that translates natural language queries into graph database queries for precise, schema-aware retrieval.
- Augment LLM responses with real-time graph data to reduce hallucination and ground answers in a curated knowledge base.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you have a Neo4j instance and need to build RAG applications grounded in structured knowledge graphs.
The package is actively maintained, has no known vulnerabilities, and offers a permissive license. Install with caution if you rely on spaCy-based NLP features on Python 3.14 (currently unsupported upstream); otherwise, low friction and well-suited for production use.
Install
neo4j-graphrag on PyPI
Before you install
Low friction: pure Python wheel with no compiled dependencies. Active maintenance with recent releases; last commit 2026-08-13. Requires Neo4j instance and at least one LLM provider extra (openai, anthropic, etc.) to function for RAG tasks.
Requires a running Neo4j instance (local or remote). APOC core library must be installed in Neo4j for knowledge graph construction. At least one LLM provider extra must be installed (e.g., openai, anthropic). Python 3.10–3.13 recommended; Python 3.14 support limited due to upstream spaCy issue.
License in practice
Apache License 2.0 (permissive): you can use, modify, and distribute this package freely in commercial and open-source projects, provided you include a copy of the license and state material changes.
Quickstart
pip install 'neo4j-graphrag[openai]'
from neo4j import GraphDatabase
from neo4j_graphrag.embeddings import OpenAIEmbeddings
from neo4j_graphrag.experimental.pipeline.kg_builder import SimpleKGPipeline
from neo4j_graphrag.llm import OpenAILLM
driver = GraphDatabase.driver('neo4j://localhost:7687', auth=('neo4j', 'password'))
embedder = OpenAIEmbeddings(model='text-embedding-3-large')
llm = OpenAILLM(model_name='gpt-4')
kg_builder = SimpleKGPipeline(llm=llm, driver=driver, embedder=embedder, schema={...})
await kg_builder.run_async(text='your text here')
Verify before relying
- Whether the package's experimental KG builder features are production-ready or intended for development/testing only
- Performance characteristics when working with large knowledge graphs or high-volume retrieval queries
- Whether APOC library installation in Neo4j is required for all features or only specific ones
Package facts
| License | Apache License, Version 2.0 permissive |
| Python support | Supports the current Python release <3.15,>=3.10.0 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 10 packagesfsspecjson-repairneo4jnumpypydanticpypdfpyyamlscipytenacitytypes-pyyaml |
| Maintenance | Actively maintained 51 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 440,023 / month, #6,649 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
Evidence: neo4j_graphrag-1.18.0-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “graph retrieval augmented generation”
- neo4j-graphragBuilds graph retrieval-augmented generation (GraphRAG) applications…
- lightrag-hkuLightRAG is a retrieval-augmented generation framework that builds…
- graphragGraphRAG extracts structured knowledge graphs from unstructured text…
Give your agent the search over MCP, or paste the wish link into any chat.
More Artificial Intelligence packages
LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.
Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.
Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.
Install it if you work with Hugging Face Hub models or datasets.
LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.
hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.
Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.
Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.
Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.
See also graphdatascience · graphrag · langchain-neo4j · neomodel · graph-retriever · ragstack-ai-knowledge-store · graphlib · lightrag-hku · graphiti-core · neo4j-driver