pdfrw
PDF file reader/writer library
Decision gist · record as of 2026-08-14
Yes, if you need lightweight, dependency-free PDF reading and basic manipulation in Python and can accept that the package is no longer actively developed. It works well for straightforward tasks like merging, subsetting, and metadata changes. No, if you require support for modern Python versions beyond 3.6, active maintenance, or handling of complex or recently encrypted PDFs.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Installation is frictionless with no runtime dependencies.
- The package is dormant (last release 2017-09-18, last commit 2024-04-29), so while the repository shows ongoing maintenance, it receives no active development.
License · maintenance · safety
MIT (permissive) — MIT license permits commercial and private use with minimal restrictions, making it suitable for most projects without licensing concerns.
last release 2017-09-18 (3252 days) · last repo commit 2024-04-29 · 1,911 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 2,543,756 downloads/mo, #3,009 on PyPI
Alternatives
Verify before relying
pip install pdfrw
from pdfrw import PdfReader, PdfWriter
reader = PdfReader('input.pdf')
writer = PdfWriter()
writer.addpages(reader.pages)
writer.write('output.pdf')- Whether the package works reliably with Python versions beyond 3.6, given its dormant status.
- Current compatibility with contemporary PDF formats and encryption standards.
- Performance characteristics for large-scale PDF processing compared to alternatives.
What it is and what it does
pdfrw is a pure-Python PDF library that reads and writes PDF files without requiring compiled dependencies or external tools. It handles common PDF manipulation tasks—merging multiple PDFs, extracting or reordering pages, rotating content, modifying metadata, and embedding one PDF into another—while preserving vector graphics without rasterization. The library can work standalone or integrate with reportlab to reuse existing PDF content in newly generated documents.
The package comes with command-line examples demonstrating practical workflows: creating booklets, extracting images, adding watermarks, creating posters, and concatenating files. It has been used in production pre-press environments and is noted as the fastest pure-Python PDF parser available. However, the project is dormant—last released 2017-09-18 and tested only through Python 3.6—so it may not support newer Python versions or handle recently introduced PDF features without modification.
Use it for
- Merge multiple PDFs into a single document or extract specific pages from an existing PDF.
- Rotate pages or reorganize them into booklet layouts for printing without specialized software.
- Add or modify PDF metadata (title, author, subject) programmatically without altering document structure.
- Extract images and embedded pages (Form XObjects) from existing PDFs to reuse in new documents.
- Generate watermarked PDFs by overlaying one PDF image on top of another.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes, if you need lightweight, dependency-free PDF reading and basic manipulation in Python and can accept that the package is no longer actively developed.
It works well for straightforward tasks like merging, subsetting, and metadata changes. No, if you require support for modern Python versions beyond 3.6, active maintenance, or handling of complex or recently encrypted PDFs.
Install
pdfrw on PyPI
Before you install
Installation is frictionless with no runtime dependencies. The package is dormant (last release 2017-09-18, last commit 2024-04-29), so while the repository shows ongoing maintenance, it receives no active development.
License in practice
MIT license permits commercial and private use with minimal restrictions, making it suitable for most projects without licensing concerns.
Quickstart
pip install pdfrw
from pdfrw import PdfReader, PdfWriter
reader = PdfReader('input.pdf')
writer = PdfWriter()
writer.addpages(reader.pages)
writer.write('output.pdf')
Verify before relying
- Whether the package works reliably with Python versions beyond 3.6, given its dormant status.
- Current compatibility with contemporary PDF formats and encryption standards.
- Performance characteristics for large-scale PDF processing compared to alternatives.
Package facts
| License | MIT permissive |
| Python support | Not specified |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | None |
| Maintenance | Dormant 3,252 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 2,543,756 / month, #3,009 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 4 - BetaIntended Audience :: DevelopersLicense :: OSI Approved :: MIT LicenseOperating System :: OS IndependentProgramming Language :: PythonProgramming Language :: Python :: 2Programming Language :: Python :: 2.6Programming Language :: Python :: 2.7Programming Language :: Python :: 3Programming Language :: Python :: 3.3Programming Language :: Python :: 3.4Programming Language :: Python :: 3.5Programming Language :: Python :: 3.6Topic :: Multimedia :: Graphics :: Graphics ConversionTopic :: PrintingTopic :: Software Development :: LibrariesTopic :: Text ProcessingTopic :: Utilities |
Evidence: pdfrw-0.4-py2.py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “pdf reader writer library”
- pdfrwpdfrw reads, writes, and manipulates PDF files in pure Python,…
- aspose-cells-pythonAspose.Cells for Python via .NET creates, reads, and manipulates…
- pdfrw2pdfrw2 reads, writes, and manipulates PDF files with operations…
Give your agent the search over MCP, or paste the wish link into any chat.
More Libraries packages
urllib3 is an HTTP client library that provides thread-safe connection pooling, SSL/TLS verification, multipart file uploads, request retries, compression support, and proxy handling for Python applications.
Requests is a Python HTTP library that simplifies sending HTTP/1.1 requests with automatic handling of headers, authentication, cookies, and response parsing.
Pluggy provides a plugin system that lets you define hook specifications and register implementations to be called in sequence, enabling extensible Python applications without tight coupling.
Install it if you're building an extensible application or framework.
Provides parsing, arithmetic, and recurrence rule computation for dates and times, with timezone support and iCalendar RFC compliance.
Install it if you need to parse flexible date strings, compute relative dates, handle timezones, or work with recurrence rules—it's the de facto choice for these tasks.
Six provides utility functions to write Python code that runs on both Python 2.7 and Python 3.3+, smoothing over language differences between the two versions.
pytest is a testing framework that lets you write test functions using plain assert statements and automatically discovers and runs them, with detailed failure reporting.
See also pdfrw2 · pypdftk · pikepdf · playa-pdf · fillpdf · PyPDF2 · PyPDF3 · PyPDF4 · PyPDFForm · gotenberg-client