micawber
a small library for extracting rich content from urls
Decision gist · record as of 2026-08-14
Yes. The package is actively maintained, has low install friction, no known vulnerabilities, and solves a real problem—extracting and embedding media metadata—with a minimal dependency footprint. The only caveat is that its license is not clearly documented; verify it in the repository before use in proprietary projects.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Python 3.8 or later.
- Low friction: pure Python wheel with a single runtime dependency (beautifulsoup4).
- Actively maintained as of 2026-07-05, with recent commits and 680 repository stars.
License · maintenance · safety
(unclear) — License status is unclear—no SPDX identifier or raw license text is recorded. Verify the actual license in the repository before using in proprietary or copyleft-sensitive projects.
last release 2026-07-05 (40 days) · last repo commit 2026-07-19 · 680 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 205,590 downloads/mo, #9,587 on PyPI
Alternatives
Verify before relying
pip install micawber
import micawber
providers = micawber.bootstrap_basic()
result = providers.request('http://www.youtube.com/watch?v=54XHDUOHuzU')
print(result['title'], result['thumbnail_url'])- Which providers (YouTube, Flickr, etc.) are included in bootstrap_basic() and whether additional providers can be registered.
- Whether the library handles rate limiting or caching for repeated requests to the same URL.
- How the library behaves when a URL is unreachable or does not return oEmbed metadata.
What it is and what it does
Micawber is a small library for fetching and embedding rich media metadata from URLs. It wraps the oEmbed standard, allowing you to query URLs (especially video and image links) and receive structured metadata like title, author, thumbnail, and embed HTML. The library ships with built-in provider rules for common platforms and can parse blocks of text or HTML, automatically replacing bare links with embedded content.
You typically use it by bootstrapping a provider registry, then either requesting metadata for a single URL or parsing a larger text block to replace all recognized links with embeds. It depends only on beautifulsoup4 for HTML parsing and requires Python 3.8+.
Use it for
- Fetch metadata (title, thumbnail, embed code) for a YouTube or Flickr link in a blog or content management system.
- Parse a user-submitted text block and automatically replace bare video URLs with embedded players.
- Build a link preview feature that shows rich metadata (author, title, image) before a user clicks.
- Extract structured data from media URLs for indexing or display in a feed or gallery.
- Convert raw HTML containing media links into HTML with embedded content widgets.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes.
The package is actively maintained, has low install friction, no known vulnerabilities, and solves a real problem—extracting and embedding media metadata—with a minimal dependency footprint. The only caveat is that its license is not clearly documented; verify it in the repository before use in proprietary projects.
Install
micawber on PyPI
Before you install
Low friction: pure Python wheel with a single runtime dependency (beautifulsoup4). Actively maintained as of 2026-07-05, with recent commits and 680 repository stars.
Requires Python 3.8 or later.
License in practice
License status is unclear—no SPDX identifier or raw license text is recorded. Verify the actual license in the repository before using in proprietary or copyleft-sensitive projects.
Quickstart
pip install micawber
import micawber
providers = micawber.bootstrap_basic()
result = providers.request('http://www.youtube.com/watch?v=54XHDUOHuzU')
print(result['title'], result['thumbnail_url'])
Verify before relying
- Which providers (YouTube, Flickr, etc.) are included in bootstrap_basic() and whether additional providers can be registered.
- Whether the library handles rate limiting or caching for repeated requests to the same URL.
- How the library behaves when a URL is unreachable or does not return oEmbed metadata.
Package facts
| License | Not declared unclear |
| Python support | Supports the current Python release >=3.8 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 1 packagebeautifulsoup4 |
| Maintenance | Actively maintained 40 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 205,590 / month, #9,587 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Programming Language :: Python :: 3Topic :: Software Development :: Libraries :: Python Modules |
Evidence: micawber-0.7.0-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “extract metadata from urls”
- micawberExtracts rich metadata (title, author, thumbnail, embed HTML) from…
- git-url-parseParses Git repository URLs to extract components like host, owner,…
- faviconFetches and parses a website's favicon from its HTML, returning icon…
Give your agent the search over MCP, or paste the wish link into any chat.
More Python Modules packages
Converts domain names between Unicode and ASCII-compatible encoding (Punycode) according to IDNA 2008 and Unicode Technical Standard 46, with security validation and broader script coverage than the standard library.
Install it if you work with internationalized domain names, need to validate domains, or use HTTP clients that depend on it transitively.
Setuptools is a Python build backend and package management tool that handles building, distributing, and installing Python packages, including support for C/C++ extension modules.
PyYAML parses and emits YAML 1.1 data format, enabling serialization and deserialization of configuration files and Python objects to and from human-readable YAML text.
Pydantic validates Python data structures against type hints, coercing and checking input at runtime to ensure it matches a declared schema.
Provides reusable metadata objects for use with PEP-593 `typing.Annotated` to express common constraints like bounds, collection sizes, and predicates on types.
Install it if you use or build libraries that need to express type constraints in a standardized, inspectable way—or if you want to annotate your own types with…
Provides runtime tools to inspect and introspect Python type annotations, enabling programmatic examination of type hints at execution time.
See also django-embed-video · mkdocs-video · goose3 · sphinxcontrib-youtube · pytubefix · sphinxcontrib-video · youtube_dl · shortzy · hbreader · ubi-reader