model2vec
Fast State-of-the-Art Static Embeddings
Decision gist · record as of 2026-08-14
Yes. Model2Vec is actively maintained, has no known vulnerabilities, uses permissive MIT licensing, and offers a clear value proposition: fast, small static embeddings with strong performance. Install if you need embedding inference speed and model compactness; the low dependency footprint and HuggingFace integration make it straightforward to adopt. Requires Python >=3.10.AI-flagged interpretation of the facts on this page — verify before relying
Before you install
- Requires Python >=3.10; pre-trained models are downloaded from HuggingFace hub on first use.
- Low friction: pure Python wheel with six common dependencies (numpy, jinja2, joblib, safetensors, tokenizers, tqdm).
- Active maintenance with recent releases; last commit 2026-08-13.
License · maintenance · safety
permissive license (permissive) — MIT License permits unrestricted use, modification, and distribution with only attribution required—no restrictions on commercial or proprietary use.
last release 2026-08-12 (2 days) · last repo commit 2026-08-13 · 2,176 stars
0 known vulnerabilities (OSV.dev, 2026-08-14) · 941,618 downloads/mo, #4,676 on PyPI
Alternatives
Verify before relying
pip install model2vec
from model2vec import StaticModel
model = StaticModel.from_pretrained("minishlab/potion-base-32M")
embeddings = model.encode(["It's dangerous to go alone!"])- Whether distillation (via model2vec[distill]) and training (via model2vec[train]) extras are included in the base install or require separate installation.
- Actual inference speed gains and embedding quality trade-offs compared to the original sentence transformer models in specific use cases.
What it is and what it does
Model2Vec is a distillation technique that transforms any sentence transformer into a compact static embedding model. It reduces model size by up to 50 times and achieves up to 500 times faster inference on CPU, with minimal performance loss. The package provides pre-trained models from HuggingFace (including multilingual variants) ready for immediate use, plus tools to distill your own models from existing sentence transformers in about 30 seconds without requiring a dataset.
The core workflow is straightforward: load a pre-trained Model2Vec model or distill one from a sentence transformer, then call encode() to generate sentence embeddings or encode_as_sequence() for token-level embeddings. These embeddings work for text classification, semantic search, clustering, and retrieval-augmented generation. The package integrates with HuggingFace hub for easy model sharing and is already integrated into Sentence Transformers and LangChain.
Use it for
- Build a semantic search or retrieval system where inference speed and model size are critical constraints.
- Distill a custom static embedding model from a sentence transformer in under a minute without training data.
- Fine-tune a classification model on top of a pre-trained Model2Vec embedding for text categorization tasks.
- Deploy embeddings in resource-constrained environments where model size and CPU inference speed matter.
- Generate multilingual embeddings for text in any of 101 languages using the potion-multilingual model.
Worth the install?
AI-flagged interpretation of the facts on this page. Verify before relying on it.
Yes.
Model2Vec is actively maintained, has no known vulnerabilities, uses permissive MIT licensing, and offers a clear value proposition: fast, small static embeddings with strong performance. Install if you need embedding inference speed and model compactness; the low dependency footprint and HuggingFace integration make it straightforward to adopt. Requires Python >=3.10.
Install
model2vec on PyPI
Before you install
Low friction: pure Python wheel with six common dependencies (numpy, jinja2, joblib, safetensors, tokenizers, tqdm). Active maintenance with recent releases; last commit 2026-08-13.
Requires Python >=3.10; pre-trained models are downloaded from HuggingFace hub on first use.
License in practice
MIT License permits unrestricted use, modification, and distribution with only attribution required—no restrictions on commercial or proprietary use.
Quickstart
pip install model2vec
from model2vec import StaticModel
model = StaticModel.from_pretrained("minishlab/potion-base-32M")
embeddings = model.encode(["It's dangerous to go alone!"])
Verify before relying
- Whether distillation (via model2vec[distill]) and training (via model2vec[train]) extras are included in the base install or require separate installation.
- Actual inference speed gains and embedding quality trade-offs compared to the original sentence transformer models in specific use cases.
Package facts
| License | permissive license permissive |
| Python support | Supports the current Python release >=3.10 |
| Install friction | Low. Pure-Python wheel |
| Runtime dependencies | 6 packagesjinja2joblibnumpysafetensorstokenizerstqdm |
| Maintenance | Actively maintained 2 days since the last release |
| Last repo commit | |
| First released | |
| Downloads | 941,618 / month, #4,676 on PyPI 30-day window, as of 2026-08-14 |
| Known vulnerabilities | None known OSV.dev, checked 2026-08-14 |
| Classifiers | Development Status :: 4 - BetaIntended Audience :: DevelopersIntended Audience :: Science/ResearchLicense :: OSI Approved :: MIT LicenseNatural Language :: EnglishProgramming Language :: Python :: 3 :: OnlyProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.12Programming Language :: Python :: 3.13Topic :: Scientific/Engineering :: Artificial IntelligenceTopic :: Software Development :: Libraries |
Evidence: model2vec-0.9.0-py3-none-any.whl
Tags
Let your AI agent find packages like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.
wish › “static text embeddings”
- model2vecModel2Vec converts sentence transformers into small, fast static…
- llama-index-embeddings-bedrockProvides Amazon Bedrock embedding models integration for LlamaIndex,…
- InstructorEmbeddingInstructorEmbedding generates task-specific text embeddings by…
Give your agent the search over MCP, or paste the wish link into any chat.
More Libraries packages
urllib3 is an HTTP client library that provides thread-safe connection pooling, SSL/TLS verification, multipart file uploads, request retries, compression support, and proxy handling for Python applications.
Requests is a Python HTTP library that simplifies sending HTTP/1.1 requests with automatic handling of headers, authentication, cookies, and response parsing.
Pluggy provides a plugin system that lets you define hook specifications and register implementations to be called in sequence, enabling extensible Python applications without tight coupling.
Install it if you're building an extensible application or framework.
Provides parsing, arithmetic, and recurrence rule computation for dates and times, with timezone support and iCalendar RFC compliance.
Install it if you need to parse flexible date strings, compute relative dates, handle timezones, or work with recurrence rules—it's the de facto choice for these tasks.
Six provides utility functions to write Python code that runs on both Python 2.7 and Python 3.3+, smoothing over language differences between the two versions.
pytest is a testing framework that lets you write test functions using plain assert statements and automatically discovers and runs them, with detailed failure reporting.
See also fastembed · sentence-transformers · transformer-smaller-training-vocab · InstructorEmbedding · pinecone-text · colpali-engine · setfit · antiberty · transformers · fasttext-numpy2