$npx skillfedfor your agent

faster-whisper

Faster Whisper transcription with CTranslate2

With conditionsPyPI Artificial IntelligenceReleased Oct 20259.0M downloads / moMITPure Python

Decision gist · record as of 2026-08-14

pure-Python wheel — faster_whisper-1.2.1-py3-none-any.whl
v1.2.1 · released 2025-10-31 · Python >=3.9 · 6 runtime deps: ctranslate2, huggingface-hub, tokenizers, onnxruntime, av, tqdm

Yes, if you need Whisper transcription and speed or memory efficiency matters. The package is mature (Beta status, 24910 stars, no known vulnerabilities), permissively licensed, and has low install friction. Maintenance is aging (287 days since last release), but the repository remains active. GPU users must have CUDA 12 libraries available; CPU-only use is simpler. Start here if you're choosing between openai/whisper and faster-whisper for the same accuracy at lower cost.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • GPU execution requires NVIDIA cuBLAS and cuDNN libraries for CUDA 12; CPU-only use works without them.
  • Python 3.9 or greater required.
  • Low friction install with a pure-Python wheel.

License · maintenance · safety

MIT (permissive) — MIT license is permissive; you can use this package freely in commercial and private projects with minimal restrictions.

last release 2025-10-31 (287 days) · last repo commit 2025-11-19 · 24,910 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 8,952,579 downloads/mo, #1,572 on PyPI

Verify before relying

pip install faster-whisper

from faster_whisper import WhisperModel

model = WhisperModel("large-v3", device="cuda", compute_type="float16")
segments, info = model.transcribe("audio.mp3", beam_size=5)
for segment in segments:
    print(f"[{segment.start:.2f}s -> {segment.end:.2f}s] {segment.text}")
  • Whether the 4x speedup claim applies to all model sizes and hardware configurations, or only specific benchmarked scenarios.
  • Real-world accuracy parity with openai/whisper across diverse audio types and languages.
  • Whether batched inference (batch_size parameter) is stable and recommended for production use.
Same gist for agents: .md · .json

What it is and what it does

Faster-whisper is a reimplementation of OpenAI's Whisper speech-to-text model using CTranslate2, a fast inference engine for Transformer models. It trades the original Whisper library for a more efficient backend, achieving measurable speed and memory improvements on both CPU and GPU while maintaining transcription accuracy. The package handles audio decoding via PyAV (bundled FFmpeg), so you don't need to install FFmpeg separately.

The package supports multiple precision modes (fp32, fp16, int8) and batch processing, allowing you to tune speed and memory trade-offs for your hardware. It can run on CPU or GPU, and works with Whisper's standard model sizes (tiny, base, small, medium, large) as well as Distil-Whisper checkpoints. The transcription API returns segments with timestamps and detected language information.

Use it for

  • Transcribe long audio or video files on GPU with reduced latency and VRAM compared to openai/whisper.
  • Run speech-to-text on CPU with int8 quantization to fit within constrained memory budgets.
  • Batch-process multiple audio files in parallel using the BatchedInferencePipeline for throughput.
  • Deploy Whisper transcription in production where inference speed and memory efficiency are critical.
  • Use Distil-Whisper checkpoints for faster, lighter-weight transcription with acceptable accuracy trade-off.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

With conditions

Yes, if you need Whisper transcription and speed or memory efficiency matters.

The package is mature (Beta status, 24910 stars, no known vulnerabilities), permissively licensed, and has low install friction. Maintenance is aging (287 days since last release), but the repository remains active. GPU users must have CUDA 12 libraries available; CPU-only use is simpler. Start here if you're choosing between openai/whisper and faster-whisper for the same accuracy at lower cost.

Install

faster-whisper on PyPI

Before you install

Low friction install with a pure-Python wheel. Maintenance is aging—last release was 287 days ago—but the repository remains active with recent commits and substantial community engagement (24910 stars). Six runtime dependencies are all established packages.

GPU execution requires NVIDIA cuBLAS and cuDNN libraries for CUDA 12; CPU-only use works without them. Python 3.9 or greater required.

License in practice

MIT license is permissive; you can use this package freely in commercial and private projects with minimal restrictions.

Quickstart

pip install faster-whisper

from faster_whisper import WhisperModel

model = WhisperModel("large-v3", device="cuda", compute_type="float16")
segments, info = model.transcribe("audio.mp3", beam_size=5)
for segment in segments:
    print(f"[{segment.start:.2f}s -> {segment.end:.2f}s] {segment.text}")

Verify before relying

  • Whether the 4x speedup claim applies to all model sizes and hardware configurations, or only specific benchmarked scenarios.
  • Real-world accuracy parity with openai/whisper across diverse audio types and languages.
  • Whether batched inference (batch_size parameter) is stable and recommended for production use.

Package facts

LicenseMIT permissive
Python supportSupports the current Python release >=3.9
Install frictionLow. Pure-Python wheel
Runtime dependencies
6 packages
ctranslate2huggingface-hubtokenizersonnxruntimeavtqdm
MaintenanceAging 287 days since the last release
Last repo commit
First released
Downloads8,952,579 / month, #1,572 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
Development Status :: 4 - BetaIntended Audience :: DevelopersIntended Audience :: Science/ResearchLicense :: OSI Approved :: MIT LicenseProgramming Language :: Python :: 3Programming Language :: Python :: 3 :: OnlyProgramming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.9Topic :: Scientific/Engineering :: Artificial Intelligence

Evidence: faster_whisper-1.2.1-py3-none-any.whl

Tags

Capabilities
speech to text transcriptionwhisper audio transcriptionfast speech recognitionaudio transcription inferencectranslate2 whisperquantized speech modelbatch audio transcription
Topics
speech-recognitionmodel-inferencequantization
PyPI keywords
openaiwhisperspeechctranslate2inferencequantizationtransformer

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “whisper audio transcription”

  • faster-whisperTranscribes audio to text using OpenAI's Whisper model, reimplemented…
  • mlx-whisperRuns OpenAI's Whisper speech recognition models on Apple silicon…
  • openai-whisperWhisper performs multilingual speech recognition, speech translation,…

Give your agent the search over MCP, or paste the wish link into any chat.

More Artificial Intelligence packages

litellm With conditions
PyPI · Artificial Intelligence · released Aug 2026

LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.

Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.

MITcompiled wheel
682.8Mdownloads / mo
huggingface-hub Worth it
PyPI · Artificial Intelligence · released Aug 2026

Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.

Install it if you work with Hugging Face Hub models or datasets.

Apache-2.0pure Python · 3.10.0+
442.4Mdownloads / mo
langchain Worth it
PyPI · Python Modules · released Aug 2026

LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.

MITpure Python
315.4Mdownloads / mo
hf-xet With conditions
PyPI · Artificial Intelligence · released Aug 2026

hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.

Apache-2.0compiled wheel · 3.8+
258.4Mdownloads / mo
tokenizers Worth it
PyPI · Artificial Intelligence · released Apr 2026

Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.

Apache-2.0compiled wheel · 3.10+
222.9Mdownloads / mo
transformers Worth it
PyPI · Artificial Intelligence · released Aug 2026

Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.

Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.

permissive licensepure Python · 3.10.0+
186.6Mdownloads / mo

See also ctranslate2 · openai-whisper · pywhispercpp · realtimestt · whisperx · whisper-timestamped · mlx-whisper · SpeechRecognition · omnivoice · pyctcdecode

Further reading