$npx skillfedfor your agent

vocos

Fourier-based neural vocoder for high-quality audio synthesis

With conditionsPyPI Artificial IntelligenceReleased Oct 2023423.6K downloads / moPure Python

Decision gist · record as of 2026-08-14

pure-Python wheel — vocos-0.1.0-py3-none-any.whl
v0.1.0 · released 2023-10-14 · 8 runtime deps: torch, torchaudio, numpy, scipy, einops, pyyaml, huggingface-hub, encodec

Yes, if you need a fast neural vocoder for mel-spectrogram or EnCodec token-to-audio synthesis and can work with a dormant codebase. Low install friction and zero known vulnerabilities make it practical for inference. Verify the license status directly in the repository before commercial use, and be aware that maintenance has stalled since 2023-10-14—expect no active support or updates.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • Requires torch and torchaudio; pre-trained models are downloaded from huggingface-hub on first use.
  • Low install friction with a pure-Python wheel.
  • Maintenance is dormant—last release was 2023-10-14 and last commit 2024-08-07—but the repository remains unarchived with moderate popularity (1150 stars).

License · maintenance · safety

(unclear) — License treatment is unclear; the repository states MIT in its LICENSE file, but the PyPI metadata does not declare it formally. Verify the LICENSE file directly before relying on the package in a commercial or license-sensitive context.

last release 2023-10-14 (1035 days) · last repo commit 2024-08-07 · 1,150 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 423,607 downloads/mo, #6,771 on PyPI

Verify before relying

pip install vocos

import torch
from vocos import Vocos

vocos = Vocos.from_pretrained("charactr/vocos-mel-24khz")
mel = torch.randn(1, 100, 256)  # B, C, T
audio = vocos.decode(mel)
  • Whether the MIT license statement in the repository's LICENSE file is the authoritative license for the PyPI package.
  • Whether the package is actively maintained or if dormancy signals a shift to a successor or fork.
  • Compatibility with modern PyTorch and torchaudio versions beyond what the fact sheet specifies.
Same gist for agents: .md · .json

What it is and what it does

Vocos is a neural vocoder—a machine learning model that converts acoustic feature representations into audio waveforms. Unlike traditional vocoders that work in the time domain, Vocos generates spectral coefficients and reconstructs audio via inverse Fourier transform, enabling fast single-pass synthesis. It is trained using a GAN objective and can accept either mel-spectrograms or EnCodec tokens as input, making it suitable for integration into text-to-speech pipelines or audio processing workflows.

The package includes pre-trained models for 24 kHz audio synthesis and supports both inference and training modes. It depends on torch, torchaudio, numpy, scipy, einops, pyyaml, huggingface-hub, and encodec. Installation is straightforward, though the dormant maintenance status (last release 2023-10-14) means bug fixes and feature updates are not actively rolling out.

Use it for

  • Convert mel-spectrograms from a text-to-speech model into high-quality audio waveforms for end-to-end TTS synthesis.
  • Reconstruct audio from EnCodec-compressed tokens at various bandwidth levels for codec-based audio processing.
  • Perform copy-synthesis by resampling an audio file to 24 kHz and reconstructing it through the vocoder.
  • Integrate with text-to-audio models as a replacement vocoder for faster or higher-quality audio generation.
  • Train a custom vocoder on domain-specific audio data using the provided training pipeline and configuration framework.

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

With conditions

Yes, if you need a fast neural vocoder for mel-spectrogram or EnCodec token-to-audio synthesis and can work with a dormant codebase.

Low install friction and zero known vulnerabilities make it practical for inference. Verify the license status directly in the repository before commercial use, and be aware that maintenance has stalled since 2023-10-14—expect no active support or updates.

Install

vocos on PyPI

Before you install

Low install friction with a pure-Python wheel. Maintenance is dormant—last release was 2023-10-14 and last commit 2024-08-07—but the repository remains unarchived with moderate popularity (1150 stars).

Requires torch and torchaudio; pre-trained models are downloaded from huggingface-hub on first use.

License in practice

License treatment is unclear; the repository states MIT in its LICENSE file, but the PyPI metadata does not declare it formally. Verify the LICENSE file directly before relying on the package in a commercial or license-sensitive context.

Quickstart

pip install vocos

import torch
from vocos import Vocos

vocos = Vocos.from_pretrained("charactr/vocos-mel-24khz")
mel = torch.randn(1, 100, 256)  # B, C, T
audio = vocos.decode(mel)

Verify before relying

  • Whether the MIT license statement in the repository's LICENSE file is the authoritative license for the PyPI package.
  • Whether the package is actively maintained or if dormancy signals a shift to a successor or fork.
  • Compatibility with modern PyTorch and torchaudio versions beyond what the fact sheet specifies.

Package facts

LicenseNot declared unclear
Python supportNot specified
Install frictionLow. Pure-Python wheel
Runtime dependencies
8 packages
torchtorchaudionumpyscipyeinopspyyamlhuggingface-hubencodec
MaintenanceDormant 1,035 days since the last release
Last repo commit
First released
Downloads423,607 / month, #6,771 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14

Evidence: vocos-0.1.0-py3-none-any.whl

Tags

Capabilities
neural vocoder audio synthesismel-spectrogram to audioGAN-based vocoderaudio waveform generationencodec token to audiofourier-based vocoderfast audio reconstruction
Topics
audio-synthesisneural-vocodergan-model

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “neural vocoder audio synthesis”

  • vocosVocos is a neural vocoder that synthesizes audio waveforms from…
  • pyworldPyWorld wraps the WORLD vocoder to decompose speech audio into pitch,…
  • TTSTTS is a deep learning library for text-to-speech synthesis that…

Give your agent the search over MCP, or paste the wish link into any chat.

More Artificial Intelligence packages

litellm With conditions
PyPI · Artificial Intelligence · released Aug 2026

LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.

Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.

MITcompiled wheel
682.8Mdownloads / mo
huggingface-hub Worth it
PyPI · Artificial Intelligence · released Aug 2026

Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.

Install it if you work with Hugging Face Hub models or datasets.

Apache-2.0pure Python · 3.10.0+
442.4Mdownloads / mo
langchain Worth it
PyPI · Python Modules · released Aug 2026

LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.

MITpure Python
315.4Mdownloads / mo
hf-xet With conditions
PyPI · Artificial Intelligence · released Aug 2026

hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.

Apache-2.0compiled wheel · 3.8+
258.4Mdownloads / mo
tokenizers Worth it
PyPI · Artificial Intelligence · released Apr 2026

Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.

Apache-2.0compiled wheel · 3.10+
222.9Mdownloads / mo
transformers Worth it
PyPI · Artificial Intelligence · released Aug 2026

Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.

Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.

permissive licensepure Python · 3.10.0+
186.6Mdownloads / mo

See also encodec · pyworld · f5-tts · snac · aot-biomaps · Gammatone · TTS · descript-audio-codec · coqui-tts · chatterbox-tts