{"enrichment":{"faq":[{"a":"Whisper supports speech to text transcription across 99 languages, making it one of the most comprehensive multilingual speech recognition models available. This broad language coverage enables users to transcribe audio content from virtually any region and convert it to text accurately, regardless of the source language.","q":"What languages does Whisper support for speech to text transcription?"},{"a":"Yes, Whisper is specifically designed with noise robustness as a core feature. It converts speech to text with high accuracy even when processing noisy recordings, making it ideal for real-world scenarios like podcast transcription, meeting recordings, and other audio sources with background noise or imperfect recording conditions.","q":"Can Whisper convert audio to text with high accuracy in noisy environments?"},{"a":"Whisper offers multiple model sizes optimized for different use cases. The available variants range from lightweight models for faster processing to high-accuracy versions for demanding applications. Setup is straightforward, and you can choose the model size based on your accuracy requirements and computational resources. Whisper also includes a turbo model variant designed for improved speed.","q":"How do I set up the Whisper ASR model and what are the available variants?"},{"a":"Yes, Whisper can translate foreign language audio to English automatically. Beyond transcription in the original language, it provides built-in translation capabilities, allowing you to convert multilingual speech directly into English text without requiring separate translation tools.","q":"Does Whisper automatically translate speech to English from other languages?"},{"a":"Whisper can generate subtitles or timestamps from media files, enabling you to create subtitle tracks in formats like SRT. This functionality makes it useful for video content creators who need to add captions with precise timing information to their videos or generate searchable transcripts.","q":"Can Whisper generate subtitles with timestamps from video files?"},{"a":"Yes, Whisper supports GPU acceleration with CUDA for efficient batch processing of large audio files. This GPU support significantly speeds up transcription when processing multiple audio files, making it practical for handling large-scale transcription projects while maintaining accuracy across the entire batch.","q":"Does Whisper support GPU acceleration for batch processing audio files?"}],"shadow_tags":["audio-processing","language-agnostic","open-source-ml","batch-capable","hardware-accelerated","subtitle-generation","voice-to-text","model-selection"],"summary_rewrite":"Whisper is OpenAI's multilingual speech recognition model for converting audio and video into text across 99 languages. It handles noisy recordings, supports translation to English, and offers multiple model sizes from lightweight to high-accuracy variants. Use it for podcasts, meeting transcription, video subtitles, and multilingual audio processing."},"files":[{"bytes":7244,"path":"optional-skills/mlops/whisper/SKILL.md","sha256":"4db9296c11967aef85a6fadcefe21b751db06bbdf25f1654900704dcce90284f","url":"https://skillfed.io/files/NousResearch/hermes-agent/whisper/4ba6ce5b/SKILL.md"}],"id":"NousResearch/hermes-agent/whisper","links":{"html":"https://skillfed.io/NousResearch/hermes-agent/whisper","md":"https://skillfed.io/NousResearch/hermes-agent/whisper.md","repo":"https://github.com/NousResearch/hermes-agent"},"meta":{"agents_supported":[],"first_seen":"2026-07-28","forks":42317,"language":"Python","last_updated":"2026-07-28","license":"MIT","name":"whisper","publisher":"NousResearch","stars":221503},"relations":{"similar":[{"id":"synthetic-sciences/openscience/whisper"},{"id":"Orchestra-Research/AI-Research-SKILLs/whisper"},{"id":"OpenLAIR/dr-claw/whisper"},{"id":"graniet/kheish/whisper"},{"id":"ThePlasmak/faster-whisper/faster-whisper"},{"id":"jamditis/claude-skills-journalism/video-transcribe"},{"id":"claude-office-skills/skills/transcription-automation"},{"id":"chubbyguan/chubbyskills/podcast-transcribe"},{"id":"different-ai/agent-bank/video-subtitle-cutter"},{"id":"datadrivenconstruction/DDC_Skills_for_AI_Agents_in_Construction/voice-to-report"}]},"slug":{"owner":"NousResearch","repo":"hermes-agent","skill":"whisper"},"version":"4ba6ce5b"}
