{"enrichment":{"faq":[{"a":"Media Transcription runs an event-driven pipeline to convert audio and video media files into searchable text transcripts. It uses MLX Whisper for speech recognition and pyannote for speaker identification, automatically extracting and processing spoken content from your media files to generate accessible text versions.","q":"What does Media Transcription do?"},{"a":"Yes. Media Transcription transcribes video to text by processing media stored on the NAS. It handles video files like MP4s alongside audio formats, extracting speech and converting it into searchable transcripts with speaker identification included.","q":"Can I transcribe video to text with Media Transcription?"},{"a":"Media Transcription lets you monitor progress in real time through its orchestration interface. You can cancel active runs or resume from partial completions, giving you full control over the transcription pipeline without losing work if a process is interrupted.","q":"How do I monitor and control transcription runs?"},{"a":"Media Transcription processes media files stored on the NAS, including MP3 audio files, MP4 video files, and other common formats. The pipeline automatically handles format conversion as needed during the transcription workflow.","q":"What audio and video formats does Media Transcription support?"},{"a":"Media Transcription uses Inngest orchestration with detached local inference processes that run independently of the main pipeline. This architecture prevents timeout failures by allowing speech recognition and speaker identification to complete without being constrained by request timeouts.","q":"How does Media Transcription prevent timeout failures?"},{"a":"Media Transcription execution is limited to the Flagg host worker, which has direct access to the media mount on the NAS. This ensures reliable, local processing of your media files with proper file system permissions and consistent performance.","q":"Where does Media Transcription run and what access does it need?"}],"shadow_tags":["speech-to-text","audio-conversion","video-processing","content-accessibility","automated-transcription","media-extraction","voice-recognition","transcript-generation"],"summary_rewrite":"Media Transcription runs a durable, event-driven pipeline to transcribe meeting media stored on the NAS, using MLX Whisper for speech recognition and pyannote for speaker identification. Monitor progress in real time, cancel runs, or resume from partial completions\u2014all orchestrated through Inngest with detached local inference processes that prevent timeout failures. Execution is limited to the Flagg host worker with direct access to the media mount."},"gist":{"api_url":"https://skillfed.io/api/skills/joelhooks/joelclaw/media-transcription.json","as_of":"2026-07-27","description":"Media Transcription converts meeting audio and video files into searchable text transcripts using durable.","install":{"manual":["git clone https://github.com/joelhooks/joelclaw","cp -r joelclaw ~/.claude/skills/media-transcription"],"primary":"npx skillfed install joelhooks/joelclaw/media-transcription","version":"4039062c"},"kind":"skill","mirror_url":"https://skillfed.io/joelhooks/joelclaw/media-transcription.md","similar":[{"id":"joelhooks/joelclaw/wzrrd-video","name":"Wzrrd Video","publisher":"joelhooks/joelclaw","url":"https://skillfed.io/joelhooks/joelclaw/wzrrd-video"}],"title":"Media Transcription by joelhooks \u2014 SkillFed","use":{"when":["Yes.","Media Transcription lets you monitor progress in real time through its orchestration interface."]},"what":{"lead":"Media Transcription converts meeting audio and video files into searchable text transcripts using durable, event-driven processing.","rest":"Media Transcription runs a durable, event-driven pipeline to transcribe meeting media stored on the NAS, using MLX Whisper for speech recognition and pyannote for speaker identification. Monitor progress in real time, cancel runs, or resume from partial completions\u2014all orchestrated through Inngest with detached local inference processes that prevent timeout failures. Execution is limited to the Flagg host worker with direct access to the media mount."}},"id":"joelhooks/joelclaw/media-transcription","install":{"mode":"external","repo":"https://github.com/joelhooks/joelclaw"},"links":{"html":"https://skillfed.io/joelhooks/joelclaw/media-transcription","md":"https://skillfed.io/joelhooks/joelclaw/media-transcription.md","repo":"https://github.com/joelhooks/joelclaw"},"meta":{"agents_supported":[],"first_seen":"2026-07-28","forks":3,"language":"TypeScript","last_updated":"2026-07-27","license":null,"name":"Media Transcription","publisher":"joelhooks","stars":61},"relations":{"similar":[{"id":"joelhooks/joelclaw/system-architecture"},{"id":"joelhooks/joelclaw/joelclaw-system-check"},{"id":"joelhooks/joelclaw/wzrrd-video"},{"id":"joelhooks/joelclaw/satellite-rig"},{"id":"joelhooks/joelclaw/three-body"},{"id":"joelhooks/joelclaw/pi-inngest"},{"id":"joelhooks/joelclaw/agent-session-capture-backup"},{"id":"joelhooks/joelclaw/session-search"},{"id":"joelhooks/joelclaw/monitor"},{"id":"joelhooks/joelclaw/memory-system"}]},"slug":{"owner":"joelhooks","repo":"joelclaw","skill":"media-transcription"},"version":"4039062c"}
