Context To Video
Context To Video transforms any text source—URLs, pasted content, PR diffs, or meeting notes—into a polished MP4 with synchronized narration, slides, and subtitles. The skill uses a free local stack (Pillow, edge-tts, ffmpeg) and optionally adds a talking-head avatar via SadTalker. Output lands in your project workspace as MP4 plus SRT subtitle file.
Context To Video converts any written content—articles, notes, or code diffs—into narrated MP4 videos with slides and subtitles.
AI-generated summary based on this skill's SKILL.md
Decision gist · record as of 2026-07-19
Context To Video converts any written content—articles, notes, or code diffs—into narrated MP4 videos with slides and subtitles. Context To Video transforms any text source—URLs, pasted content, PR diffs, or meeting notes—into a polished MP4 with synchronized narration, slides, and subtitles. The skill uses a free local stack (Pillow, edge-tts, ffmpeg) and optionally adds a talking-head avatar via SadTalker. Output lands in your project workspace as MP4 plus SRT subtitle file.
Use it when
- Yes.
- Context To Video accepts multiple input formats including URLs, pasted text content, PR diffs, and meeting notes.
Install
aktsmm/Agent-Skills/context-to-video · repository language: Python
generated, unverified - the skill's exact subdirectory could not be determined; check the repository on GitHub
Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.
Frequently asked questions
AI-generated answers based on this skill's SKILL.md and metadata
What does Context To Video do?
Context To Video transforms any text source—URLs, pasted content, PR diffs, or meeting notes—into a polished MP4 with synchronized narration, slides, and subtitles. The skill uses a free local stack (Pillow, edge-tts, ffmpeg) and optionally adds a talking-head avatar via SadTalker. Output lands in your project workspace as MP4 plus SRT subtitle file.
Can I convert written content to video automatically?
Yes. Context To Video converts written context or text content into video format automatically without requiring manual video editing. Simply provide your text—whether as a URL, pasted content, PR diff, or meeting notes—and the skill generates a complete MP4 with narration, slides, and subtitles.
What input formats does Context To Video accept?
Context To Video accepts multiple input formats including URLs, pasted text content, PR diffs, and meeting notes. You can feed it any written material, and the skill will process it into video output.
What output files does Context To Video produce?
Context To Video produces two output files: an MP4 video file with synchronized narration, slides, and subtitles, plus an SRT subtitle file. Both files land in your project workspace for easy access and distribution.
Does Context To Video require paid software?
No. Context To Video uses a free local stack built on Pillow, edge-tts, and ffmpeg. You can optionally add a talking-head avatar via SadTalker, but the core video generation requires no paid tools or subscriptions.
Can I repurpose written content as video with this skill?
Yes. Context To Video is designed to repurpose written content as video for broader distribution. Transform text-based material into engaging video output that reaches audiences who prefer video over reading.
Let your AI agent find skills like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.
wish › “Convert written context or text content into video format automatically”
Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →
Related skills
Motion Explainer transforms a topic into a complete short-form video with voiceover and captions. The pipeline handles research, scripting, audio generation, keyframe design, and video assembly—defaulting to the gflow backend for cost-efficient production. Choose between Mixed Media collage or paper-diorama cinematography styles.
Build academic presentation decks from research papers with full control over outline and visuals. The skill handles script drafting, slide generation via nanobanana, optional text-to-speech narration, and video assembly with ffmpeg—you direct the structure and emphasis throughout.
Video Watch processes video URLs and local files to generate artifacts that GitHub Copilot can analyze, including captions, sampled frames, contact sheets, and a prompt packet. Use it when you need Copilot to inspect, summarize, or diagnose video content like demos, recordings, or tutorials. The skill handles multiple detail modes to balance speed and visual coverage.
Feishu Voice TTS transforms text into speech using edge-tts and delivers it as Feishu audio messages, bypassing the platform's text fallback for direct file sends. The skill handles transcoding to Opus format, file upload, and message delivery through Feishu's open API.
This integrated skill bundles six specialized capabilities for producing teaching media: image generation, technical article illustrations, subject infographics, animated lessons with voiceover, short-form video covers, and article-to-video conversion. Route requests to the appropriate sub-skill, or chain multiple capabilities into complete production pipelines for knowledge points or long-form content.
This skill transforms individual subject concepts into two synchronized animated formats rendered by HyperFrames: a narrated teaching video (~90s, 1080p with Chinese voiceover and subtitles) and a silent looping animation (~32s, 1080p for embedding). Both outputs share a single HTML structure, unified storyboard, and thematic color palette across nine subject domains.
More skills wjs-segmenting-video (MIT) · Local Media Transcription (NOASSERTION)