skillfed

ai-avatar-video

Route your audio and image inputs across RunComfy's avatar models to generate talking-head videos with natural lip-sync and gestures. The skill picks the right model for your intent—whether you need a photoreal presenter, stylized character animation, or cinematic multi-modal composition—and provides the exact CLI command and prompting pattern for each.

AI Avatar & Talking Head Video creates talking-head videos by syncing portraits or characters to audio voiceovers via RunComfy.

AI-generated summary based on this skill's SKILL.md

31 9 MIT updated by prime-skills

Install

prime-skills/runcomfy-agent-skills/ai-avatar-video

git clone https://github.com/prime-skills/runcomfy-agent-skills
cp -r runcomfy-agent-skills/ai-avatar-video ~/.claude/skills/ai-avatar-video
npx skillfed install prime-skills/runcomfy-agent-skills/ai-avatar-video

Frequently asked questions

AI-generated answers based on this skill's SKILL.md and metadata

How do I create a talking head video from audio with ai-avatar-video?

ai-avatar-video routes your audio and portrait image across RunComfy's avatar models to generate talking-head videos with natural lip-sync and gestures. Provide your audio file and a reference image, and the skill picks the right model for your intent—whether photoreal or stylized—then outputs the exact CLI command and prompting pattern you need to run.

Can ai-avatar-video generate an avatar video directly from a script?

Yes. ai-avatar-video supports generating avatar videos from written scripts without pre-recorded audio. The skill routes your script across available models, selects the best fit for your needs, and provides the CLI command and prompting guidance to synthesize speech and animate your avatar in one workflow.

What's the difference between ai-avatar-video and alternatives like HeyGen?

ai-avatar-video is MIT-licensed and runs on RunComfy's open avatar models, giving you full control over model selection and prompting. It routes across multiple avatar architectures to match your intent—photoreal presenter, stylized character, or cinematic composition—and outputs reproducible CLI commands rather than a closed SaaS interface.

Can I use ai-avatar-video to lip-sync a character to audio?

Yes. ai-avatar-video specializes in audio-driven lip-sync for both photoreal portraits and stylized characters. Provide your audio file and character image, and the skill routes to the appropriate model, handles the sync, and delivers the command and prompting pattern to animate gestures and mouth movement together.

How does ai-avatar-video choose which avatar model to use?

ai-avatar-video analyzes your intent—whether you need a cinematic multi-modal shot, a simple talking head, UGC-style animation, or stylized character work—and routes to the best-fit model from RunComfy's avatar suite. It then provides the exact CLI command and prompting strategy for that model, ensuring optimal results for your use case.

What inputs does ai-avatar-video need to produce a video?

ai-avatar-video accepts audio files, portrait or character images, and optional reference materials for cinematic shots. You can also provide a written script instead of pre-recorded audio. The skill routes these inputs to the right model, then outputs the CLI command and full prompting pattern to generate your talking-head or avatar video.

SKILL.md

rendered from the published skill — quoted content, verbatim

AI Avatar & Talking Head Video

Put words in a

(truncated - see the full file via the links below)

Read as markdown · JSON record · Browse the source repository

File tree — 1 file
ai-avatar-video/SKILL.md

Related skills

Tags

digital-human mouth-sync voiceover-sync synthetic-media video-generation character-animation speech-driven multimodal-composition presenter-creation dubbed-content