$npx skillfedfor your agent

ai-avatar-video

Route your audio and image inputs across RunComfy's avatar models to generate talking-head videos with natural lip-sync and gestures. The skill picks the right model for your intent—whether you need a photoreal presenter, stylized character animation, or cinematic multi-modal composition—and provides the exact CLI command and prompting pattern for each.

AI Avatar & Talking Head Video creates talking-head videos by syncing portraits or characters to audio voiceovers via RunComfy.

AI-generated summary based on this skill's SKILL.md

★ 31  9 MITupdated by prime-skills

Decision gist · record as of 2026-05-15

AI Avatar & Talking Head Video creates talking-head videos by syncing portraits or characters to audio voiceovers via RunComfy. Route your audio and image inputs across RunComfy's avatar models to generate talking-head videos with natural lip-sync and gestures. The skill picks the right model for your intent—whether you need a photoreal presenter, stylized character animation, or cinematic multi-modal composition—and provides the exact CLI command and prompting pattern for each.

manual: git clone https://github.com/prime-skills/runcomfy-agent-skills → cp -r runcomfy-agent-skills/ai-avatar-video ~/.claude/skills/ai-avatar-video
ai-avatar-video/SKILL.md · version cda90345

Use it when

  • Yes.
  • ai-avatar-video is MIT-licensed and runs on RunComfy's open avatar models, giving you full control over model selection and prompting.

Verify before relying

Read SKILL.md below before installing (1 file). Open directory: indexed for reading, not audited.

Same gist for agents: .md · .json

Install

prime-skills/runcomfy-agent-skills/ai-avatar-video

Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.

Frequently asked questions

AI-generated answers based on this skill's SKILL.md and metadata

How do I create a talking head video from audio with ai-avatar-video?

ai-avatar-video routes your audio and portrait image across RunComfy's avatar models to generate talking-head videos with natural lip-sync and gestures. Provide your audio file and a reference image, and the skill picks the right model for your intent—whether photoreal or stylized—then outputs the exact CLI command and prompting pattern you need to run.

Can ai-avatar-video generate an avatar video directly from a script?

Yes. ai-avatar-video supports generating avatar videos from written scripts without pre-recorded audio. The skill routes your script across available models, selects the best fit for your needs, and provides the CLI command and prompting guidance to synthesize speech and animate your avatar in one workflow.

What's the difference between ai-avatar-video and alternatives like HeyGen?

ai-avatar-video is MIT-licensed and runs on RunComfy's open avatar models, giving you full control over model selection and prompting. It routes across multiple avatar architectures to match your intent—photoreal presenter, stylized character, or cinematic composition—and outputs reproducible CLI commands rather than a closed SaaS interface.

Can I use ai-avatar-video to lip-sync a character to audio?

Yes. ai-avatar-video specializes in audio-driven lip-sync for both photoreal portraits and stylized characters. Provide your audio file and character image, and the skill routes to the appropriate model, handles the sync, and delivers the command and prompting pattern to animate gestures and mouth movement together.

How does ai-avatar-video choose which avatar model to use?

ai-avatar-video analyzes your intent—whether you need a cinematic multi-modal shot, a simple talking head, UGC-style animation, or stylized character work—and routes to the best-fit model from RunComfy's avatar suite. It then provides the exact CLI command and prompting strategy for that model, ensuring optimal results for your use case.

What inputs does ai-avatar-video need to produce a video?

ai-avatar-video accepts audio files, portrait or character images, and optional reference materials for cinematic shots. You can also provide a written script instead of pre-recorded audio. The skill routes these inputs to the right model, then outputs the CLI command and full prompting pattern to generate your talking-head or avatar video.

SKILL.md

Rendered from the published skill. Quoted content, verbatim.

AI Avatar & Talking Head Video

Put words in a

(truncated - see the full file via the links below)

File tree — 1 file
ai-avatar-video/SKILL.md

Let your AI agent find skills like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.

wish › “Create a talking-head video by syncing a portrait or character to an audio voiceover”

Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →

Related skills

ai-video-generation
by prime-skills · prime-skills/runcomfy-agent-skills

Route text and image prompts to the right AI video model via RunComfy's CLI. The skill handles model selection across HappyHorse, Kling, Seedance, Veo, Wan, and others—each optimized for different outputs like in-pass audio, multi-shot character consistency, or physics-accurate motion. Get the exact `runcomfy run` command and prompt patterns for your use case.

MITupdated May 2026
★ 31repo stars
Dreamina Video
by AceDataCloud · AceDataCloud/Skills

Dreamina Video animates static portraits into speaking digital humans by pairing images with audio tracks, producing lip-synced output powered by ByteDance OmniHuman 1.5. The skill supports mask-based targeting for multi-person photos and async task polling to handle longer processing jobs without timeout.

no license declared → metadata onlyupdated Jul 2026
★ 13repo stars
ai-image-generation
by prime-skills · prime-skills/runcomfy-agent-skills

Create and edit images using RunComfy's CLI with access to over a dozen AI models including FLUX 2, Google Nano Banana, OpenAI GPT Image 2, ByteDance Seedream, and others. The skill intelligently routes your request to the right model based on your intent—whether you need photoreal portraits, fast iteration, precise typography, or open-weights workflows—and provides the exact command to run.

MITupdated May 2026
★ 31repo stars
runcomfy-cli
by prime-skills · prime-skills/runcomfy-agent-skills

RunComfy CLI is a single binary that connects you to hundreds of AI model endpoints—image generation, video creation, editing, face swap, lip-sync, and more—all from the command line. Install once, authenticate once, then invoke any model with `runcomfy run` and JSON inputs. This skill teaches installation, login, model discovery, invocation modes (sync, poll, no-wait), JSON scripting, and error handling.

MITupdated May 2026
★ 31repo stars
happy-video-gen
by iamzhihuix · iamzhihuix/happy-claude-skills

Create short videos from text descriptions or still images by routing to your choice of 10 providers—OpenAI Sora, Google Veo, Runway, Pika, Luma, and others—all through a single command-line interface. Supports text-to-video, image-to-video, and optional last-frame control where available, with configurable duration, aspect ratio, and resolution.

MITupdated Apr 2026
★ 305repo stars
seedance-ai-avatar
by rediumvex · rediumvex/ai-video-generator-claude

This skill architects detailed Seedance 2.0 prompts for avatar-driven video content, handling everything from photorealistic digital humans to stylized 3D characters and abstract entities. It covers character definition, uncanny valley avoidance, hook patterns, environment design, and technical specs to produce broadcast-ready prompts for spokesperson videos, product demos, and virtual influencer content.

MITupdated Jun 2026
★ 299repo stars
Tags
digital-humanmouth-syncvoiceover-syncsynthetic-mediavideo-generationcharacter-animationspeech-drivenmultimodal-compositionpresenter-creationdubbed-content