video-gen
Create AI-powered videos from text descriptions, images, or reference materials using three distinct model families. Choose HappyHorse for versatile text-to-video and editing, SeeDance for frame-based transitions, or PixVerse for specialized features like lip-sync, motion transfer, and marketing templates.
video-gen creates AI videos from text prompts using HappyHorse, SeeDance, or PixVerse models.
AI-generated summary based on this skill's SKILL.md
Decision gist · record as of 2026-07-16
video-gen creates AI videos from text prompts using HappyHorse, SeeDance, or PixVerse models. Create AI-powered videos from text descriptions, images, or reference materials using three distinct model families. Choose HappyHorse for versatile text-to-video and editing, SeeDance for frame-based transitions, or PixVerse for specialized features like lip-sync, motion transfer, and marketing templates.
Use it when
- Yes, video-gen supports converting static images into animated videos.
- video-gen enables AI-powered editing of existing videos, including style changes, background modifications, and effects application.
Verify before relying
Read SKILL.md below before installing (4 files). Open directory: indexed for reading, not audited.
Similar skills
Install
marswaveai/skills/video-gen · repository language: HTML
Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.
Frequently asked questions
AI-generated answers based on this skill's SKILL.md and metadata
How do I generate video from text with video-gen?
video-gen lets you create AI-powered videos from text prompts using three model families. HappyHorse excels at versatile text-to-video generation and editing tasks. SeeDance specializes in frame-based transitions, while PixVerse offers specialized features like lip-sync and motion transfer. Select your preferred model based on your creative needs.
Can video-gen animate a static image to video?
Yes, video-gen supports converting static images into animated videos. You can enhance the animation by providing optional reference materials to guide the output. This capability works across multiple model families, giving you flexibility in how you approach image-to-video creation.
What video editing capabilities does video-gen offer?
video-gen enables AI-powered editing of existing videos, including style changes, background modifications, and effects application. HappyHorse is particularly strong for general editing tasks, while PixVerse provides advanced options like motion transfer and lip-sync adjustments for more specialized editing needs.
Does video-gen support lip sync and motion transfer?
Yes, video-gen's PixVerse model family specializes in lip-sync and motion transfer capabilities. These features allow you to synchronize audio with video and transfer motion between clips, making PixVerse ideal for creating professional-quality videos with precise audio-visual alignment.
Can I create promotional videos with video-gen?
video-gen supports marketing and promotional video creation through multi-image fusion capabilities. You can combine multiple images with AI enhancement to produce polished promotional content. This feature is particularly useful for brands and creators looking to generate marketing materials efficiently.
SKILL.md
Rendered from the published skill. Quoted content, verbatim.
When to Use
- User wants to generate an AI video from a text description
- User wants to animate a still image (first-frame)
- User has reference images to guide video generation
- User wants to edit an existing video (change style, background, etc.)
- User wants to lip-sync a video to audio or TTS (PixVerse only) — "对口型", "口型同步"
- User wants a marketing ad / promo mix video (PixVerse agent)
- User says "生成视频", "做视频", "video generation", "text to video", "视频编辑", "pixverse", "口型"
When NOT to Use
- User wants an explainer video with narration and AI visuals (use
/explainer) - User wants to transcribe audio/video to text (use
/asr) - User wants to generate an image (use
/image-gen)
Purpose
Generate AI videos using the ListenHub CLI. Supports three model
(truncated - see the full file via the links below)
File tree — 4 files
video-gen/SKILL.md
video-gen/references/happyhorse-api.md
video-gen/references/pixverse-api.md
video-gen/shared
Let your AI agent find skills like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.
wish › “Generate AI video from text prompt using HappyHorse, SeeDance, or PixVerse”
Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →
Related skills
Happyhorse Video lets you create and modify videos through multiple input methods—text descriptions, single images, or up to 9 reference images—all via the AceDataCloud API. Choose from text-to-video generation, first-frame animation, reference-guided creation, or editing existing footage with optional style guidance. Output ranges from 720P to 1080P at 3–15 seconds, with controls for aspect ratio, audio preservation, and reproducible seeds.
Generate Video turns text descriptions into short videos, optionally anchored by still images as opening or closing frames. It supports text-to-video generation, image-to-video animation, and frame-to-frame morphing, with configurable resolution, duration, aspect ratio, and reproducible output via seed control.
HappyHorse 1.0 is a video generation and editing skill powered by Alibaba's models, accessible via the inference.sh CLI. Create physically realistic videos from text prompts, animate still images, preserve characters across multiple reference photos, or edit existing footage with natural language instructions—all at 720P or 1080P up to 15 seconds.
Create AI-generated images by describing what you want. Select your preferred model, resolution, and aspect ratio, then optionally add reference images for style guidance. The skill handles the generation and saves your images locally.
Create short videos from text descriptions or still images by routing to your choice of 10 providers—OpenAI Sora, Google Veo, Runway, Pika, Luma, and others—all through a single command-line interface. Supports text-to-video, image-to-video, and optional last-frame control where available, with configurable duration, aspect ratio, and resolution.
Create videos from text descriptions or images using ByteDance Seedance models via the Volcengine Ark API. Choose from multiple model variants optimized for speed, quality, or specific use cases like text-to-video or image-to-video generation. A Python CLI tool handles task creation, polling, and downloads automatically.
More skills Video (unlicensed)