qianwen-video-generation
Generate videos through multiple input modes—text descriptions, single images, frame transitions, character role-play, or video editing—powered by Qianwen's Wan models. All operations run asynchronously; submit your request and poll for completion. The skill auto-detects your task and routes to the right model, with detailed reference guides for prompt engineering, polling patterns, and media workflows.
Qianwen Video Generation creates videos from text prompts, images, or frame pairs using Wan models with asynchronous processing.
AI-generated summary based on this skill's SKILL.md
Decision gist · record as of 2026-06-17
Qianwen Video Generation creates videos from text prompts, images, or frame pairs using Wan models with asynchronous processing. Generate videos through multiple input modes—text descriptions, single images, frame transitions, character role-play, or video editing—powered by Qianwen's Wan models. All operations run asynchronously; submit your request and poll for completion. The skill auto-detects your task and routes to the right model, with detailed reference guides for prompt engineering, polling patterns, and media workflows.
Use it when
- Yes.
- qianwen-video-generation can create smooth transitions and animations between two images by treating them as first and last frames.
Verify before relying
Read SKILL.md below before installing (15 files). Open directory: indexed for reading, not audited.
Install
QianWen-AI/qianwen-ai/qianwen-video-generation · repository language: Python
Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.
Frequently asked questions
AI-generated answers based on this skill's SKILL.md and metadata
How do I generate video from a text description using qianwen-video-generation?
qianwen-video-generation converts text prompts into videos through its primary text-to-video capability. Submit your text description as a request, and the skill routes it to the appropriate Wan model. Since all operations run asynchronously, you'll receive a request ID and must poll for completion. The skill includes detailed prompt engineering guides to help you craft descriptions that produce the best results.
Can qianwen-video-generation animate a static image into a moving video?
Yes. qianwen-video-generation supports converting static images into animated videos. You can upload a single image, and the skill will generate motion and animation around it. This is one of several input modes the skill auto-detects; it also handles frame transitions, text prompts, and video editing workflows all within the same platform.
What does qianwen-video-generation do with first frame and last frame video transitions?
qianwen-video-generation can create smooth transitions and animations between two images by treating them as first and last frames. The skill synthesizes the motion and content flow between these reference points. This capability is part of its multi-input approach, which also includes text-to-video, image animation, character role-play, and video repainting workflows.
Does qianwen-video-generation support video editing and recomposition with AI?
Yes. qianwen-video-generation includes video editing and repainting capabilities, allowing you to recompose and edit existing video content using AI. The skill auto-detects your task type and routes requests to the right model. All operations are asynchronous; submit your request and poll for completion using the reference guides provided for polling patterns and media workflows.
Can I generate character-driven videos with role-play scenarios in qianwen-video-generation?
qianwen-video-generation supports character-driven video generation with role-play scenarios. This is one of five primary input modes the skill handles, alongside text prompts, image animation, frame transitions, and video editing. The skill's auto-detection routes your request to the appropriate Wan model, and detailed guides cover prompt engineering for character-focused content.
What license does qianwen-video-generation use?
qianwen-video-generation is released under the Apache-2.0 license, allowing broad use, modification, and distribution under the terms of that open-source license.
SKILL.md
Rendered from the published skill. Quoted content, verbatim.
> Agent setup: If your agent doesn't auto-load skills (e.g. Claude Code), > see agent-compatibility.md once per session.
Qwen Video Generation
Generate videos using Wan models. All tasks are asynchronous — submit, then poll until completion. This skill is part of QianWen-AI/qianwen-ai.
> ⚠️ Critical Parameter Differences by Mode: > - kf2v (First+Last Frame): Duration is fixed at 5 seconds — other values will fail. Output is
(truncated - see the full file via the links below)
File tree — 15 files
skills/video/qianwen-video-generation/SKILL.md
skills/video/qianwen-video-generation/references/agent-compatibility.md
skills/video/qianwen-video-generation/references/api-guide.md
skills/video/qianwen-video-generation/references/examples.md
skills/video/qianwen-video-generation/references/execution-guide.md
skills/video/qianwen-video-generation/references/merge-media.md
skills/video/qianwen-video-generation/references/polling-guide.md
skills/video/qianwen-video-generation/references/prompt-guide.md
skills/video/qianwen-video-generation/references/request-fields.md
skills/video/qianwen-video-generation/references/sources.md
skills/video/qianwen-video-generation/references/workflows.md
skills/video/qianwen-video-generation/scripts/gossamer.py
skills/video/qianwen-video-generation/scripts/qianwen_lib.py
skills/video/qianwen-video-generation/scripts/video.py
skills/video/qianwen-video-generation/scripts/video_lib.py
Let your AI agent find skills like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.
wish › “Generate a video from a text prompt or description”
Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →
Related skills
Create videos asynchronously using QwenCloud's Wan models across multiple modes: text-to-video, image-to-video, first-and-last-frame transitions, role-play character animation, and video editing. Submit your request and poll for completion—no synchronous waiting required.
This skill wraps Aliyun's Wan video generation models through the DashScope SDK, enabling both text-to-video and image-to-video workflows. It standardizes video.generate requests with support for prompt control, duration, frame rate, resolution, seed, and motion parameters across multiple Wan model variants.
This skill walks you through QianWen API credential setup and validation. It handles both standard and Token Plan key types, manages environment variables and .env files, and provides verification steps to confirm your authentication is working correctly.
Wan Video lets you create videos programmatically through AceDataCloud's API, supporting text-to-video, image-to-video, and reference video transfer workflows. Choose from multiple models optimized for different generation types, with output resolutions from 480P to 1080P and optional audio generation.
Qwen Audio TTS turns written text into natural-sounding speech using Qwen's TTS engine. Choose from multiple voices and models—including the fast qwen3-tts-flash for standard tasks or instruction-guided variants for tone control—then output audio directly to file.
qianwen-vision lets you process images and videos through Qwen's vision models to extract text, understand visual content, and reason about complex scenes. Use it for OCR, chart/table analysis, multi-image comparison, and video comprehension across multiple model options tuned for speed, precision, or reasoning depth.
More skills qianwen-text (Apache-2.0) · aliyun-wan-i2v (MIT) · qianwen-image-generation (Apache-2.0) · aliyun-wan-videoedit (MIT) · Happyhorse (unlicensed)