$npx skillfedfor your agent

aliyun-wan-i2v

Aliyun Wan I2V taps DashScope's Wan 2.7 model to synthesize videos from static images asynchronously. It handles single-image generation, frame interpolation between two images, video extension from clips, and audio-driven animation with lip-sync capabilities.

Aliyun Wan I2V generates videos from images using DashScope's Wan 2.7 model with async processing.

AI-generated summary based on this skill's SKILL.md

397 34 MITupdated by cinience

Decision gist · record as of 2026-07-18

Aliyun Wan I2V generates videos from images using DashScope's Wan 2.7 model with async processing. Aliyun Wan I2V taps DashScope's Wan 2.7 model to synthesize videos from static images asynchronously. It handles single-image generation, frame interpolation between two images, video extension from clips, and audio-driven animation with lip-sync capabilities.

manual: git clone https://github.com/cinience/alicloud-skills → cp -r alicloud-skills/skills/ai/video/aliyun-wan-i2v ~/.claude/skills/aliyun-wan-i2v
skills/ai/video/aliyun-wan-i2v/SKILL.md · version 973bf490

Use it when

  • Aliyun Wan I2V converts images to video by submitting requests to the DashScope async video synthesis API.
  • Yes, aliyun-wan-i2v supports audio-driven video synthesis with lip-sync capabilities.

Verify before relying

Read SKILL.md below before installing (4 files). Open directory: indexed for reading, not audited.

Same gist for agents: .md · .json

Install

cinience/alicloud-skills/aliyun-wan-i2v · repository language: Python

Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.

Frequently asked questions

AI-generated answers based on this skill's SKILL.md and metadata

What can aliyun-wan-i2v do with image to video generation?

Aliyun Wan I2V leverages DashScope's Wan 2.7 model to synthesize videos from static images asynchronously. The skill generates full video sequences from a single image, interpolates frames between two images, extends existing video clips, and creates audio-driven animations with lip-sync capabilities for dynamic content creation.

How does aliyun-wan-i2v convert image to video?

Aliyun Wan I2V converts images to video by submitting requests to the DashScope async video synthesis API. You provide a static image, and the Wan 2.7 model generates a video sequence. The skill supports single-image generation, frame interpolation between multiple images, and video extension from existing clips.

Can aliyun-wan-i2v create audio-driven video with lip-sync?

Yes, aliyun-wan-i2v supports audio-driven video synthesis with lip-sync capabilities. Beyond static image-to-video generation, the skill can animate content synchronized to audio input, enabling you to create videos where motion and lip movements align with provided audio tracks.

What is the license for aliyun-wan-i2v?

Aliyun Wan I2V is released under the MIT license, allowing free use, modification, and distribution with minimal restrictions.

How does aliyun-wan-i2v handle video interpolation between images?

Aliyun Wan I2V interpolates video between multiple frames by using the DashScope Wan 2.7 model to generate smooth transitions. The skill can extend video clips and create intermediate frames, enabling seamless video continuation and frame-to-frame animation between static images.

What parameters can I configure in aliyun-wan-i2v?

Aliyun Wan I2V allows configuration of video generation parameters through the DashScope async API. You can manage settings for image input, interpolation modes, video extension options, audio-sync parameters, and other synthesis controls to customize the generated video output.

SKILL.md

Rendered from the published skill. Quoted content, verbatim.

Wan 2.7 Image-to-Video

Validation

mkdir -p output/aliyun-wan-i2v
python -m py_compile skills/ai/video/aliyun-wan-i2v/scripts/generate_i2v.py && echo "py_compile_ok" > output/aliyun-wan-i2v/validate.txt

Pass criteria: command exits 0 and output/aliyun-wan-i2v/validate.txt is generated.

Output And Evidence

  • Save task IDs, polling responses, and final video URLs to output/aliyun-wan-i2v/.
  • Keep at least one end-to-end run log for troubleshooting.

Prerequisites

  • Install SDK (recommended in a venv):
python3 -m venv .venv
. .venv/bin/activate
python -m pip install dashscope
  • Set DASHSCOPE_API_KEY in your environment, or add dashscope_api_key to ~/.alibabacloud/credentials.

Critical model names

  • wan2.7-i2v — supports first-frame, first+last frame, video continuation, and audio-driven generation

Capabilities

|

(truncated - see the full file via the links below)

File tree — 4 files
skills/ai/video/aliyun-wan-i2v/SKILL.md
skills/ai/video/aliyun-wan-i2v/references/api_reference.md
skills/ai/video/aliyun-wan-i2v/references/sources.md
skills/ai/video/aliyun-wan-i2v/scripts/generate_i2v.py

Let your AI agent find skills like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.

wish › “Generate video from static image using AI model”

Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →

Related skills

aliyun-wan-videoedit
by cinience · cinience/alicloud-skills

This skill harnesses Alibaba's Wan 2.7 video editing model to transform video appearance through style transfer (clay, anime, etc.) or content-aware edits guided by text prompts and optional reference images. It handles async task creation, polling, and media validation across multiple aspect ratios and resolutions.

MITupdated Jul 2026
★ 397repo stars
aliyun-happyhorse-r2v
by cinience · cinience/alicloud-skills

This skill generates videos by blending multiple reference images into a single output, with prompts that map each image to a character placeholder (character1, character2, etc.). Supports 1–9 images per video, customizable resolution, aspect ratio, and duration via the DashScope async API.

MITupdated Jul 2026
★ 397repo stars
aliyun-video-style-repaint
by cinience · cinience/alicloud-skills

Aliyun Video Style Repaint applies one of eight preset artistic styles to your videos via Alibaba's DashScope API. Choose from Japanese manga, American comics, 3D cartoon, Chinese ink painting, paper art, and other effects, then submit your video for async processing and download the styled result.

MITupdated Jul 2026
★ 397repo stars
aliyun-happyhorse-t2v
by cinience · cinience/alicloud-skills

This skill wraps Alibaba Cloud's HappyHorse 1.0 text-to-video model through the async DashScope video-synthesis API. Submit a text prompt and configure resolution (720P or 1080P), aspect ratio, duration (3–15 seconds), and optional seed for reproducibility; the skill polls for task completion and returns your generated MP4 video URL.

MITupdated Jul 2026
★ 397repo stars
qianwen-video-generation
by QianWen-AI · QianWen-AI/qianwen-ai

Generate videos through multiple input modes—text descriptions, single images, frame transitions, character role-play, or video editing—powered by Qianwen's Wan models. All operations run asynchronously; submit your request and poll for completion. The skill auto-detects your task and routes to the right model, with detailed reference guides for prompt engineering, polling patterns, and media workflows.

Apache-2.0updated Jun 2026
★ 38repo stars
aliyun-zimage-turbo
by cinience · cinience/alicloud-skills

Z-Image Turbo enables fast text-to-image generation through Alibaba Cloud's DashScope multimodal API. Control output dimensions, randomization seed, and optional prompt enhancement while managing costs based on your feature selection.

MITupdated Jul 2026
★ 397repo stars

More skills aliyun-wan-r2v (MIT) · Dashscope (AGPL-3.0) · aliyun-vidu-video (MIT) · Image To Video (unlicensed) · aliyun-wan-video (MIT)

Tags
video-synthesisframe-interpolationaudio-driven-videolip-sync-generationasync-apimedia-input-handlingvideo-continuationprompt-rewriting