qwencloud-image-generation
Create images from text descriptions or edit existing ones with Wan and Qwen Image models. This skill handles text-to-image generation, style transfer, subject consistency across reference images, and interleaved text-image output for tutorials and guides.
qwencloud-image-generation creates images from text prompts and edits existing images using Wan and Qwen models.
AI-generated summary based on this skill's SKILL.md
Decision gist · record as of 2026-07-22
qwencloud-image-generation creates images from text prompts and edits existing images using Wan and Qwen models. Create images from text descriptions or edit existing ones with Wan and Qwen Image models. This skill handles text-to-image generation, style transfer, subject consistency across reference images, and interleaved text-image output for tutorials and guides.
Use it when
- Yes.
- qwencloud-image-generation excels at creating product images, illustrations, posters, and artistic designs.
Verify before relying
Read SKILL.md below before installing (10 files). Open directory: indexed for reading, not audited.
Install
QwenCloud/qwencloud-ai/qwencloud-image-generation · repository language: Python
Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.
Frequently asked questions
AI-generated answers based on this skill's SKILL.md and metadata
How does qwencloud-image-generation generate images from text prompts?
qwencloud-image-generation uses advanced AI models to synthesize images directly from your text descriptions. You provide a detailed prompt describing what you want to create—whether it's a product photo, illustration, poster, or artistic design—and the skill generates a corresponding image using Wan and Qwen Image models. The generated images maintain visual coherence and follow your specifications.
Can I edit photos with style transfer using this skill?
Yes. qwencloud-image-generation supports image editing through style transfer and artistic transformation. You can apply different artistic styles to existing images or transform them while maintaining subject consistency. The skill preserves key elements of your original image while applying the desired stylistic changes you specify.
What types of designs can qwencloud-image-generation create?
qwencloud-image-generation excels at creating product images, illustrations, posters, and artistic designs. Whether you need marketing materials, creative artwork, or professional product photography, the skill can generate high-quality images tailored to your needs. It's designed to handle diverse creative use cases across commercial and artistic domains.
Does qwencloud-image-generation support interleaved text and image output?
Yes. qwencloud-image-generation can generate interleaved text-image content, making it ideal for creating tutorials, guides, and educational materials. You can combine generated or edited images with accompanying text to produce comprehensive visual documentation and instructional content.
Can I use reference photos to maintain subject consistency?
qwencloud-image-generation supports subject consistency image generation using reference photos. This allows you to create variations or transformations of images while keeping key subjects consistent across multiple outputs. It's particularly useful for generating product variations or maintaining character consistency across multiple generated scenes.
What's the license for qwencloud-image-generation?
qwencloud-image-generation is released under the Apache-2.0 license, allowing you to use, modify, and distribute the skill freely while maintaining appropriate attribution and license notices.
SKILL.md
Rendered from the published skill. Quoted content, verbatim.
> Agent setup: If your agent doesn't auto-load skills (e.g. Claude Code), > see agent-compatibility.md once per session.
Qwen Image Generation
Generate and edit images using Wan and Qwen Image models. Supports text-to-image,
(truncated - see the full file via the links below)
File tree — 10 files
skills/image/qwencloud-image-generation/SKILL.md
skills/image/qwencloud-image-generation/references/agent-compatibility.md
skills/image/qwencloud-image-generation/references/api-guide.md
skills/image/qwencloud-image-generation/references/execution-guide.md
skills/image/qwencloud-image-generation/references/prompt-guide.md
skills/image/qwencloud-image-generation/references/sources.md
skills/image/qwencloud-image-generation/scripts/gossamer.py
skills/image/qwencloud-image-generation/scripts/image.py
skills/image/qwencloud-image-generation/scripts/image_lib.py
skills/image/qwencloud-image-generation/scripts/qwencloud_lib.py
Let your AI agent find skills like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.
wish › “Generate images from text descriptions using AI models”
Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →
Related skills
Create images from text descriptions or edit existing ones using Wan and Qwen Image models. Supports style transfer, subject consistency across multiple reference images, text rendering in images, and interleaved text-image output for tutorials and guides. Choose from multiple models optimized for different tasks—from fast drafts to high-resolution 4K generation.
Aliyun Wan Image lets you create and modify images using DashScope's Wan 2.7 models via text prompts, image editing, interactive region selection, or multi-image sequences. Supports output up to 4K on the professional model and includes color palette customization.
This skill harnesses Alibaba Cloud DashScope to create, modify, and interpret images through multiple AI models. It handles text-to-image generation, image editing with local or remote files, and visual content analysis—routing each task to the appropriate model and script automatically.
Access Qwen's language models for text generation, multi-turn conversations, code writing, and function calling through an OpenAI-compatible interface. The skill supports multiple Qwen variants optimized for different tasks—from general-purpose models to specialized code and reasoning versions—with flexible model selection and streaming output.
Qwen Vision lets you understand images and videos through Qwen's specialized VL and QVQ models. Extract text via OCR, analyze charts and tables, perform visual reasoning, and compare multiple images—all with built-in support for thinking mode and high-resolution processing.
Turn written text into high-quality spoken audio using QwenCloud's TTS models. Choose between fast standard synthesis (qwen3-tts-flash) or instruction-guided style control (qwen3-tts-instruct-flash), or opt for premium quality via CosyVoice. Select from multiple voices and languages to match your content needs.
More skills aliyun-qwen-image (MIT) · qwencloud-video-generation (Apache-2.0) · web-asset-generator (MIT)