baoyu-imagine
baoyu-imagine generates images from text prompts across 10+ AI providers including OpenAI GPT Image 2, Google, Azure OpenAI, and others. It supports reference images for identity preservation, batch generation, custom aspect ratios, and quality presets, with flexible configuration via local or user-home settings.
baoyu-imagine generates images from text prompts using OpenAI, Google, Azure, and multiple other AI providers.
AI-generated summary based on this skill's SKILL.md
Decision gist · record as of 2026-07-28
baoyu-imagine generates images from text prompts using OpenAI, Google, Azure, and multiple other AI providers. baoyu-imagine generates images from text prompts across 10+ AI providers including OpenAI GPT Image 2, Google, Azure OpenAI, and others. It supports reference images for identity preservation, batch generation, custom aspect ratios, and quality presets, with flexible configuration via local or user-home settings.
Use it when
- Yes.
- baoyu-imagine creates multiple images in parallel batch mode for efficiency.
Verify before relying
Read SKILL.md below before installing (36 files). Open directory: indexed for reading, not audited.
Install
guanyang/open-agent-hub/baoyu-imagine · repository language: TypeScript
Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.
Frequently asked questions
AI-generated answers based on this skill's SKILL.md and metadata
How do I generate an image from text with baoyu-imagine?
baoyu-imagine generates images from text prompts using multiple AI providers including OpenAI GPT Image 2, Google, Azure OpenAI, and 10+ others. Simply provide your text prompt and select your preferred provider; the tool handles the API calls and returns your generated image.
Can baoyu-imagine create images with reference photos?
Yes. baoyu-imagine supports reference images to preserve identity during generation. You can provide a reference photo alongside your text prompt, and the tool will use it to guide the image generation process across supported providers.
Does baoyu-imagine support batch image generation?
baoyu-imagine creates multiple images in parallel batch mode for efficiency. This allows you to generate several images at once rather than one at a time, significantly speeding up workflows when you need multiple variations or outputs.
Which AI providers does baoyu-imagine work with?
baoyu-imagine supports 10+ AI image providers including OpenAI, Google, Azure OpenAI, Replicate, DashScope, and others. You can switch between providers based on your needs, API availability, and preferences.
Can I control image quality and aspect ratio in baoyu-imagine?
Yes. baoyu-imagine lets you control image quality, size, and aspect ratio during generation. You can set custom aspect ratios and quality presets through flexible configuration via local or user-home settings.
What license does baoyu-imagine use?
baoyu-imagine is released under the MIT license, allowing free use, modification, and distribution with minimal restrictions.
SKILL.md
Rendered from the published skill. Quoted content, verbatim.
Image Generation (AI SDK)
Official API-based image generation. Supports OpenAI GPT Image 2, Azure OpenAI, Google, OpenRouter, DashScope (阿里通义万象), Z.AI GLM-Image, MiniMax, Jimeng (即梦), Seedream (豆包) and Replicate.
User Input Tools
When this skill prompts the user, follow this tool-selection rule (priority order):
- Prefer built-in user-input tools exposed by the current agent runtime — e.g.,
AskUserQuestion,request_user_input,clarify,ask_user, or any equivalent. - Fallback: if no such tool exists, emit a numbered
(truncated - see the full file via the links below)
File tree — 15 files
skills/baoyu-imagine/SKILL.md
skills/baoyu-imagine/references/codex-image2-fallback.md
skills/baoyu-imagine/references/codex-oauth-vs-openai-api-key.md
skills/baoyu-imagine/references/config/first-time-setup.md
skills/baoyu-imagine/references/config/preferences-schema.md
skills/baoyu-imagine/references/providers/dashscope.md
skills/baoyu-imagine/references/providers/minimax.md
skills/baoyu-imagine/references/providers/openrouter.md
skills/baoyu-imagine/references/providers/replicate.md
skills/baoyu-imagine/references/providers/zai.md
skills/baoyu-imagine/references/usage-examples.md
skills/baoyu-imagine/scripts/build-batch.test.ts
skills/baoyu-imagine/scripts/build-batch.ts
skills/baoyu-imagine/scripts/main.test.ts
skills/baoyu-imagine/scripts/main.ts
Let your AI agent find skills like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.
wish › “Generate images from text prompts using multiple AI providers”
Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →
Related skills
Generate images from text prompts using your choice of 11+ AI providers including OpenAI GPT Image 2, Google, Azure, and DashScope. Supports reference images for identity preservation, custom aspect ratios, batch processing, and prompt files. Configure your default provider and model once, then generate single or multiple images with flexible quality and size options.
Tuzi Image Gen creates images from text prompts across multiple AI providers, with Tuzi as the default. Customize output with aspect ratios, quality presets, reference images, and model selection—supporting synchronous and async generation workflows.
Baoyu Image Gen creates images from text prompts using multiple AI providers—OpenAI, Google, DashScope, and Replicate. Choose your preferred provider and model, set aspect ratios and quality levels, and optionally reference existing images to guide generation. Sequential processing is the default; parallel generation is available on request.
Canghe Image Gen creates images from text prompts across multiple AI providers including OpenAI, Google, DashScope, and Canghe. Configure your preferred provider, quality level, and aspect ratio, with support for reference images and batch generation when needed.
happy-audio-gen synthesizes natural speech from any text across six major TTS providers through a single interface. Route here whenever users ask to read text aloud, create narration, dub scripts, or generate voice-overs—the skill auto-detects available credentials and handles long-form content by chunking transparently. Output formats include MP3, WAV, OGG, and FLAC.
Create AI-generated images by describing what you want. Select your preferred model, resolution, and aspect ratio, then optionally add reference images for style guidance. The skill handles the generation and saves your images locally.
More skills happy-image-gen (MIT) · ai-image-generation (MIT) · aliyun-wan-image (MIT) · baoyu-compress-image (MIT) · Baoyu Image Gen (unlicensed)