happy-image-gen
happy-image-gen unifies image generation across eight providers—OpenAI, Google, Replicate, Stability AI, FAL, Ark, Bailian, and SiliconFlow—under a single command-line interface. Create still images from text prompts or transform existing images with reference-driven edits. The skill auto-detects available API keys and respects your configuration defaults, so you can switch providers without rewriting commands.
happy-image-gen generates still images from text prompts using OpenAI, Google, Replicate, Stability AI, and four other providers.
AI-generated summary based on this skill's SKILL.md
Decision gist · record as of 2026-04-20
happy-image-gen generates still images from text prompts using OpenAI, Google, Replicate, Stability AI, and four other providers. happy-image-gen unifies image generation across eight providers—OpenAI, Google, Replicate, Stability AI, FAL, Ark, Bailian, and SiliconFlow—under a single command-line interface. Create still images from text prompts or transform existing images with reference-driven edits. The skill auto-detects available API keys and respects your configuration defaults, so you can switch providers without rewriting commands.
Use it when
- Yes.
- happy-image-gen supports transforming or restyling existing images using reference-driven edits.
Verify before relying
Read SKILL.md below before installing (20 files). Open directory: indexed for reading, not audited.
Install
iamzhihuix/happy-claude-skills/happy-image-gen · repository language: TypeScript
Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.
Frequently asked questions
AI-generated answers based on this skill's SKILL.md and metadata
How do I generate an image of a cat with happy-image-gen?
happy-image-gen lets you create still images from text prompts across eight providers—OpenAI, Google, Replicate, Stability AI, FAL, Ark, Bailian, and SiliconFlow. Simply provide your text description (e.g., "a cat") and the skill generates the image using your configured provider. The tool auto-detects available API keys, so if you've set up credentials, you can start generating immediately without additional setup.
Can I create image using DALL-E or other specific models?
Yes. happy-image-gen supports creating images with specific providers and models including DALL-E, Flux, SDXL, and others. You can specify which provider or model you want to use in your command, and the skill will route your request accordingly. The tool unifies all eight providers under a single interface, so you can switch between them without rewriting your workflow.
Does happy-image-gen support image-to-image editing?
happy-image-gen supports transforming or restyling existing images using reference-driven edits. You can upload an existing image and provide instructions or reference styles to apply, allowing you to modify photos and artwork with AI-powered transformations across supported providers.
How do I configure API keys for multiple image generation services?
happy-image-gen auto-detects available API keys from your environment and respects your configuration defaults. Set up credentials for the providers you want to use (OpenAI, Google, Replicate, Stability AI, FAL, Ark, Bailian, or SiliconFlow), and the skill will recognize them automatically. You can then switch between providers without managing keys manually in each command.
What custom options does happy-image-gen offer for rendering?
happy-image-gen lets you render images with custom aspect ratios and quality settings. These options allow you to fine-tune output dimensions and fidelity according to your needs, giving you control over the final image characteristics beyond just the text prompt.
Which AI providers does happy-image-gen support?
happy-image-gen unifies eight providers under one interface: OpenAI, Google, Replicate, Stability AI, FAL, Ark, Bailian, and SiliconFlow. This multi-provider approach lets you generate images from text prompts or transform existing images, choosing the provider that best fits your use case or preference.
SKILL.md
Rendered from the published skill. Quoted content, verbatim.
happy-image-gen
Generates still images across 8 providers through one CLI: bun scripts/main.ts .... The same CLI handles text-to-image and image-to-image (reference-driven) edits.
Quick usage
```bash bun
(truncated - see the full file via the links below)
File tree — 15 files
skills/happy-image-gen/SKILL.md
skills/happy-image-gen/assets/EXTEND.template.md
skills/happy-image-gen/package.json
skills/happy-image-gen/references/aspect_ratio_map.md
skills/happy-image-gen/references/config/extend-schema.md
skills/happy-image-gen/references/config/first-time-setup.md
skills/happy-image-gen/references/error_codes.md
skills/happy-image-gen/references/providers.md
skills/happy-image-gen/scripts/main.ts
skills/happy-image-gen/scripts/providers/ark.ts
skills/happy-image-gen/scripts/providers/bailian.ts
skills/happy-image-gen/scripts/providers/fal.ts
skills/happy-image-gen/scripts/providers/google.ts
skills/happy-image-gen/scripts/providers/openai.ts
skills/happy-image-gen/scripts/providers/replicate.ts
Let your AI agent find skills like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.
wish › “Generate still images from text prompts across multiple AI providers”
Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →
Related skills
Create short videos from text descriptions or still images by routing to your choice of 10 providers—OpenAI Sora, Google Veo, Runway, Pika, Luma, and others—all through a single command-line interface. Supports text-to-video, image-to-video, and optional last-frame control where available, with configurable duration, aspect ratio, and resolution.
happy-audio-gen synthesizes natural speech from any text across six major TTS providers through a single interface. Route here whenever users ask to read text aloud, create narration, dub scripts, or generate voice-overs—the skill auto-detects available credentials and handles long-form content by chunking transparently. Output formats include MP3, WAV, OGG, and FLAC.
Create AI-generated images by describing what you want. Select your preferred model, resolution, and aspect ratio, then optionally add reference images for style guidance. The skill handles the generation and saves your images locally.
Image Generation produces images and videos from text descriptions, defaulting to Google's Gemini for high-quality results or fal.ai FLUX.2 klein 4B in cheap mode for faster, lower-cost output. Customize aspect ratios, resolution, and include reference images for editing. Video generation uses Grok Imagine by default or switches to fal.ai LTX-2 for budget-conscious workflows.
Generate images from text prompts using your choice of 11+ AI providers including OpenAI GPT Image 2, Google, Azure, and DashScope. Supports reference images for identity preservation, custom aspect ratios, batch processing, and prompt files. Configure your default provider and model once, then generate single or multiple images with flexible quality and size options.
baoyu-imagine generates images from text prompts across 10+ AI providers including OpenAI GPT Image 2, Google, Azure OpenAI, and others. It supports reference images for identity preservation, batch generation, custom aspect ratios, and quality presets, with flexible configuration via local or user-home settings.
More skills Baoyu Image Gen (unlicensed) · Tuzi Image Gen (unlicensed) · Baoyu Image Gen (unlicensed) · Canghe Image Gen (unlicensed) · generate-brand-assets (MIT)