$npx skillfedfor your agent

Gemini Imagegen

This skill harnesses Google's Gemini image generation model to create and modify visuals from text descriptions. It supports multiple resolutions (1K–4K), aspect ratios, iterative refinement through multi-turn chat, and composition from multiple reference images. Best for logos, product mockups, stylized art, and photorealistic scenes.

Gemini Imagegen generates images from text prompts using Google's Gemini API with support for editing, refinement, and multiple aspect ratios.

AI-generated summary based on this skill's SKILL.md

448 122 unlicensed, metadata onlyupdated by davekilleen

Decision gist · record as of 2026-07-27

Gemini Imagegen generates images from text prompts using Google's Gemini API with support for editing, refinement, and multiple aspect ratios. This skill harnesses Google's Gemini image generation model to create and modify visuals from text descriptions. It supports multiple resolutions (1K–4K), aspect ratios, iterative refinement through multi-turn chat, and composition from multiple reference images. Best for logos, product mockups, stylized art, and photorealistic scenes.

manual: git clone https://github.com/davekilleen/Dex → cp -r Dex ~/.claude/skills/gemini-imagegen

Use it when

  • Gemini Imagegen excels at logos, product mockups, stylized art, and photorealistic scenes.
  • Yes.
Same gist for agents: .md · .json

Install

davekilleen/Dex/gemini-imagegen · repository language: Python

generated, unverified - the skill's exact subdirectory could not be determined; check the repository on GitHub

Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.

Frequently asked questions

AI-generated answers based on this skill's SKILL.md and metadata

How do I generate images with Gemini?

Gemini Imagegen uses Google's Gemini image generation model to create visuals from text descriptions. You provide a prompt describing what you want, and the skill generates images at resolutions from 1K to 4K with your chosen aspect ratio. It works seamlessly in Claude Code for programmatic image creation in your workflows.

What can Gemini Imagegen create?

Gemini Imagegen excels at logos, product mockups, stylized art, and photorealistic scenes. The skill supports iterative refinement through multi-turn chat, letting you adjust and improve results. You can also compose visuals from multiple reference images to blend concepts and styles into new creations.

Can I integrate Gemini image generation into Claude Code?

Yes. Gemini Imagegen integrates directly into Claude Code, giving you access to image generation capabilities within your coding workflows. This lets you create visual content programmatically as part of larger automation scripts or applications without leaving your development environment.

What resolutions and aspect ratios does Gemini Imagegen support?

Gemini Imagegen supports multiple resolutions ranging from 1K to 4K, with flexible aspect ratio options. This range covers everything from social media graphics to high-resolution prints, letting you tailor output dimensions to your specific project needs.

How does Gemini Imagegen refine images through iteration?

Gemini Imagegen uses multi-turn chat to enable iterative refinement. After generating an initial image, you can describe adjustments, request style changes, or ask for variations. The skill processes your feedback and produces refined versions, making it easy to perfect visuals without starting from scratch.

Let your AI agent find skills like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.

wish › “Generate images using Google Gemini API”

Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →

Related skills

Nano Banana
by rebyteai-template · rebyteai-template/rebyte-skills

Nano Banana turns text descriptions into images using Google's Gemini technology, with two model options: Flash for quick iterations up to 1024px, and Pro for professional-grade output up to 4K. Customize aspect ratios and sizes for social media, presentations, or print materials.

no license declared → metadata onlyupdated Jun 2026
★ 14repo stars
threejs-image-generator
by majidmanzarpour · majidmanzarpour/threejs-game-skills

Create polished 2D assets for Three.js games—concept art, textures, UI elements, backgrounds, and reference images for 3D conversion. The skill wraps Gemini's image generation and editing capabilities, letting you produce game-ready materials directly or feed them into procedural 3D workflows. Organize outputs by asset type and integrate with threejs-3d-generator for end-to-end game production.

MITupdated Jul 2026
★ 1,140repo stars
Gemini 3 Image Prompt
by modu-ai · modu-ai/cowork-plugins

This skill transforms your image description into a structured prompt following Google's official 5-component guide for Gemini 3 Pro Image (Nano Banana Pro). It guides you through preset selection, fine-tuning, and aspect ratio choices, then generates ready-to-paste prompts for Gemini, GPT-image-2, and Midjourney v8.1 in parallel.

no license declared → metadata onlyupdated Jun 2026
★ 261repo stars
nano-banana-pro
by nicepkg · nicepkg/ai-workflow

Nano Banana Pro harnesses Google's Gemini 3 Pro model to generate images from text descriptions, edit existing images, and apply style transformations. The skill excels at data-accurate infographics, precise text rendering, and context-aware generation using reference images, with flexible output sizing and aspect ratio controls.

MITfor claude-codeupdated Jan 2026
★ 270repo stars
nano-banana-pro
by swarmclawai · swarmclawai/swarmclaw

Nano Banana Pro harnesses Gemini 3 Pro's image capabilities to generate, edit, and composite images through simple command-line scripts. Supports text-to-image generation at multiple resolutions, single-image editing, and multi-image composition with customizable aspect ratios.

MITupdated Jun 2026
★ 628repo stars
Nano Banana Pro
by intellectronica · intellectronica/agent-skills

Nano Banana Pro taps Google's Gemini 3 Pro Image API to create new images from text descriptions or modify existing ones. Pick your resolution—1K, 2K, or 4K—and let the skill handle generation or editing tasks like style changes, element removal, or color adjustments.

CC0-1.0updated Apr 2026
★ 279repo stars

More skills nano-banana-pro (MIT) · nano-banana-pro (MIT)

Tags
image-synthesisvisual-generationapi-integrationai-artcontent-creation