Gemini Imagegen
This skill harnesses Google's Gemini image generation model to create and modify visuals from text descriptions. It supports multiple resolutions (1K–4K), aspect ratios, iterative refinement through multi-turn chat, and composition from multiple reference images. Best for logos, product mockups, stylized art, and photorealistic scenes.
Gemini Imagegen generates images from text prompts using Google's Gemini API with support for editing, refinement, and multiple aspect ratios.
AI-generated summary based on this skill's SKILL.md
Decision gist · record as of 2026-07-27
Gemini Imagegen generates images from text prompts using Google's Gemini API with support for editing, refinement, and multiple aspect ratios. This skill harnesses Google's Gemini image generation model to create and modify visuals from text descriptions. It supports multiple resolutions (1K–4K), aspect ratios, iterative refinement through multi-turn chat, and composition from multiple reference images. Best for logos, product mockups, stylized art, and photorealistic scenes.
Use it when
- Gemini Imagegen excels at logos, product mockups, stylized art, and photorealistic scenes.
- Yes.
Similar skills
Install
davekilleen/Dex/gemini-imagegen · repository language: Python
generated, unverified - the skill's exact subdirectory could not be determined; check the repository on GitHub
Open directory. Skills are indexed for reading, not audited. Review a skill's body before installing it.
Frequently asked questions
AI-generated answers based on this skill's SKILL.md and metadata
How do I generate images with Gemini?
Gemini Imagegen uses Google's Gemini image generation model to create visuals from text descriptions. You provide a prompt describing what you want, and the skill generates images at resolutions from 1K to 4K with your chosen aspect ratio. It works seamlessly in Claude Code for programmatic image creation in your workflows.
What can Gemini Imagegen create?
Gemini Imagegen excels at logos, product mockups, stylized art, and photorealistic scenes. The skill supports iterative refinement through multi-turn chat, letting you adjust and improve results. You can also compose visuals from multiple reference images to blend concepts and styles into new creations.
Can I integrate Gemini image generation into Claude Code?
Yes. Gemini Imagegen integrates directly into Claude Code, giving you access to image generation capabilities within your coding workflows. This lets you create visual content programmatically as part of larger automation scripts or applications without leaving your development environment.
What resolutions and aspect ratios does Gemini Imagegen support?
Gemini Imagegen supports multiple resolutions ranging from 1K to 4K, with flexible aspect ratio options. This range covers everything from social media graphics to high-resolution prints, letting you tailor output dimensions to your specific project needs.
How does Gemini Imagegen refine images through iteration?
Gemini Imagegen uses multi-turn chat to enable iterative refinement. After generating an initial image, you can describe adjustments, request style changes, or ask for variations. The skill processes your feedback and produces refined versions, making it easy to perfect visuals without starting from scratch.
Let your AI agent find skills like this
Example. Real query, live index.
You found this page by searching. An agent finds it by wishing: SkillFed indexes 56,283 agent skills by what they can do, searchable in plain language.
wish › “Generate images using Google Gemini API”
Give your agent the search over MCP, or paste the wish link into any chat. No install? Search from any chat →
Related skills
Nano Banana turns text descriptions into images using Google's Gemini technology, with two model options: Flash for quick iterations up to 1024px, and Pro for professional-grade output up to 4K. Customize aspect ratios and sizes for social media, presentations, or print materials.
Create polished 2D assets for Three.js games—concept art, textures, UI elements, backgrounds, and reference images for 3D conversion. The skill wraps Gemini's image generation and editing capabilities, letting you produce game-ready materials directly or feed them into procedural 3D workflows. Organize outputs by asset type and integrate with threejs-3d-generator for end-to-end game production.
This skill transforms your image description into a structured prompt following Google's official 5-component guide for Gemini 3 Pro Image (Nano Banana Pro). It guides you through preset selection, fine-tuning, and aspect ratio choices, then generates ready-to-paste prompts for Gemini, GPT-image-2, and Midjourney v8.1 in parallel.
Nano Banana Pro harnesses Google's Gemini 3 Pro model to generate images from text descriptions, edit existing images, and apply style transformations. The skill excels at data-accurate infographics, precise text rendering, and context-aware generation using reference images, with flexible output sizing and aspect ratio controls.
Nano Banana Pro harnesses Gemini 3 Pro's image capabilities to generate, edit, and composite images through simple command-line scripts. Supports text-to-image generation at multiple resolutions, single-image editing, and multi-image composition with customizable aspect ratios.
Nano Banana Pro taps Google's Gemini 3 Pro Image API to create new images from text descriptions or modify existing ones. Pick your resolution—1K, 2K, or 4K—and let the skill handle generation or editing tasks like style changes, element removal, or color adjustments.
More skills nano-banana-pro (MIT) · nano-banana-pro (MIT)