
Generate Image
OfficialFreeCreate images quickly using AI providers.
Free · Opens the source repo
What Generate Image does
Generate Image is a skill designed for developers and designers looking to create visual assets efficiently. By leveraging AI capabilities from OpenAI and Google Gemini, this skill simplifies the process of image generation, allowing users to focus on creativity rather than technical setup. It supports generating textures, icons, sprites, artwork, and mockups, making it a versatile tool for various projects.
The skill operates through a straightforward workflow. It first checks for API keys for either OpenAI or Google Gemini. If a key is present, it automatically selects the appropriate provider based on user preference or context. In cases where no key is set, the skill guides the user through a simple onboarding process to obtain and configure the necessary API keys. This ensures that even those new to image generation can start creating images without extensive technical knowledge.
Users can specify their image requirements through prompts, and the skill handles the API interactions, including sending requests and processing responses. Images are saved directly to the specified output path, which can be customized to suit the project's directory structure. This feature is particularly useful for game developers and designers who need to maintain organized asset folders.
Overall, Generate Image is an essential tool for anyone needing to produce high-quality images quickly. Its integration with leading AI providers ensures that users can choose between quality and speed based on their specific needs, making it suitable for both rapid prototyping and polished final products.
When to use it
Use this skill when you need to generate images or visual assets quickly and efficiently, especially when working on design or development projects.
When not to use it
This skill may not be suitable if you require highly customized image generation features beyond what the APIs offer or if you need to work offline.
What you can build with it
Creating Game Textures
Generate seamless and tileable textures for games by using specific prompts tailored for game assets.
Rapid Prototyping of Visual Assets
Quickly create mockups and visual assets for presentations or client feedback, enhancing the design workflow.
Iterative Design Processes
Use the fast generation capabilities of Google Gemini to create multiple iterations of an image for design refinement.
How to install Generate Image
View source1. Install with the skills CLI
npx skills add github/awesome-copilot/generate-image --agent claude-code2. Or install it manually
Download the skill folder and drop it into ~/.claude/skills/ for all projects, or .claude/skills/ to scope it to one repo. Restart Claude Code so it picks up the new skill.
Anthropic's agentic coding CLI, and the reference implementation of Agent Skills. Drop a skill folder into ~/.claude/skills and Claude Code loads it automatically whenever a task matches the skill's description. Claude Code docs
Inside SKILL.md
Written by githubGenerate Image
You are an image generation assistant. When invoked, follow the workflow below.
Workflow
- Check for API keys — check whether
SKILL_IMAGE_GEN_OPENAI_KEYand/orSKILL_IMAGE_GEN_GEMINI_KEYare set in the environment. - If one key is set — use that provider. No need to ask.
- If both are set — pick based on context (OpenAI for polish, Gemini for speed), or ask if the user has a preference.
- If no keys are set — run the Onboarding section.
- Generate the image using the appropriate API reference.
- Tell the user where the image was saved.
Onboarding
Only run this if no keys are set. Guide the user conversationally.
- Ask which provider they'd like to use:
- OpenAI (gpt-image-2) — High quality, excellent text rendering, paid per image
- Google Gemini (Nano Banana) — Fast, free tier available, great for iteration
- Direct them to get an API key:
- OpenAI → https://platform.openai.com/api-keys
- Gemini → https://aistudio.google.com/apikey
- Once they provide the key, set
SKILL_IMAGE_GEN_OPENAI_KEYorSKILL_IMAGE_GEN_GEMINI_KEYin the current session and persist it to the appropriate shell profile. - Proceed to generate the image they originally asked for.
API Reference: OpenAI
Method: POST
URL: https://api.openai.com/v1/images/generations
Headers:
Authorization: Bearer <SKILL_IMAGE_GEN_OPENAI_KEY>Content-Type: application/json
Body (JSON):
{
"model": "gpt-image-2",
"prompt": "<user prompt>",
"n": 1,
"size": "1024x1024",
"quality": "medium"
}
| Field | Default | Options |
|---|---|---|
| model | gpt-image-2 | gpt-image-2, gpt-image-1 |
| size | 1024x1024 | 1024x1024, 1024x1536, 1536x1024, auto |
| quality | medium | low, medium, high |
Response: data[0].b64_json contains the base64-encoded image. Decode it and save to the output path. If data[0].url is present instead, download the image from that URL.
API Reference: Google Gemini (Nano Banana)
Method: POST
URL: https://generativelanguage.googleapis.com/v1beta/models/<model>:generateContent
Headers:
x-goog-api-key: <SKILL_IMAGE_GEN_GEMINI_KEY>Content-Type: application/json
Body (JSON):
{
"contents": [{"parts": [{"text": "Generate an image: <user prompt>"}]}],
"generationConfig": {"responseModalities": ["TEXT", "IMAGE"]}
}
| Field | Default | Options |
|---|---|---|
| model (in URL) | gemini-2.0-flash-exp | gemini-2.0-flash-exp, gemini-2.5-flash-image |
Response: Find candidates[0].content.parts[] — look for a part with inlineData.data (base64 image) and inlineData.mimeType. Decode and save.
Error cases: error key (API error), promptFeedback.blockReason (safety block), finishReason: "SAFETY" (filtered).
Agent Guidelines
- Choose the output path intelligently — save to the project's relevant directory (e.g.,
assets/,images/, or the current directory). - For game textures, enrich prompts with "seamless", "tileable", "game asset".
- For batch generation, make multiple API calls in parallel.
- If the user asks to switch providers or what options are available, explain both and help them set up.
- Always create the output directory before saving.
- Ensure special characters in the user's prompt are properly escaped in the JSON body.
Frequently asked questions about Generate Image
Similar skills
Algorithmic Art
Create generative art using p5.js and algorithmic philosophies.
Fal.ai Media Generation
Create images, videos, and audio with AI.
p5.js Production Pipeline
Create stunning generative art and interactive visuals.
ASCII Video Production
Transform videos into striking ASCII art animations.
Pixel Art
Transform images into retro pixel art and animations.
Pretext Creative Demos
Build engaging browser demos with innovative text layouts.
