Generates visual assets like icons, textures, and mockups using AI.
Install
mkdir -p .claude/skills/generate-image-m00sp && curl -L -o skill.zip "https://agentskills.codes/api/skills/download/14333" && unzip -o skill.zip -d .claude/skills/generate-image-m00sp && rm skill.zipInstalls to .claude/skills/generate-image-m00sp
Activation
This is the description your AI agent reads to decide when to run this skill — the better it matches your request, the more reliably it fires.
Generate images using AI. Use when asked to generate, create, or make images, textures, icons, sprites, artwork, visual assets, or mockups. Supports OpenAI (gpt-image-2) and Google Gemini (Nano Banana). Requires an API key for the chosen provider.Key capabilities
- →Generate images using OpenAI or Google Gemini
- →Check for available API keys for image generation providers
- →Guide users through API key setup if none are configured
- →Save generated images to a specified output path
- →Enrich prompts for game textures with specific keywords
How it works
The skill checks for configured API keys for OpenAI or Google Gemini, then uses the selected provider to generate an image based on the user's prompt. If no keys are set, it guides the user through the onboarding process to obtain and set an API key.
Inputs & outputs
When to use generate-image
- →Generating icons for UI
- →Creating placeholder textures
- →Making visual mockups
About this skill
Generate Image
You are an image generation assistant. When invoked, follow the workflow below.
Workflow
- Check for API keys — check whether
SKILL_IMAGE_GEN_OPENAI_KEYand/orSKILL_IMAGE_GEN_GEMINI_KEYare set in the environment. - If one key is set — use that provider. No need to ask.
- If both are set — pick based on context (OpenAI for polish, Gemini for speed), or ask if the user has a preference.
- If no keys are set — run the Onboarding section.
- Generate the image using the appropriate API reference.
- Tell the user where the image was saved.
Onboarding
Only run this if no keys are set. Guide the user conversationally.
- Ask which provider they'd like to use:
- OpenAI (gpt-image-2) — High quality, excellent text rendering, paid per image
- Google Gemini (Nano Banana) — Fast, free tier available, great for iteration
- Direct them to get an API key:
- OpenAI → https://platform.openai.com/api-keys
- Gemini → https://aistudio.google.com/apikey
- Once they provide the key, set
SKILL_IMAGE_GEN_OPENAI_KEYorSKILL_IMAGE_GEN_GEMINI_KEYin the current session and persist it to the appropriate shell profile. - Proceed to generate the image they originally asked for.
API Reference: OpenAI
Method: POST
URL: https://api.openai.com/v1/images/generations
Headers:
Authorization: Bearer <SKILL_IMAGE_GEN_OPENAI_KEY>Content-Type: application/json
Body (JSON):
{
"model": "gpt-image-2",
"prompt": "<user prompt>",
"n": 1,
"size": "1024x1024",
"quality": "medium"
}
| Field | Default | Options |
|---|---|---|
| model | gpt-image-2 | gpt-image-2, gpt-image-1 |
| size | 1024x1024 | 1024x1024, 1024x1536, 1536x1024, auto |
| quality | medium | low, medium, high |
Response: data[0].b64_json contains the base64-encoded image. Decode it and save to the output path. If data[0].url is present instead, download the image from that URL.
API Reference: Google Gemini (Nano Banana)
Method: POST (uses Generative Language Image or Content endpoint depending on model)
URL (examples):
- Text+image/content endpoint:
https://generativelanguage.googleapis.com/v1beta/models/<model>:generateContent - Image-specific endpoint (newer image models may use):
https://generativelanguage.googleapis.com/v1/models/<model>:generateImage
Authentication (recommended):
- Preferred:
Authorization: Bearer <ACCESS_TOKEN>(use OAuth2 service account or short-lived access token). - API Key (legacy / limited): can be passed as
x-goog-api-key: <KEY>header or?key=<KEY>query param, but some image-generation models require OAuth/Bearer tokens.
Headers:
Authorization: Bearer <ACCESS_TOKEN>(preferred)Content-Type: application/json
Body (JSON) — content endpoint (text+image):
{
"contents": [{"parts": [{"text": "Generate an image: <user prompt>"}]}],
"generationConfig": {"responseModalities": ["IMAGE"]}
}
Body (JSON) — image-specific endpoint (if supported by model):
{
"prompt": "<user prompt>",
"imageConfig": {"modes": ["RGB"], "size": "1024x1024"}
}
| Field | Default | Options |
|---|---|---|
| model (in URL) | gemini-2.0-flash-exp or an image-capable model | e.g., gemini-2.5-flash-image, image-bison |
Response (guidance):
- For content endpoint: inspect
candidates[0].content.parts[]for entries withinlineData.data(base64 image) andinlineData.mimeType. - For image endpoint: response commonly includes
imagesordatafield carrying base64-encoded image(s) or URLs — decode/save accordingly.
Notes & troubleshooting:
- Many Google image models now require OAuth Bearer tokens; prefer service-account-based short-lived tokens rather than long-lived API keys.
- If
401/403errors occur, verify token scopes and that the calling project has the Generative API enabled. - If requests return
400, confirm request body shape and model name. - Log outgoing request headers and body (with keys redacted) to aid debugging.
Error cases: error key (API error), promptFeedback.blockReason (safety block), finishReason: "SAFETY" (filtered).
Migration guidance:
- Update integrations to support both
Authorization: Bearerandx-goog-api-keybut prefer and document using OAuth where image models require it. - Add tests that mock both endpoints and assert request shape, headers, and image decoding behavior.
Agent Guidelines
- Choose the output path intelligently — save to the project's relevant directory (e.g.,
assets/,images/, or the current directory). - For game textures, enrich prompts with "seamless", "tileable", "game asset".
- For batch generation, make multiple API calls in parallel.
- If the user asks to switch providers or what options are available, explain both and help them set up.
- Always create the output directory before saving.
- Ensure special characters in the user's prompt are properly escaped in the JSON body.
When not to use it
- →When no image generation is required
- →When the task is not related to visual content creation
Limitations
- →Requires an API key for the chosen provider
- →Image generation quality depends on the selected provider and prompt details
How it compares
This skill automates the selection and configuration of image generation providers, simplify the process compared to manual API calls and setup.
Compared to similar skills
generate-image side by side with the closest alternatives in the catalog.
| Skill | Installs | Updated | Safety | Difficulty |
|---|---|---|---|---|
| generate-image (this skill) | 0 | 1mo | No flags | Beginner |
| jimeng-mcp-skill | 19 | 4mo | Caution | Intermediate |
| gemini-logo-remover | 9 | 8mo | Review | Beginner |
| ai-image | 9 | 9mo | Review | Intermediate |
Try saying
Example prompts that trigger this skill in your AI assistant.
You might also like
jimeng-mcp-skill
wwwzhouhui
使用jimeng-mcp-server进行AI图像和视频生成。当用户请求从文本生成图像、合成多张图片、从文本描述创建视频或为静态图像添加动画时使用此技能。支持四大核心能力:文生图、图像合成、文生视频、图生视频。需要jimeng-mcp-server在本地运行或通过SSE/HTTP访问。
gemini-logo-remover
bear2u
Remove Gemini logos, watermarks, or AI-generated image markers using OpenCV inpainting. Use this skill when the user asks to remove Gemini logo, AI watermark, or any logo/watermark from images.
ai-image
tyrchen
Generate AI images using OpenAI's gpt-image-1 model with customizable aspect ratios and artistic themes. Use when the user wants to create images, generate artwork, or mentions image generation with specific styles like Ghibli, futuristic, Pixar, oil painting, or Chinese painting.
baoyu-article-illustrator
JimLiu
Analyzes article structure, identifies positions requiring visual aids, generates illustrations with Type × Style two-dimension approach. Use when user asks to "illustrate article", "add images", "generate images for article", or "为文章配图".
nano-banana-pro-prompts-recommend-skill
YouMind-OpenLab
Recommend suitable prompts from 6000+ Nano Banana Pro image generation prompts based on user needs. Use this skill when users want to: - Generate images with AI (Nano Banana Pro model) - Find inspiration for image generation prompts - Get prompt recommendations for specific use cases (portraits, landscapes, product photos, etc.) - Create illustrations for articles, videos, podcasts, or other content - Translate and understand prompt techniques
image-generation
onyx-dot-app
Generate images using nano banana.