Claude Skills for Image Generation
Image generation skills cover the full raster workflow: text to image, editing and inpainting, model choice, style control, upscale and restore, and brand assets. Skills in this topic generate or edit bitmap visuals such as photos, illustrations, textures, sprites, mockups, and transparent-background cutouts. They are not for work better handled by editing existing SVG, vector, or code-native assets, extending an established icon or logo system, or building visuals directly in HTML, CSS, or canvas. The practical order is prompt, model, style control, then edit and upscale; editing often matters more than generation. This list is ranked by observed search demand and reviewed for relevance. It is not an endorsement. Use it to compare approaches, then read each skill's own documentation before adopting it.
8 to start with
48 reviewed options
Start with these.
Image work is prompt, model, style control, then edit and upscale — editing often matters more than generation.
These are the closest matches for this work. Start with the first one and open it to install.
- Rank 01View Agent SkillDirect
imagegen
OfficialGenerate or edit raster images when the task benefits from AI-created bitmap visuals such as photos, illustrations, textures, sprites, mockups, or transparent-background cutouts. Use when Codex should create a brand-new image, transform an existing image, or derive visual variants from references, and the output should be a bitmap asset rather than repo-native code or vector. Do not use when the task is better handled by editing existing SVG/vector/code-native assets, extending an established icon or logo system, or building the visual directly in HTML/CSS/canvas.
- Rank 02View Agent SkillDirect
ai-image-generation
NotableGenerate AI images with GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve and 50+ models via inference.sh CLI. Models: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Capabilities: text-to-image, image-to-image, inpainting, LoRA, image editing, upscaling, text rendering. Use for: AI art, product mockups, concept art, social media graphics, marketing visuals, illustrations. Triggers: flux, image generation, ai image, text to image, stable diffusion, generate image, ai art, midjourney alternative, dall-e alternative, text2img, t2i, i
142,155 · popularity - Rank 03View Agent SkillDirect
image-edit
Routes image editing on RunComfy to the right model (Nano Banana Edit, GPT Image 2 Edit, Flux Kontext, Z-Image Inpaint) with documented prompting patterns for batch, text rewrite, and mask edits.
421,211 · popularity - Rank 04View Agent SkillDirect417,503 · popularity
gpt-image-edit
Notable - Rank 05View Agent SkillDirect
eachlabs-image-edit
Edits, transforms, upscales, and enhances images using 130+ AI models via the EachLabs API, covering style transfer, background removal, inpainting, face swap, and try-on.
759 · popularity - Rank 06View Agent SkillDirect
image
NotableWhen the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets. Also use when the user mentions 'AI image generation,' 'generate an image,' 'create a graphic,' 'product mockup,' 'hero image,' 'social media graphic,' 'banner image,' 'cover photo,' 'profile banner,' 'listing screenshot,' 'Flux,' 'Flux Kontext,' 'Midjourney,' 'DALL-E,' 'GPT Image,' 'ChatGPT Images,' 'Ideogram,' 'Gemini image,' 'Nano Banana,' 'Recraft,' 'Stable Diffusion,' 'Canva,' 'Figma,' 'image optimization,' 'com
69,825 · popularity - Rank 07View Agent SkillDirect
qianwen-image-generation
Generates and edits images with Wan and Qwen Image models, supporting text-to-image, style transfer, subject consistency, multi-image composition, and text rendering.
Needs (stated): Python
4,826 · popularity - Rank 08View Agent SkillDirect
qwencloud-image-generation
Generates and edits images with Wan and Qwen Image models, supporting text-to-image, style transfer, subject consistency, and text rendering.
Needs (stated): Python
1,421 · popularity
Explore 40 more Agent Skills
- Rank 09View Agent SkillDirect
stable-diffusion-image-generation
Generates images with Stable Diffusion via HuggingFace Diffusers, covering text-to-image, image-to-image, inpainting, ControlNet, and LoRA.
880 · popularity - Rank 10View Agent SkillDirect
eachlabs-image-generation
Generates new images from text prompts using 60+ AI models (Flux, GPT Image, Gemini, Imagen, Seedream) via the EachLabs Predictions API.
294 · popularity - Rank 11View Agent SkillDirect
zenmux-image-generation
Generates or edits images through ZenMux with GPT Image, Nano Banana, Qwen Image, Seedream, FLUX and more, optimizing prompts and managing project config.
287 · popularity - Rank 12View Agent SkillDirect
ai-image-editing
Routes AI image-editing tasks (inpainting, background removal, upscaling, outpainting, restoration, retouch) to the right engine with a task-first framework and honesty guidance.
288 · popularity - Rank 13View Agent SkillDirect
layer-image-editing
Edits existing images on Layer: inpainting, outpainting, reframing, relighting, restyling, multi-angle, layer splitting, background removal, vectorization, and upscaling.
504 · popularity - Rank 14View Agent SkillDirect
fal-text-to-image
Generates, remixes, and edits images using fal.ai models (FLUX, Recraft, Imagen), covering text-to-image, image-to-image, and masked inpainting.
184 · popularity - Rank 15View Agent SkillDirect
generate-image
Generates and edits images through the OpenRouter Image API (Gemini, Seedream, Recraft, GPT-Image), covering photos, illustrations, logos, and compositing from references.
Needs (stated): Credentials · Network access · Python
1,656 · popularity - Rank 16View Agent SkillDirect65,496 · popularity
gpt-image-2
Notable - Rank 17View Agent SkillDirect
image-generation
Generates images from structured JSON prompts with optional reference images via a Python script, covering character design, scenes, and products.
2,762 · popularity - Rank 18View Agent SkillDirect
ai-image-prompts-skill
Recommends curated prompts from a 10,000+ real-world image generation prompt library, working with any text-to-image model and including sample images.
978 · popularity - Rank 19View Agent SkillDirect
linkfox-multimodal-generate-image
Generates and edits images for product visuals via the Linkfox API, covering text-to-image, image-to-image, background replacement, style transfer, and model swapping.
382 · popularity - Rank 20View Agent SkillDirect
image-prompt
Turns a social need into an image brief, teaches model-agnostic prompt craft, and routes to the right image tool, with an AI-generation gate and accessibility notes.
310 · popularity - Rank 21View Agent SkillDirect
marketing-image-generation
Generates marketing images, ad creatives, social cards, blog headers, and product mockups, with a concrete visual-direction workflow and AI-slop avoidance guidance.
217 · popularity - Rank 22View Agent SkillDirect
zimage-generation
Generates images using ModelScope's Z-Image API via a Python script, triggered by Zimage or ModelScope requests.
864 · popularity - Rank 23View Agent SkillDirect
short-drama-image-prompts
Writes copy-ready image prompts for short-drama characters, looks, locations, props, and state variants, including reference and edit prompt guidance.
367 · popularity - Rank 24View Agent SkillPossible
brandkit
NotablePremium brand-kit image generation skill for creating high-end brand-guidelines boards, logo systems, identity decks, and visual-world presentations. Trained for minimalist, cinematic, editorial, dark-tech, luxury, cultural, security, gaming, developer-tool, and consumer-app brand systems. Optimized for intentional logo concepting, refined composition, sparse typography, strong symbolic meaning, premium mockups, art-directed imagery, and flexible grid layouts.
297,293 · popularity - Rank 25View Agent SkillPossible
imagegen-frontend-mobile
NotableElite mobile app image-generation skill for creating premium, app-native screen concepts and flows. Designed for iOS, Android, and cross-platform mobile products. Prioritizes clean hierarchy, comfortably readable text, strong multi-screen consistency, controlled color palettes, non-generic creative direction, textured surfaces, image-led composition, tasteful custom iconography, and clean phone mockup framing. By default, screens should be shown inside a subtle premium iPhone or similar phone mockup with a visible frame, while the main focus stays on the app content itself. This skill generate
284,892 · popularity - Rank 26View Agent SkillPossible
chatgpt-image-ad
Generates Meta image-ad creatives via ChatGPT Image 2 through the Arcads API, with edge-safe layouts and glyph-safety guards.
164 · popularity - Rank 27View Agent SkillPossible
nano-banana-image-ad
Generates Meta image-ad creatives via Nano Banana 2/Pro through the Arcads API, with edge-safe layouts and multi-reference support.
164 · popularity - Rank 28View Agent SkillPossible
syncfusion-react-image-editor
Implements the Syncfusion React Image Editor component for cropping, annotations, filters, transformations, and export in React apps.
639 · popularity - Rank 29View Agent SkillPossible
syncfusion-blazor-image-editor
Implements the Syncfusion Blazor Image Editor component for cropping, annotations, filters, transformations, and export.
329 · popularity - Rank 30View Agent SkillPossible
syncfusion-angular-image-editor
Implements the Syncfusion Angular Image Editor component for cropping, annotations, filters, transformations, and export.
271 · popularity - Rank 31View Agent SkillPossible
syncfusion-maui-image-editor
Implements the Syncfusion .NET MAUI ImageEditor for cropping, transformations, annotations, filters, and saving edited images.
160 · popularity - Rank 32View Agent SkillPossible
physical-ai-defect-image-generation
Orchestrates defect image generation with NVIDIA Cosmos AnomalyGen on OSMO for PCBA, metal, and glass inspection, including image-edit augmentation.
1,936 · popularity - Rank 33View Agent SkillPossible
image-editing
Guides image editing and enhancement for Xiaohongshu content using mobile apps, covering adjustments, color correction, retouching, and text overlays.
422 · popularity - Rank 34View Agent SkillPossible
venice-image-edit
Catalogue entry advertising Venice.ai image edits, upscaling, and background removal, pointing to the upstream bundle for the full workflow.
2,495 · popularity - Rank 35View Agent SkillPossible
build
Implements features and sets up CE.SDK Web projects for building photo, video, and design editors across frameworks.
140 · popularity - Rank 36View Agent SkillPossible
fal-image-edit
Thin catalogue entry describing AI-powered image editing with style transfer and object removal, pointing to an upstream fal.ai skill.
285 · popularity - Rank 37View Agent SkillPossible
og-image-generation
Generates dynamic social preview (OG) images using Next.js file conventions and the next/og library.
234 · popularity - Rank 38View Agent SkillPossible
neta
Routing skill that indexes Neta capabilities and points to sub-skills including neta-creative for image/video/song generation.
152 · popularity - Rank 39View Agent SkillPossible
editor
Analyzes, generates, and edits video/audio/images with Diffusion Studio, composing video compositions.
1,447 · popularity - Rank 40View Agent SkillPossible
enhance-prompt
NotableTransforms vague UI ideas into polished, Stitch-optimized prompts. Enhances specificity, adds UI/UX keywords, injects design system context, and structures output for better generation results.
56,037 · popularity - Rank 41View Agent SkillPossible
image-optimization
Optimizes images for web with compression, modern formats, and responsive techniques to reduce file size.
648 · popularity - Rank 42View Agent SkillPossible
hatch-pet
OfficialCreate, repair, validate, visually QA, and package Codex-compatible animated pets and pet spritesheets from character art, generated images, company or prospect brand cues, or visual references. Use when a user wants a lightweight-worker Codex pet workflow, a non-pixel custom pet style, a prospect or company mascot pet, or a full 8x9 animated pet atlas with transparent unused cells, QA contact sheets, and pet.json packaging. This skill composes the installed $imagegen system skill for visual generation and uses bundled scripts for deterministic spritesheet assembly.
Needs (stated): Credentials
- Rank 43View Agent SkillPossible525,176 · popularity
image-to-video
Notable - Rank 44View Agent SkillPossible192,138 · popularity
seedance-2-5-image-to-video
Notable - Rank 45View Agent SkillPossible
ai-video-generation
NotableGenerate AI videos with Google Veo, Seedance 2.0, HappyHorse, Wan, Grok and 40+ models via inference.sh CLI. Models: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Capabilities: text-to-video, image-to-video, reference-to-video, video editing, lipsync, avatar animation, video upscaling, foley sound. Use for: social media videos, marketing content, explainer videos, product demos, AI avatars. Triggers: video generation, ai video, text to video, image to video, veo, animate image, video from image, ai animation, video generator, genera
142,615 · popularity - Rank 46View Agent SkillPossible
ai-avatar-video
NotableCreate AI avatar and talking head videos via inference.sh CLI. Recommended: P-Video-Avatar (fastest, cheapest, built-in TTS). Also: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (100+ languages, emotion steering for characters), ElevenLabs, Kokoro. Capabilities: audio-driven avatars, text-to-avatar, lipsync videos, talking head generation, virtual presenters, UGC content. Use for: AI presenters, explainer videos, virtual influencers, dubbing, marketing videos, UGC ads, gaming avatars, NPC dialogue. Triggers: ai avatar, talking head, lipsync, avatar video, virtual presenter, ai spokesperson
142,016 · popularity - Rank 47View Agent SkillPossible
chanjing-digital-human
Integrates the Chanjing platform for digital-human video synthesis and AI Creation image/video assets within FrameVideo.
150 · popularity - Rank 48View Agent SkillPossible
glmv-prompt-gen
Analyzes images/videos and generates professional prompts for text-to-image and text-to-video tools like Midjourney, Stable Diffusion, DALL-E, Sora, and Kling.
Needs (stated): API key · Python
300 · popularity
Frequently asked questions
- When should I use an image generation skill instead of editing SVG or code-native assets?
- Use one when the output should be a bitmap asset, such as a photo, illustration, texture, sprite, mockup, or transparent-background cutout. Avoid it when the task is better handled by editing existing SVG, vector, or code-native assets, extending an established icon or logo system, or building the visual directly in HTML, CSS, or canvas.
- What capabilities do skills in this topic typically cover?
- They span text-to-image, image-to-image, inpainting, LoRA, image editing, upscaling, and text rendering, along with style transfer, background removal, face swap, and try-on. Some route editing requests to specific models with documented prompting patterns for batch, text rewrite, and mask edits.
- Does generation or editing matter more in an image workflow?
- The workflow runs prompt, model, style control, then edit and upscale, and editing often matters more than generation. Many skills focus on transforming an existing image or deriving visual variants from references rather than creating a brand-new image.
- How is this ranked list put together, and is it an endorsement?
- The list is ranked by observed search demand and reviewed for relevance. It is not an endorsement, so check each skill's documentation before relying on it.
How to install an Agent Skill
Start with the first skill on this page. Claude Code reads a SKILL.md once it is in the skills directory:
- Claude Code — copy the SKILL.md for imagegen into ~/.claude/skills/imagegen/ for yourself, or .claude/skills/imagegen/ to share it with a project.
- From the registry — run npx skills add openai/skills@imagegen.
- Or download imagegen from its source and place that SKILL.md in the directory above.
How to choose Check the source and the skill’s own instructions before you install. Official and Notable describe provenance, not safety.