Claude Skills for Image Generation

Image generation skills cover the full raster workflow: text to image, editing and inpainting, model choice, style control, upscale and restore, and brand assets. Skills in this topic generate or edit bitmap visuals such as photos, illustrations, textures, sprites, mockups, and transparent-background cutouts. They are not for work better handled by editing existing SVG, vector, or code-native assets, extending an established icon or logo system, or building visuals directly in HTML, CSS, or canvas. The practical order is prompt, model, style control, then edit and upscale; editing often matters more than generation. This list is ranked by observed search demand and reviewed for relevance. It is not an endorsement. Use it to compare approaches, then read each skill's own documentation before adopting it.

8 to start with
48 reviewed options

Start with these.

Image work is prompt, model, style control, then edit and upscale — editing often matters more than generation.

These are the closest matches for this work. Start with the first one and open it to install.

48 reviewed options

  1. Rank 04
    Direct

    gpt-image-edit

    Notable
    • Text to image
    • Models
    417,503 · popularity
    View Agent Skill
  2. Rank 05
    Direct

    eachlabs-image-edit

    Edits, transforms, upscales, and enhances images using 130+ AI models via the EachLabs API, covering style transfer, background removal, inpainting, face swap, and try-on.

    • Editing and inpainting
    • Models
    • Style control
    759 · popularity
    View Agent Skill
  3. Rank 06
    Direct

    image

    Notable

    When the user wants to create, generate, edit, or optimize images for marketing — blog heroes, social graphics, product mockups, profile banners, listing visuals, or brand assets. Also use when the user mentions 'AI image generation,' 'generate an image,' 'create a graphic,' 'product mockup,' 'hero image,' 'social media graphic,' 'banner image,' 'cover photo,' 'profile banner,' 'listing screenshot,' 'Flux,' 'Flux Kontext,' 'Midjourney,' 'DALL-E,' 'GPT Image,' 'ChatGPT Images,' 'Ideogram,' 'Gemini image,' 'Nano Banana,' 'Recraft,' 'Stable Diffusion,' 'Canva,' 'Figma,' 'image optimization,' 'com

    • Editing and inpainting
    • Models
    • Style control
    69,825 · popularity
    View Agent Skill
  4. Rank 07
    Direct

    qianwen-image-generation

    Generates and edits images with Wan and Qwen Image models, supporting text-to-image, style transfer, subject consistency, multi-image composition, and text rendering.

    • Text to image
    • Editing and inpainting
    • Style control

    Needs (stated): Python

    4,826 · popularity
    View Agent Skill
  5. Rank 08
    Direct

    qwencloud-image-generation

    Generates and edits images with Wan and Qwen Image models, supporting text-to-image, style transfer, subject consistency, and text rendering.

    • Text to image
    • Editing and inpainting
    • Style control

    Needs (stated): Python

    1,421 · popularity
    View Agent Skill
Explore 40 more Agent Skills
  1. Rank 09
    Direct

    stable-diffusion-image-generation

    Generates images with Stable Diffusion via HuggingFace Diffusers, covering text-to-image, image-to-image, inpainting, ControlNet, and LoRA.

    • Text to image
    • Editing and inpainting
    • Models
    880 · popularity
    View Agent Skill
  2. Rank 10
    Direct

    eachlabs-image-generation

    Generates new images from text prompts using 60+ AI models (Flux, GPT Image, Gemini, Imagen, Seedream) via the EachLabs Predictions API.

    • Text to image
    • Editing and inpainting
    • Models
    294 · popularity
    View Agent Skill
  3. Rank 11
    Direct

    zenmux-image-generation

    Generates or edits images through ZenMux with GPT Image, Nano Banana, Qwen Image, Seedream, FLUX and more, optimizing prompts and managing project config.

    • Text to image
    • Editing and inpainting
    • Models
    287 · popularity
    View Agent Skill
  4. Rank 12
    Direct

    ai-image-editing

    Routes AI image-editing tasks (inpainting, background removal, upscaling, outpainting, restoration, retouch) to the right engine with a task-first framework and honesty guidance.

    • Text to image
    • Editing and inpainting
    • Models
    288 · popularity
    View Agent Skill
  5. Rank 13
    Direct

    layer-image-editing

    Edits existing images on Layer: inpainting, outpainting, reframing, relighting, restyling, multi-angle, layer splitting, background removal, vectorization, and upscaling.

    • Text to image
    • Editing and inpainting
    • Upscale and restore
    504 · popularity
    View Agent Skill
  6. Rank 14
    Direct

    fal-text-to-image

    Generates, remixes, and edits images using fal.ai models (FLUX, Recraft, Imagen), covering text-to-image, image-to-image, and masked inpainting.

    • Text to image
    • Editing and inpainting
    • Models
    184 · popularity
    View Agent Skill
  7. Rank 15
    Direct

    generate-image

    Generates and edits images through the OpenRouter Image API (Gemini, Seedream, Recraft, GPT-Image), covering photos, illustrations, logos, and compositing from references.

    • Text to image
    • Editing and inpainting
    • Models

    Needs (stated): Credentials · Network access · Python

    1,656 · popularity
    View Agent Skill
  8. Rank 16
    Direct

    gpt-image-2

    Notable
    • Text to image
    • Models
    • Style control
    65,496 · popularity
    View Agent Skill
  9. Rank 17
    Direct

    image-generation

    Generates images from structured JSON prompts with optional reference images via a Python script, covering character design, scenes, and products.

    • Text to image
    • Style control
    2,762 · popularity
    View Agent Skill
  10. Rank 18
    Direct

    ai-image-prompts-skill

    Recommends curated prompts from a 10,000+ real-world image generation prompt library, working with any text-to-image model and including sample images.

    • Text to image
    • Models
    978 · popularity
    View Agent Skill
  11. Rank 19
    Direct

    linkfox-multimodal-generate-image

    Generates and edits images for product visuals via the Linkfox API, covering text-to-image, image-to-image, background replacement, style transfer, and model swapping.

    • Text to image
    • Editing and inpainting
    • Style control
    382 · popularity
    View Agent Skill
  12. Rank 20
    Direct

    image-prompt

    Turns a social need into an image brief, teaches model-agnostic prompt craft, and routes to the right image tool, with an AI-generation gate and accessibility notes.

    • Text to image
    • Models
    • Style control
    310 · popularity
    View Agent Skill
  13. Rank 21
    Direct

    marketing-image-generation

    Generates marketing images, ad creatives, social cards, blog headers, and product mockups, with a concrete visual-direction workflow and AI-slop avoidance guidance.

    • Text to image
    • Style control
    • Brand assets
    217 · popularity
    View Agent Skill
  14. Rank 22
    Direct

    zimage-generation

    Generates images using ModelScope's Z-Image API via a Python script, triggered by Zimage or ModelScope requests.

    • Text to image
    864 · popularity
    View Agent Skill
  15. Rank 23
    Direct

    short-drama-image-prompts

    Writes copy-ready image prompts for short-drama characters, looks, locations, props, and state variants, including reference and edit prompt guidance.

    • Text to image
    • Brand assets
    367 · popularity
    View Agent Skill
  16. Rank 24
    Possible

    brandkit

    Notable

    Premium brand-kit image generation skill for creating high-end brand-guidelines boards, logo systems, identity decks, and visual-world presentations. Trained for minimalist, cinematic, editorial, dark-tech, luxury, cultural, security, gaming, developer-tool, and consumer-app brand systems. Optimized for intentional logo concepting, refined composition, sparse typography, strong symbolic meaning, premium mockups, art-directed imagery, and flexible grid layouts.

    • Style control
    • Brand assets
    297,293 · popularity
    View Agent Skill
  17. Rank 25
    Possible

    imagegen-frontend-mobile

    Notable

    Elite mobile app image-generation skill for creating premium, app-native screen concepts and flows. Designed for iOS, Android, and cross-platform mobile products. Prioritizes clean hierarchy, comfortably readable text, strong multi-screen consistency, controlled color palettes, non-generic creative direction, textured surfaces, image-led composition, tasteful custom iconography, and clean phone mockup framing. By default, screens should be shown inside a subtle premium iPhone or similar phone mockup with a visible frame, while the main focus stays on the app content itself. This skill generate

    • Style control
    284,892 · popularity
    View Agent Skill
  18. Rank 26
    Possible

    chatgpt-image-ad

    Generates Meta image-ad creatives via ChatGPT Image 2 through the Arcads API, with edge-safe layouts and glyph-safety guards.

    • Text to image
    • Models
    • Style control
    164 · popularity
    View Agent Skill
  19. Rank 27
    Possible

    nano-banana-image-ad

    Generates Meta image-ad creatives via Nano Banana 2/Pro through the Arcads API, with edge-safe layouts and multi-reference support.

    • Text to image
    • Models
    • Brand assets
    164 · popularity
    View Agent Skill
  20. Rank 28
    Possible

    syncfusion-react-image-editor

    Implements the Syncfusion React Image Editor component for cropping, annotations, filters, transformations, and export in React apps.

    • Editing and inpainting
    639 · popularity
    View Agent Skill
  21. Rank 29
    Possible

    syncfusion-blazor-image-editor

    Implements the Syncfusion Blazor Image Editor component for cropping, annotations, filters, transformations, and export.

    • Editing and inpainting
    329 · popularity
    View Agent Skill
  22. Rank 30
    Possible

    syncfusion-angular-image-editor

    Implements the Syncfusion Angular Image Editor component for cropping, annotations, filters, transformations, and export.

    • Editing and inpainting
    271 · popularity
    View Agent Skill
  23. Rank 31
    Possible

    syncfusion-maui-image-editor

    Implements the Syncfusion .NET MAUI ImageEditor for cropping, transformations, annotations, filters, and saving edited images.

    • Editing and inpainting
    160 · popularity
    View Agent Skill
  24. Rank 32
    Possible

    physical-ai-defect-image-generation

    Orchestrates defect image generation with NVIDIA Cosmos AnomalyGen on OSMO for PCBA, metal, and glass inspection, including image-edit augmentation.

    1,936 · popularity
    View Agent Skill
  25. Rank 33
    Possible

    image-editing

    Guides image editing and enhancement for Xiaohongshu content using mobile apps, covering adjustments, color correction, retouching, and text overlays.

    • Editing and inpainting
    • Style control
    422 · popularity
    View Agent Skill
  26. Rank 34
    Possible

    venice-image-edit

    Catalogue entry advertising Venice.ai image edits, upscaling, and background removal, pointing to the upstream bundle for the full workflow.

    • Upscale and restore
    2,495 · popularity
    View Agent Skill
  27. Rank 35
    Possible

    build

    Implements features and sets up CE.SDK Web projects for building photo, video, and design editors across frameworks.

    140 · popularity
    View Agent Skill
  28. Rank 36
    Possible

    fal-image-edit

    Thin catalogue entry describing AI-powered image editing with style transfer and object removal, pointing to an upstream fal.ai skill.

    • Editing and inpainting
    • Style control
    285 · popularity
    View Agent Skill
  29. Rank 37
    Possible

    og-image-generation

    Generates dynamic social preview (OG) images using Next.js file conventions and the next/og library.

    234 · popularity
    View Agent Skill
  30. Rank 38
    Possible

    neta

    Routing skill that indexes Neta capabilities and points to sub-skills including neta-creative for image/video/song generation.

    • Style control
    152 · popularity
    View Agent Skill
  31. Rank 39
    Possible

    editor

    Analyzes, generates, and edits video/audio/images with Diffusion Studio, composing video compositions.

    1,447 · popularity
    View Agent Skill
  32. Rank 40
    Possible

    enhance-prompt

    Notable

    Transforms vague UI ideas into polished, Stitch-optimized prompts. Enhances specificity, adds UI/UX keywords, injects design system context, and structures output for better generation results.

    • Text to image
    • Style control
    • Brand assets
    56,037 · popularity
    View Agent Skill
  33. Rank 41
    Possible

    image-optimization

    Optimizes images for web with compression, modern formats, and responsive techniques to reduce file size.

    648 · popularity
    View Agent Skill
  34. Rank 42
    Possible

    hatch-pet

    Official

    Create, repair, validate, visually QA, and package Codex-compatible animated pets and pet spritesheets from character art, generated images, company or prospect brand cues, or visual references. Use when a user wants a lightweight-worker Codex pet workflow, a non-pixel custom pet style, a prospect or company mascot pet, or a full 8x9 animated pet atlas with transparent unused cells, QA contact sheets, and pet.json packaging. This skill composes the installed $imagegen system skill for visual generation and uses bundled scripts for deterministic spritesheet assembly.

    • Text to image
    • Style control

    Needs (stated): Credentials

    View Agent Skill
  35. Rank 43
    Possible

    image-to-video

    Notable
    • Text to image
    525,176 · popularity
    View Agent Skill
  36. Rank 44
    Possible

    seedance-2-5-image-to-video

    Notable
    • Text to image
    192,138 · popularity
    View Agent Skill
  37. Rank 45
    Possible

    ai-video-generation

    Notable

    Generate AI videos with Google Veo, Seedance 2.0, HappyHorse, Wan, Grok and 40+ models via inference.sh CLI. Models: Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, HunyuanVideo. Capabilities: text-to-video, image-to-video, reference-to-video, video editing, lipsync, avatar animation, video upscaling, foley sound. Use for: social media videos, marketing content, explainer videos, product demos, AI avatars. Triggers: video generation, ai video, text to video, image to video, veo, animate image, video from image, ai animation, video generator, genera

    • Text to image
    • Upscale and restore
    142,615 · popularity
    View Agent Skill
  38. Rank 46
    Possible

    ai-avatar-video

    Notable

    Create AI avatar and talking head videos via inference.sh CLI. Recommended: P-Video-Avatar (fastest, cheapest, built-in TTS). Also: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (100+ languages, emotion steering for characters), ElevenLabs, Kokoro. Capabilities: audio-driven avatars, text-to-avatar, lipsync videos, talking head generation, virtual presenters, UGC content. Use for: AI presenters, explainer videos, virtual influencers, dubbing, marketing videos, UGC ads, gaming avatars, NPC dialogue. Triggers: ai avatar, talking head, lipsync, avatar video, virtual presenter, ai spokesperson

    • Text to image
    • Style control
    142,016 · popularity
    View Agent Skill
  39. Rank 47
    Possible

    chanjing-digital-human

    Integrates the Chanjing platform for digital-human video synthesis and AI Creation image/video assets within FrameVideo.

    • Editing and inpainting
    150 · popularity
    View Agent Skill
  40. Rank 48
    Possible

    glmv-prompt-gen

    Analyzes images/videos and generates professional prompts for text-to-image and text-to-video tools like Midjourney, Stable Diffusion, DALL-E, Sora, and Kling.

    • Text to image
    • Models

    Needs (stated): API key · Python

    300 · popularity
    View Agent Skill
Frequently asked questions

Frequently asked questions

When should I use an image generation skill instead of editing SVG or code-native assets?
Use one when the output should be a bitmap asset, such as a photo, illustration, texture, sprite, mockup, or transparent-background cutout. Avoid it when the task is better handled by editing existing SVG, vector, or code-native assets, extending an established icon or logo system, or building the visual directly in HTML, CSS, or canvas.
What capabilities do skills in this topic typically cover?
They span text-to-image, image-to-image, inpainting, LoRA, image editing, upscaling, and text rendering, along with style transfer, background removal, face swap, and try-on. Some route editing requests to specific models with documented prompting patterns for batch, text rewrite, and mask edits.
Does generation or editing matter more in an image workflow?
The workflow runs prompt, model, style control, then edit and upscale, and editing often matters more than generation. Many skills focus on transforming an existing image or deriving visual variants from references rather than creating a brand-new image.
How is this ranked list put together, and is it an endorsement?
The list is ranked by observed search demand and reviewed for relevance. It is not an endorsement, so check each skill's documentation before relying on it.

How to install an Agent Skill

Start with the first skill on this page. Claude Code reads a SKILL.md once it is in the skills directory:

  • Claude Code — copy the SKILL.md for imagegen into ~/.claude/skills/imagegen/ for yourself, or .claude/skills/imagegen/ to share it with a project.
  • From the registry — run npx skills add openai/skills@imagegen.
  • Or download imagegen from its source and place that SKILL.md in the directory above.

How to choose Check the source and the skill’s own instructions before you install. Official and Notable describe provenance, not safety.