Claude Skills for Video Generation

Video generation skills cover the tooling an agent needs to turn a written prompt or a still image into a finished clip. That includes text-to-video and image-to-video entry points, access to specific models, shot control, render pipeline steps, and audio or lipsync work. Skills in this topic often wrap a CLI or MCP interface, handle login and region-aware install, and expose multiple backends behind one call. Use the ranked list as a starting point, not a recommendation. Ranking reflects observed search demand, and each entry was reviewed for relevance to the topic. Read the skill description to confirm which models, tiers, and capabilities it actually exposes before you load it, and check whether it requires an API key or account setup.

8 to start with
51 reviewed options

Start with these.

AI video starts with a prompt or a still: choose a model, control the shot, then render and sync audio.

These are the closest matches for this work. Start with the first one and open it to install.

Start from a written prompt.

51 reviewed options

  1. Rank 04
    Direct

    kling-3-0

    Covers all six Kling 3.0 endpoints on RunComfy across Standard, Pro, and 4K tiers for text-to-video and image-to-video.

    • Text to video
    • Image to video
    • Models
    397,152 · popularity
    View Agent Skill
  2. Rank 05
    Direct

    videoagent-video-studio

    Generates short AI videos from text or images across 7 backends (minimax, kling, veo, hunyuan, grok, seedance) with zero API key setup.

    • Text to video
    • Image to video
    • Models
    10,464 · popularity
    View Agent Skill
  3. Rank 06
    Direct

    qianwen-video-generation

    Generates videos with Wan and HappyHorse models supporting text-to-video, image-to-video, first+last frame, reference role-play, and VACE editing.

    • Text to video
    • Image to video
    • Models

    Needs (stated): Python

    4,692 · popularity
    View Agent Skill
  4. Rank 07
    Direct

    qwencloud-video-generation

    Generates videos with Wan models supporting text-to-video, image-to-video, first+last frame, reference role-play, and VACE editing.

    • Text to video
    • Image to video
    • Models

    Needs (stated): Python

    1,376 · popularity
    View Agent Skill
  5. Rank 08
    Direct

    eachlabs-video-generation

    Generates videos from text, images, or references using 165+ AI models via the EachLabs Predictions API, covering t2v, i2v, transitions, and avatars.

    • Text to video
    • Image to video
    • Models
    467 · popularity
    View Agent Skill
Explore 43 more Agent Skills
  1. Rank 09
    Direct

    video-generation

    End-to-end AI video production through Hyper MCP covering text-to-video, image-to-video, scene chaining, captions, voiceover, and editing.

    • Text to video
    • Image to video
    • Models
    171 · popularity
    View Agent Skill
  2. Rank 10
    Direct

    text-to-video

    Generates dynamic videos from text descriptions using meitu-cli with high-quality and high-motion modes.

    • Text to video

    Needs (stated): Credentials

    88 · popularity
    View Agent Skill
  3. Rank 11
    Direct

    aliyun-kling-video

    Generates videos with Kling v3 models on DashScope covering text-to-video, image-to-video, reference-to-video, smart storyboard, and video editing.

    • Text to video
    • Image to video
    • Models
    72 · popularity
    View Agent Skill
  4. Rank 12
    Direct

    kling-video-generator

    Generates videos from text, images, or other videos using the Kling 3.0 Omni model covering t2v, i2v, editing, multi-shot, and audio-synced video.

    • Text to video
    • Image to video
    • Models

    Needs (stated): Python

    81 · popularity
    View Agent Skill
  5. Rank 13
    Direct

    fal-text-to-video

    Complete fal.ai text-to-video system covering Kling, Sora 2, LTX, Runway, and Luma with model endpoints and prompt structure.

    • Text to video
    • Image to video
    • Models
    78 · popularity
    View Agent Skill
  6. Rank 14
    Direct

    video

    Notable

    When the user wants to create, generate, or produce video content using AI tools or programmatic frameworks. Also use when the user mentions 'video production,' 'AI video,' 'Remotion,' 'Hyperframes,' 'HeyGen,' 'Synthesia,' 'Veo,' 'Sora,' 'Runway,' 'Kling,' 'Seedance,' 'Hailuo,' 'MiniMax,' 'Pika,' 'Hunyuan,' 'Wan,' 'video generation,' 'AI avatar,' 'talking head video,' 'programmatic video,' 'video template,' 'explainer video,' 'product demo video,' 'video pipeline,' 'copy this edit,' 'match this video style,' 'reverse-engineer this video,' 'edit like this reference,' or 'make me a video.' Use t

    • Models
    • Render pipeline
    • Audio
    69,699 · popularity
    View Agent Skill
  7. Rank 15
    Direct

    ai-avatar-video

    Notable

    Create AI avatar and talking head videos via inference.sh CLI. Recommended: P-Video-Avatar (fastest, cheapest, built-in TTS). Also: OmniHuman, Fabric, PixVerse. Audio: Inworld TTS-2 (100+ languages, emotion steering for characters), ElevenLabs, Kokoro. Capabilities: audio-driven avatars, text-to-avatar, lipsync videos, talking head generation, virtual presenters, UGC content. Use for: AI presenters, explainer videos, virtual influencers, dubbing, marketing videos, UGC ads, gaming avatars, NPC dialogue. Triggers: ai avatar, talking head, lipsync, avatar video, virtual presenter, ai spokesperson

    • Audio
    142,016 · popularity
    View Agent Skill
  8. Rank 16
    Direct

    higgsfield-video-explainer

    Notable
    77,553 · popularity
    View Agent Skill
  9. Rank 17
    Direct

    agnes-ai-generation

    Calls Agnes AI APIs for text, image, and video generation including text-to-video, image-to-video, and keyframe video via a Python script.

    • Text to video
    • Image to video
    • Shot control

    Needs (stated): API key

    2,312 · popularity
    View Agent Skill
  10. Rank 18
    Direct

    kling-3-prompting

    Teaches writing better Kling 3.0 prompts for text-to-video, image-to-video, keyframes, multi-shot, and dialogue scenes via an interactive builder.

    • Text to video
    • Image to video
    • Models
    574 · popularity
    View Agent Skill
  11. Rank 19
    Direct

    kling-official

    Official Kling direct API guidance for OpenMontage providers covering video, image, TTS, avatar, and lip-sync endpoints.

    • Models
    • Audio
    484 · popularity
    View Agent Skill
  12. Rank 20
    Direct

    kling

    Guides Kling-led generative video production for native 4K, multi-shot storyboarding, and motion transfer with consent and disclosure gates.

    • Image to video
    • Models
    • Audio
    283 · popularity
    View Agent Skill
  13. Rank 21
    Direct

    kling-video

    Generates AI videos using Kling models for text-to-video, image-to-video, video-to-video, and AI avatar videos via the Refly CLI.

    • Text to video
    • Image to video
    • Models
    265 · popularity
    View Agent Skill
  14. Rank 22
    Direct

    ai-video-production-master

    Expert script-to-video production pipelines for Apple Silicon covering stock footage, T2V, I2V, LoRA training, and cloud GPU orchestration.

    • Text to video
    • Image to video
    • Models
    389 · popularity
    View Agent Skill
  15. Rank 23
    Direct

    comfyui-video-production

    Plans and orchestrates end-to-end video production pipelines in ComfyUI covering img2vid, txt2vid, vid2vid, and multi-shot with validation gates.

    • Image to video
    • Models
    • Shot control
    361 · popularity
    View Agent Skill
  16. Rank 24
    Direct

    model-routing

    Chooses default fal.ai endpoint IDs for genmedia production skills including video generation, image generation, and editing.

    • Text to video
    • Image to video
    • Models
    331 · popularity
    View Agent Skill
  17. Rank 25
    Direct

    muapi-product-video-ad-maker

    Creates a cinematic product video ad from a product photo using muapi image edit and image-to-video models.

    • Image to video
    • Render pipeline
    2,291 · popularity
    View Agent Skill
  18. Rank 26
    Direct

    chanjing-digital-human

    Integrates the Chanjing platform for digital-human video synthesis, voice selection, AI Creation assets, and BGM/SFX downloads.

    • Render pipeline
    • Audio
    150 · popularity
    View Agent Skill
  19. Rank 27
    Direct

    narrator-ai-video-generation

    Generates AI-narrated video content using narrator-ai-cli for movie commentary, short dramas, and film analysis.

    Needs (stated): API key

    137 · popularity
    View Agent Skill
  20. Rank 28
    Direct

    framevideo-ai-production

    Chinese AI production subflow for FrameVideo covering script-to-storyboard, shot aggregation, prompt polishing, and Chanjing AIGC generation.

    • Text to video
    • Shot control
    • Audio
    97 · popularity
    View Agent Skill
  21. Rank 29
    Direct

    chanjing-one-click-video-creation

    One-click short video renderer that generates complete videos with script, storyboard, digital-human narration, and AI footage via Chanjing API.

    • Shot control
    • Render pipeline
    • Audio
    77 · popularity
    View Agent Skill
  22. Rank 30
    Possible

    remotion-best-practices

    Notable

    Router for all Remotion skills

    • Render pipeline
    • Audio
    View Agent Skill
  23. Rank 31
    Possible

    remotion-create

    Notable

    Create a new Remotion video

    • Render pipeline
    View Agent Skill
  24. Rank 32
    Possible

    remotion-captions

    Notable

    Transcribing, displaying and animating captions

    • Audio
    View Agent Skill
  25. Rank 33
    Possible

    remotion-docs

    Notable

    Search Remotion documentation

    View Agent Skill
  26. Rank 34
    Possible

    remotion-interactivity

    Notable

    Structure Remotion markup for interactivity

    View Agent Skill
  27. Rank 35
    Possible

    general-video

    Notable
    • Shot control
    284,394 · popularity
    View Agent Skill
  28. Rank 36
    Possible

    music-to-video

    Notable

    Turn a music track (an audio file, a video to pull audio from, or a track generated from a mood brief) into a beat-synced video — lyric video, slideshow, or kinetic promo. The music drives all pacing; any user-supplied images/videos are cut onto the same beat grid, and a complete video needs zero assets. Narrated pieces → the input-matched workflow (see /hyperframes). Unclear → /hyperframes.

    • Shot control
    • Render pipeline
    201,379 · popularity
    View Agent Skill
  29. Rank 37
    Possible

    changelog-video

    Notable

    Turn a weekly changelog .md into a finished branded changelog video (square 1080, ~45-60s, Annie VO, animated brand background, mock-UI visualizations, lowkey captions). Use when the user provides a changelog/digest markdown and wants the weekly video, or says "changelog video". Self-contained — fonts, background, lexicon, and scripts ship in this skill.

    • Render pipeline
    • Audio
    60,100 · popularity
    View Agent Skill
  30. Rank 38
    Possible

    remotion-video-creation

    Best practices for Remotion video creation in React with 29 domain-specific rules covering 3D, animations, audio, captions, and transitions.

    • Audio
    8,418 · popularity
    View Agent Skill
  31. Rank 39
    Possible

    fal-kling-o3

    Catalogue entry advertising Kling O3 image and video generation via fal.ai, pointing to an upstream bundle.

    • Models
    2,476 · popularity
    View Agent Skill
  32. Rank 40
    Possible

    short-video-production

    Guides short vertical video production for Xiaohongshu covering hooks, pacing, filming, and editing.

    753 · popularity
    View Agent Skill
  33. Rank 41
    Possible

    video-production

    Plans and routes programmable or automated video production across code-first, template-first, and hybrid pipelines including Remotion.

    • Shot control
    • Render pipeline
    • Audio
    472 · popularity
    View Agent Skill
  34. Rank 42
    Possible

    music-video-generation

    Generates music videos, visualizers, and lyric videos synchronized to audio via the each::sense API.

    352 · popularity
    View Agent Skill
  35. Rank 43
    Possible

    youtube-video-generation

    Generates YouTube videos and Shorts including faceless videos, explainers, and tutorials via the each::sense API.

    • Audio
    320 · popularity
    View Agent Skill
  36. Rank 44
    Possible

    ugc video generation

    Generates UGC-style videos including testimonials, unboxings, and reviews via the each::sense API.

    246 · popularity
    View Agent Skill
  37. Rank 45
    Possible

    codex-storyboard-video-production

    Creates and manages AI video storyboard projects with automated asset generation through the Codex Storyboard workspace.

    • Shot control
    196 · popularity
    View Agent Skill
  38. Rank 46
    Possible

    nsfw video generation

    Generates adult video content via the each::sense API with safety checker disabled.

    192 · popularity
    View Agent Skill
  39. Rank 47
    Possible

    ra-video-production-director

    End-to-end video production orchestration routing to HyperFrames, Remotion, captions, TTS, and other skills for a content-creation workspace.

    • Render pipeline
    • Audio
    171 · popularity
    View Agent Skill
  40. Rank 48
    Possible

    wjs-converting-text-to-video

    Turns a WeChat article into a narrated short MP4 video with TTS voiceover and HyperFrames animation.

    • Text to video
    • Audio
    150 · popularity
    View Agent Skill
  41. Rank 49
    Possible

    neta-creative

    Neta API creative skill for generating images, videos, songs, and MVs, and deconstructing creative ideas from existing works.

    116 · popularity
    View Agent Skill
  42. Rank 50
    Possible

    promo-video

    Creates professional promotional videos using Remotion with AI voiceover and background music.

    • Render pipeline
    • Audio
    94 · popularity
    View Agent Skill
  43. Rank 51
    Possible

    explainer-video-production

    Provider-independent explainer video production covering scripting, storyboarding, claim review, localization, and QA.

    • Shot control
    • Render pipeline
    • Audio
    71 · popularity
    View Agent Skill
What this topic covers

What this topic covers

Each sub-topic groups Agent Skills for a specific situation.

Text to video
Start from a written prompt.
Frequently asked questions

Frequently asked questions

What is the difference between text-to-video and image-to-video skills?
Text-to-video skills start from a written prompt, while image-to-video skills animate a still image. Some skills support both, along with reference-to-video and video editing.
Do video generation skills require an API key?
It depends on the skill. Some wrap a CLI that needs login and region-aware install, while others advertise zero API key setup across multiple backends.
Which models can these skills call?
Descriptions in this topic reference models such as Veo 3.1, Veo 3, Seedance 2.0, HappyHorse 1.0, Wan 2.5, Grok Imagine Video, OmniHuman, Fabric, and HunyuanVideo. Availability varies by skill.
How is this list ranked?
Entries are ordered by observed search demand and reviewed for relevance to video generation. Inclusion is not an endorsement.

How to install an Agent Skill

Start with the first skill on this page. Claude Code reads a SKILL.md once it is in the skills directory:

  • Claude Code — copy the SKILL.md for kling-cli into ~/.claude/skills/kling-cli/ for yourself, or .claude/skills/kling-cli/ to share it with a project.
  • From the registry — run npx skills add klingai-tech/skills@kling-cli.
  • Or download kling-cli from its source and place that SKILL.md in the directory above.

How to choose Check the source and the skill’s own instructions before you install. Official and Notable describe provenance, not safety.