Skill-Details
kling-3-prompting
Vermittelt das Schreiben besserer Kling 3.0 Prompts für Text-to-Video, Image-to-Video, Keyframes, Multi-Shot und Dialogszenen über einen interaktiven Builder.
Vor Nutzung prüfen
Die automatische Prüfung bewertet Relevanz, nicht Sicherheit oder Empfehlung. Lies vor der Nutzung die Quellanweisungen.
SKILL.md
Dieser Auszug wurde bei der Prüfung gespeichert. Die externe Quelle enthält die vollständige und aktuelle Version.
---
name: kling-3-prompting
description: >
Write better prompts for Kling 3.0 AI video generation. Use when the user wants
to create, write, improve, or refine prompts — text-to-video, image-to-video,
keyframes, multi-shot sequences, or dialogue scenes.
metadata:
author: aedev-tools
version: "1.0"
---
## Overview
Kling 3.0 is a unified multimodal video model. It understands **cinematic direction**, not keyword lists. Write prompts like a director — describe what the audience sees, hears, and feels over time.
**Core shift:** Description → Direction. Think "direct a scene" not "describe an image."
## Interactive Builder Workflow
When invoked, guide the user through these steps using `AskUserQuestion`:
```dot
digraph builder {
"1. Generation mode?" [shape=diamond];
"Text-to-Video" [shape=box];
"Image-to-Video" [shape=box];
"Multi-Shot Sequence" [shape=box];
"Keyframe Transition" [shape=box];
"2. Gather scene details" [shape=box];
"3. Assemble prompt" [shape=box];
"4. Present & refine" [shape=box];
"1. Generation mode?" -> "Text-to-Video";
"1. Generation mode?" -> "Image-to-Video";
"1. Generation mode?" -> "Multi-Shot Sequence";
"1. Generation mode?" -> "Keyframe Transition";
"Text-to-Video" -> "2. Gather scene details";
"Image-to-Video" -> "2. Gather scene details";
"Multi-Shot Sequence" -> "2. Gather scene details";
"Keyframe Transition" -> "2. Gather scene details";
"2. Gather scene details" -> "3. Assemble prompt";
"3. Assemble prompt" -> "4. Present & refine";
}
```
### Step 1: Determine Generation Mode
Ask the user which mode:
- **Text-to-Video** — prompt from scratch
- **Image-to-Video** — animate a reference image
- **Multi-Shot Sequence** — 2-6 shot storyboard (up to 15s)
- **Keyframe Transition** — start frame → end frame with interpolated motion
### Step 2: Gather Scene Details
Ask about each element (adapt questions to mode):
| Element | Question | Why it matters |
|---------|----------|----------------|
| **Subject** | Who/what is the focus? Specific appearance details? | Anchors consistency — define distinguishing traits early |
| **Action** | What happens? Describe the timeline (first → then → finally) | Kling 3.0 excels at sequential action over 15s arcs |
| **Environment** | Where? Be specific (not "a street" but "narrow Tokyo alley, steam from grates") | Grounds the scene physically |
| **Camera** | Shot type and movement? (See camera reference below) | Cinematic language produces far better results |
| **Lighting** | What light sources? Name them specifically | "Flickering neon" beats "dramatic lighting" |
| **Mood/Emotion** | What should the audience feel? | Drives color grade, pacing, music |
| **Audio** | Dialogue? Ambient sound? Music? | Kling 3.0 generates native audio + lip-sync |
| **Duration** | How long? (3-15s) | Longer = describe progression over time |
| **Aspect Ratio** | 16:9 / 9:16 / 1:1 / 21:9? | 16:9 cinematic, 9:16 social, 21:9 ultra-wide |
**Image-to-Video:** Focus on how the scene *evolves from* the image — movement, camera motion, environmental change. The model preserves identity/layout from the source.
**Keyframes:** Ask for start and end frame descriptions. Frames should match in color, style, and lighting. Prompt sparingly — Kling infers motion well.
**Multi-Shot:** Define each shot separately with its own framing, subject, action, and duration. Label shots explicitly.
### Step 3: Assemble the Prompt
Use the **Master Formula**:
```
[Scene/Environment] + [Subject & Appearance] + [Action Timeline] + [Camera Movement] + [Audio & Atmosphere] + [Technical Specs]
```
**Writing rules:**
- Use cinematic motion verbs: dolly push, whip-pan, crash zoom, rack focus, tracking shot — NOT "moves" or "goes"
- Name real light sources: neon signs, candlelight, golden hour, LED panels — NOT "dramatic lighting"
- Include texture for credibility: grain, lens flares, condensation, fabric sheen, smoke, sweat
- Describe temporal flow: beginning → Vollständige Quelle auf GitHub lesen (öffnet externe Seite)