Skill 详情
glmv-prompt-gen
分析图像/视频,并为 Midjourney、Stable Diffusion、DALL-E、Sora 和 Kling 等文生图和文生视频工具生成专业提示词。
声明的前提(自述): API key · Python
使用前先检查
自动化审核只检查相关性,不代表安全审查或推荐。使用前请阅读来源中的说明。
SKILL.md
这段内容是审核时保存的快照。外部来源才是完整且最新的版本。
---
name: glmv-prompt-gen
description:
Analyze images/videos and generate professional prompts for text-to-image and
text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika).
Use when the user wants to generate prompts from reference images/videos, create
AI art prompts, or get prompt engineering suggestions from visual content.
metadata:
openclaw:
requires:
env:
- ZHIPU_API_KEY
bins:
- python
primaryEnv: ZHIPU_API_KEY
emoji: "✨"
homepage: https://github.com/zai-org/GLM-V/tree/main/skills/glmv-prompt-gen
---
# GLM-V Prompt Generation Skill
Analyze reference images or videos and generate professional prompts for AI image/video generation tools.
## When to Use
- Generate prompts for text-to-image tools (Midjourney, Stable Diffusion, DALL-E, etc.)
- Generate prompts for text-to-video tools (Sora, Runway, Kling, Pika, etc.)
- User mentions "生成prompt", "文生图prompt", "文生视频prompt", "prompt工程", "参考图生成prompt", "generate prompt"
- User provides an image/video and wants to recreate or remix it
- Extract prompt ideas from reference visual content
## Supported Input Types
| Type | Formats | Max Size | Max Count | Base64 |
| ----- | -------------- | ----------------- | --------- | ------------- |
| Image | jpg, png, jpeg | 5MB / 6000×6000px | 50 | ✅ |
| Video | mp4, mkv, mov | 200MB | — | ❌ (URL only) |
> ⚠️ Images and videos cannot be used in the same request.
> ⚠️ Videos only support URLs — local paths and base64 are NOT supported.
### 📋 Output Display Rules (MANDATORY)
After running the script, **you must display the full prompt output exactly as returned**. Do not summarize, truncate, or only say "prompt generated". Users need the complete prompt (especially the English prompt) for direct copy/paste.
- Show the full output: content analysis + prompt + prompt breakdown
- In `auto` mode, show both text-to-image and text-to-video prompts
- English prompts are core output and must be shown completely
- If output was saved (`-o`), provide the file path and show file content
## Output Modes
| Mode | Description |
| ------- | -------------------------------------------------- |
| `image` | Generate prompts for text-to-image tools (default) |
| `video` | Generate prompts for text-to-video tools |
| `auto` | Generate prompts for both image and video |
## Resource Links
| Resource | Link |
| --------------- | --------------------------------------------------------------------------------------------------------------------------------- |
| **Get API Key** | [https://bigmodel.cn/usercenter/proj-mgmt/apikeys](https://bigmodel.cn/usercenter/proj-mgmt/apikeys) |
| **API Docs** | [Chat Completions / 对话补全](https://docs.bigmodel.cn/api-reference/%E6%A8%A1%E5%9E%8B-api/%E5%AF%B9%E8%AF%9D%E8%A1%A5%E5%85%A8) |
## Prerequisites
### API Key Setup / API Key 配置(Required / 必需)
This script reads the key from the `ZHIPU_API_KEY` environment variable and shares it with other Zhipu skills.
脚本通过 `ZHIPU_API_KEY` 环境变量获取密钥,与其他智谱技能共用同一个 key。
**Get Key / 获取 Key:** Visit [Zhipu Open Platform API Keys / 智谱开放平台 API Keys](https://bigmodel.cn/usercenter/proj-mgmt/apikeys) to create or copy your key.
**Setup options / 配置方式(任选一种):**
1. **OpenClaw config (recommended) / OpenClaw 配置(推荐):** Set in `openclaw.json` under `skills.entries.glmv-prompt-gen.env`:
```json
"glmv-prompt-gen": { "enabled": true, "env": { "ZHIPU_API_KEY": "你的密钥" } }
```
2. **Shell environment variable / Shell 环境变量:** Add to `~/.zshrc`:
```bash
export ZHIPU_API_KEY="你的密钥"
```
> 💡 If you already configured another Zhipu skill (for example `zhipu-tools` or `glmv-caption`), they s