Visual Design — Image Generation & Editing
Generate new images or edit existing ones via the backend image API.
Be patient, it takes about 2 minutes to generate an image each time.
Prerequisites
API Key Configuration (Required First)
This skill requires a SKYWORK_API_KEY to be configured in OpenClaw.
If you don't have an API key yet, please visit:
https://skywork.ai
For detailed setup instructions, see:
references/apikey-fetch.md
Usage
Run the script using absolute path (do NOT cd to skill directory):
Generate new image:
python3 <SKILL_DIR>/scripts/generate_image.py --prompt "description" --filename "output.png" [--aspect-ratio 3:4] [--resolution 1K|2K|4K]
Edit existing image:
python3 <SKILL_DIR>/scripts/generate_image.py --prompt "edit instructions" --filename "output.png" --input-image "source.png" [--aspect-ratio 3:4] [--resolution 2K]
Edit with multiple reference images:
python3 <SKILL_DIR>/scripts/generate_image.py --prompt "combine these styles" --filename "output.png" -i "ref1.png" -i "ref2.png"
Always run from the user's working directory so images save there.
When to Generate vs Edit
- Generation (
--prompt only): Creating new images from scratch — posters, logos, illustrations, photos, infographics.
- Editing (
--prompt + --input-image): User provides existing image(s) and wants modifications — style changes, element addition/removal, color adjustments, format conversion.
- Notice: Edit api supports character resemblance of up to 4 characters and the fidelity of up to 10 objects in a single workflow
If the user uploads/references images and wants changes, always use --input-image.
Resolution
- 1K — ~1024px, fast drafts
- 2K (default) — ~2048px, good for most deliverables
- 4K — ~4096px, final high-res output
Map user requests: "low/draft" → 1K, "normal/medium/2K" → 2K, "high-res/hi-res/4K/ultra" → 4K.
Aspect Ratio
Supported ratios: 1:1, 2:3, 3:2, 3:4, 4:3, 4:5, 5:4, 9:16, 16:9, 21:9.
Selection guidance:
- 1:1 — Social media avatars, icons, album covers
- 3:4 / 4:3 — General posters, presentations
- 4:5 / 5:4 — Instagram posts, portraits
- 9:16 / 16:9 — Mobile stories / desktop wallpapers, video covers
- 2:3 / 3:2 — Print posters, book covers
- 21:9 — Ultra-wide banners, cinema format
If the user doesn't specify, omit --aspect-ratio and let the API decide.
Filename Convention
Pattern: yyyy-mm-dd-hh-mm-ss-descriptive-name.png
Examples:
- "A serene Japanese garden" →
2026-03-10-14-23-05-japanese-garden.png
- "sunset over mountains" →
2026-03-10-15-30-12-sunset-mountains.png
- Unclear context →
2026-03-10-17-12-48-x9k2.png
Preflight
command -v python3 (must exist)
- If editing: verify each
--input-image file exists
Common Failures
Error: Input image not found: → wrong path; verify --input-image file exists
HTTP error 500 → backend service error; retry or check server status
Request failed: Connection refused → backend service unavailable; ensure the service is running
- Insufficient benefit: The script or log may show a message like
Insufficient benefit. Please upgrade your account at {url}, meaning the user's benefit level does not meet the requirement for this skill.
How to reply when benefit is insufficient
When you detect the above, reply in the user's current language — do not echo the English message. Use this pattern:
- Convey: "Sorry, image generation failed. This skill requires upgrading your Skywork membership to use." then a single call-to-action link.
- Format: One short sentence in the user's language + a link like
[Upgrade now →](url) or the equivalent in their language.
- URL: Extract the upgrade URL from the log/script output (e.g. the
at https://... part).
Note: Only suggest upgrading when the error is Insufficient benefit. For auth errors like NO_TOKEN / INVALID_TOKEN / 401 / “invalid API key”, keep the error code / raw message and guide users to update SKYWORK_API_KEY. Do not suggest upgrading membership.
Output
- Script prints the local file path and the OSS URL.
- Depending on the platform, use the most appropriate way to deliver the image (e.g. send as image message, display inline, or print the URLs). By default, return both the local path and OSS URL to the user. The OSS URL ensures cross-platform accessibility.
Design Scenarios
Match the user's request to a scenario and read the corresponding file for specialized workflow:
- E-commerce product image: See scenarios/e-commerce.md
- Storyboard: See scenarios/storyboard.md
- Infographic: See scenarios/infographic.md
- Logo: See scenarios/logo.md
- Branding / VI: See scenarios/branding.md
- Brochure: See scenarios/brochure.md
- Social media: See scenarios/social-media.md
- Poster: See scenarios/poster.md
Prompt Engineering
Prompts Best Practices
Follow these principles for quality prompts using the image API for generation or editing:
- Describe the scene, don't just list keywords. A narrative, descriptive paragraph produces much better results than disconnected words. The model's core strength is deep language understanding.
- Weak: "cat, sunset, beach"
- Strong: "A ginger tabby cat sitting on a sandy beach at golden hour, facing the camera with soft warm backlighting, shallow depth of field, ocean waves blurred in the background"
- Be hyper-specific. The more detail you provide, the more control you have. Include all visual details: style, colors, composition, lighting, background, textures.
- Provide context and intent. Explain the purpose of the image — the model's understanding of context influences the output.
- Use step-by-step instructions for complex scenes with many elements. Break the prompt into layers: foreground, middle ground, background.
- Use "semantic negative prompts." Instead of "no cars," describe positively: "an empty, deserted street with no signs of traffic."
- Control the camera. Use photographic and cinematic terms: "wide-angle shot", "macro shot", "low-angle perspective", "bird's eye view", "rule of thirds", "shallow depth of field".
- Time perception. If the result needs real-time timeliness, mention the current time context in the prompt.
- Text in images. Place text content within double quotation marks:
A movie poster with the title "INCEPTION" in large silver metallic letters at the top
- Clearly specify and emphasize the elements that require modification. Describe reference images by their order (first image, second image), not by filename.
1---2name: skywork-design3description: Generate or edit images via backend Skywork Image API. Use for any image creation, poster design, logo design, visual asset generation, or image modification request. Supports text-to-image and image-to-image editing with aspect ratio and resolution control.4---5
6# Visual Design — Image Generation & Editing
7
8Generate new images or edit existing ones via the backend image API.
9Be patient, it takes about 2 minutes to generate an image each time.
10
11---
12
13## Prerequisites
14
15### API Key Configuration (Required First)
16This skill requires a **SKYWORK_API_KEY** to be configured in OpenClaw.
17
18If you don't have an API key yet, please visit:
19**https://skywork.ai**
20
21For detailed setup instructions, see:
22[references/apikey-fetch.md](references/apikey-fetch.md)
23
24## Usage
25
26Run the script using absolute path (do NOT cd to skill directory):
27
28**Generate new image:**
29```bash
30python3 <SKILL_DIR>/scripts/generate_image.py --prompt "description" --filename "output.png" [--aspect-ratio 3:4] [--resolution 1K|2K|4K]
31```
32
33**Edit existing image:**
34```bash
35python3 <SKILL_DIR>/scripts/generate_image.py --prompt "edit instructions" --filename "output.png" --input-image "source.png" [--aspect-ratio 3:4] [--resolution 2K]
36```
37
38**Edit with multiple reference images:**
39```bash
40python3 <SKILL_DIR>/scripts/generate_image.py --prompt "combine these styles" --filename "output.png" -i "ref1.png" -i "ref2.png"
41```
42
43Always run from the user's working directory so images save there.
44
45## When to Generate vs Edit
46
47- **Generation** (`--prompt` only): Creating new images from scratch — posters, logos, illustrations, photos, infographics.
48- **Editing** (`--prompt` + `--input-image`): User provides existing image(s) and wants modifications — style changes, element addition/removal, color adjustments, format conversion.
49 - Notice: Edit api supports character resemblance of up to 4 characters and the fidelity of up to 10 objects in a single workflow
50
51If the user uploads/references images and wants changes, always use `--input-image`.
52
53## Resolution
54
55- **1K** — ~1024px, fast drafts
56- **2K** (default) — ~2048px, good for most deliverables
57- **4K** — ~4096px, final high-res output
58
59Map user requests: "low/draft" → 1K, "normal/medium/2K" → 2K, "high-res/hi-res/4K/ultra" → 4K.
60
61## Aspect Ratio
62
63Supported ratios: `1:1`, `2:3`, `3:2`, `3:4`, `4:3`, `4:5`, `5:4`, `9:16`, `16:9`, `21:9`.
64
65Selection guidance:
66- **1:1** — Social media avatars, icons, album covers
67- **3:4 / 4:3** — General posters, presentations
68- **4:5 / 5:4** — Instagram posts, portraits
69- **9:16 / 16:9** — Mobile stories / desktop wallpapers, video covers
70- **2:3 / 3:2** — Print posters, book covers
71- **21:9** — Ultra-wide banners, cinema format
72
73If the user doesn't specify, omit `--aspect-ratio` and let the API decide.
74
75## Filename Convention
76
77Pattern: `yyyy-mm-dd-hh-mm-ss-descriptive-name.png`
78
79Examples:
80- "A serene Japanese garden" → `2026-03-10-14-23-05-japanese-garden.png`
81- "sunset over mountains" → `2026-03-10-15-30-12-sunset-mountains.png`
82- Unclear context → `2026-03-10-17-12-48-x9k2.png`
83
84## Preflight
85
86- `command -v python3` (must exist)
87- If editing: verify each `--input-image` file exists
88
89## Common Failures
90
91- `Error: Input image not found:` → wrong path; verify `--input-image` file exists
92- `HTTP error 500` → backend service error; retry or check server status
93- `Request failed: Connection refused` → backend service unavailable; ensure the service is running
94- **Insufficient benefit**: The script or log may show a message like `Insufficient benefit. Please upgrade your account at {url}`, meaning the user's benefit level does not meet the requirement for this skill.
95
96### How to reply when benefit is insufficient
97
98When you detect the above, **reply in the user's current language** — do not echo the English message. Use this pattern:
99
100- Convey: "Sorry, image generation failed. This skill requires upgrading your Skywork membership to use." then a single call-to-action link.
101- **Format**: One short sentence in the user's language + a link like `[Upgrade now →](url)` or the equivalent in their language.
102- **URL**: Extract the upgrade URL from the log/script output (e.g. the `at https://...` part).
103
104> Note: Only suggest upgrading when the error is **Insufficient benefit**. For auth errors like `NO_TOKEN` / `INVALID_TOKEN` / `401` / “invalid API key”, keep the error code / raw message and guide users to update `SKYWORK_API_KEY`. **Do not** suggest upgrading membership.
105
106## Output
107
108- Script prints the local file path and the OSS URL.
109- Depending on the platform, use the most appropriate way to deliver the image (e.g. send as image message, display inline, or print the URLs). By default, return both the local path and OSS URL to the user. The OSS URL ensures cross-platform accessibility.
110
111## Design Scenarios
112
113Match the user's request to a scenario and read the corresponding file for specialized workflow:
114
115- **E-commerce product image**: See [scenarios/e-commerce.md](scenarios/e-commerce.md)
116- **Storyboard**: See [scenarios/storyboard.md](scenarios/storyboard.md)
117- **Infographic**: See [scenarios/infographic.md](scenarios/infographic.md)
118- **Logo**: See [scenarios/logo.md](scenarios/logo.md)
119- **Branding / VI**: See [scenarios/branding.md](scenarios/branding.md)
120- **Brochure**: See [scenarios/brochure.md](scenarios/brochure.md)
121- **Social media**: See [scenarios/social-media.md](scenarios/social-media.md)
122- **Poster**: See [scenarios/poster.md](scenarios/poster.md)
123
124## Prompt Engineering
125
126### Prompts Best Practices
127
128Follow these principles for quality prompts using the image API for generation or editing:
129
130- **Describe the scene, don't just list keywords.** A narrative, descriptive paragraph produces much better results than disconnected words. The model's core strength is deep language understanding.
131 - Weak: "cat, sunset, beach"
132 - Strong: "A ginger tabby cat sitting on a sandy beach at golden hour, facing the camera with soft warm backlighting, shallow depth of field, ocean waves blurred in the background"
133- **Be hyper-specific.** The more detail you provide, the more control you have. Include all visual details: style, colors, composition, lighting, background, textures.
134- **Provide context and intent.** Explain the purpose of the image — the model's understanding of context influences the output.
135- **Use step-by-step instructions** for complex scenes with many elements. Break the prompt into layers: foreground, middle ground, background.
136- **Use "semantic negative prompts."** Instead of "no cars," describe positively: "an empty, deserted street with no signs of traffic."
137- **Control the camera.** Use photographic and cinematic terms: "wide-angle shot", "macro shot", "low-angle perspective", "bird's eye view", "rule of thirds", "shallow depth of field".
138- **Time perception.** If the result needs real-time timeliness, mention the current time context in the prompt.
139- **Text in images.** Place text content within double quotation marks:
140 > A movie poster with the title "INCEPTION" in large silver metallic letters at the top
141- Clearly specify and emphasize the elements that require modification. Describe reference images by their order (first image, second image), not by filename.