OpenAI Image Gen
Generate a handful of “random but structured” prompts and render them via the OpenAI Images API.
Run
Note: Image generation can take longer than common exec timeouts (for example 30 seconds).
When invoking this skill via Understudy’s exec tool, set a higher timeout to avoid premature termination/retries (e.g., exec timeout=300).
python3 {baseDir}/scripts/gen.py
open ~/Projects/tmp/openai-image-gen-*/index.html # if ~/Projects/tmp exists; else ./tmp/...
Useful flags:
# GPT image models with various options
python3 {baseDir}/scripts/gen.py --count 16 --model gpt-image-1
python3 {baseDir}/scripts/gen.py --prompt "ultra-detailed studio photo of a lobster astronaut" --count 4
python3 {baseDir}/scripts/gen.py --size 1536x1024 --quality high --out-dir ./out/images
python3 {baseDir}/scripts/gen.py --model gpt-image-1.5 --background transparent --output-format webp
# DALL-E 3 (note: count is automatically limited to 1)
python3 {baseDir}/scripts/gen.py --model dall-e-3 --quality hd --size 1792x1024 --style vivid
python3 {baseDir}/scripts/gen.py --model dall-e-3 --style natural --prompt "serene mountain landscape"
# DALL-E 2
python3 {baseDir}/scripts/gen.py --model dall-e-2 --size 512x512 --count 4
Model-Specific Parameters
Different models support different parameter values. The script automatically selects appropriate defaults based on the model.
Size
- GPT image models (
gpt-image-1, gpt-image-1-mini, gpt-image-1.5): 1024x1024, 1536x1024 (landscape), 1024x1536 (portrait), or auto
- dall-e-3:
1024x1024, 1792x1024, or 1024x1792
- dall-e-2:
256x256, 512x512, or 1024x1024
Quality
- GPT image models:
auto, high, medium, or low
- dall-e-3:
hd or standard
- dall-e-2:
standard only
Other Notable Differences
- dall-e-3 only supports generating 1 image at a time (
n=1). The script automatically limits count to 1 when using this model.
- GPT image models support additional parameters:
--background: transparent, opaque, or auto (default)
--output-format: png (default), jpeg, or webp
- Note:
stream and moderation are available via API but not yet implemented in this script
- dall-e-3 has a
--style parameter: vivid (hyper-real, dramatic) or natural (more natural looking)
Output
*.png, *.jpeg, or *.webp images (output format depends on model + --output-format)
prompts.json (prompt → file mapping)
index.html (thumbnail gallery)
1---2name: openai-image-gen3description: Batch-generate images via OpenAI Images API. Random prompt sampler + `index.html` gallery.4---56# OpenAI Image Gen78Generate a handful of “random but structured” prompts and render them via the OpenAI Images API.910## Run1112Note: Image generation can take longer than common exec timeouts (for example 30 seconds).13When invoking this skill via Understudy’s exec tool, set a higher timeout to avoid premature termination/retries (e.g., exec timeout=300).1415```bash16python3 {baseDir}/scripts/gen.py17open ~/Projects/tmp/openai-image-gen-*/index.html # if ~/Projects/tmp exists; else ./tmp/...18```1920Useful flags:2122```bash23# GPT image models with various options24python3 {baseDir}/scripts/gen.py --count 16 --model gpt-image-125python3 {baseDir}/scripts/gen.py --prompt "ultra-detailed studio photo of a lobster astronaut" --count 426python3 {baseDir}/scripts/gen.py --size 1536x1024 --quality high --out-dir ./out/images27python3 {baseDir}/scripts/gen.py --model gpt-image-1.5 --background transparent --output-format webp2829# DALL-E 3 (note: count is automatically limited to 1)30python3 {baseDir}/scripts/gen.py --model dall-e-3 --quality hd --size 1792x1024 --style vivid31python3 {baseDir}/scripts/gen.py --model dall-e-3 --style natural --prompt "serene mountain landscape"3233# DALL-E 234python3 {baseDir}/scripts/gen.py --model dall-e-2 --size 512x512 --count 435```3637## Model-Specific Parameters3839Different models support different parameter values. The script automatically selects appropriate defaults based on the model.4041### Size4243- **GPT image models** (`gpt-image-1`, `gpt-image-1-mini`, `gpt-image-1.5`): `1024x1024`, `1536x1024` (landscape), `1024x1536` (portrait), or `auto`44 - Default: `1024x1024`45- **dall-e-3**: `1024x1024`, `1792x1024`, or `1024x1792`46 - Default: `1024x1024`47- **dall-e-2**: `256x256`, `512x512`, or `1024x1024`48 - Default: `1024x1024`4950### Quality5152- **GPT image models**: `auto`, `high`, `medium`, or `low`53 - Default: `high`54- **dall-e-3**: `hd` or `standard`55 - Default: `standard`56- **dall-e-2**: `standard` only57 - Default: `standard`5859### Other Notable Differences6061- **dall-e-3** only supports generating 1 image at a time (`n=1`). The script automatically limits count to 1 when using this model.62- **GPT image models** support additional parameters:63 - `--background`: `transparent`, `opaque`, or `auto` (default)64 - `--output-format`: `png` (default), `jpeg`, or `webp`65 - Note: `stream` and `moderation` are available via API but not yet implemented in this script66- **dall-e-3** has a `--style` parameter: `vivid` (hyper-real, dramatic) or `natural` (more natural looking)6768## Output6970- `*.png`, `*.jpeg`, or `*.webp` images (output format depends on model + `--output-format`)71- `prompts.json` (prompt → file mapping)72- `index.html` (thumbnail gallery)