workflow_qwenvl_image_to_text
Imported workflow skill generated from QwenVL Image to Text.json.
Family
txt2img
Inputs
- Optional runtime overrides supported by
run(...):promptnegative_promptwidth,heightseed,steps,cfgsampler_name,scheduler,denoiseserver,headers,api_prefix
Outputs
- Returns JSON with:
statusprompt_idoutput_images
Model Requirements
- None detected from loader nodes.
Custom Node Requirements
comfyui-qwenvl
Links
- https://discord.com/invite/gggpkVgBf3
- https://github.com/1038lab/ComfyUI-QwenVL
- https://www.youtube.com/@pixaroma
Routing Metadata
- Family:
txt2img - Input modalities:
image - Output modalities:
application/json - Model families:
qwen, wan - Node count:
5 - Complexity score:
1 - Resource profile:
low - Estimated runtime:
fast (usually under 30s on modern GPU) - Max latent resolution hint:
NonexNone - Max sampler steps hint:
None
Detected Models
- None detected.
Detected Custom Nodes
comfyui-qwenvl
Runtime Warnings
- Uses custom nodes; missing nodes can cause validation/runtime failures.