Qwen3-VL LoRA Caption

Workflow allows you to use a Qwen3 VL model (2B-Instruct is good enough) to create description based prompt of an image is sees as well as tag based. The two are combined in the end to create a high quality and varied prompt. You will need to install any missing nodes so you can use it properly. First run will also download automatically the Qwen3 model that you select (default: 2B-Instruct)
Workflow Information
This workflow uses the following nodes:
- AILab_QwenVL
- MarkdownNote
- PreviewImage
- PrimitiveString
- WWAA_BuildString
- WWAA_DisplayAny
- WWAA_ImageLoader
- WWAA_PromptWriter
- WWAA_SearchReplaceMulti
Download Workflow
WWAA-Qwen3-VL-LoRA-Caption-v1.zip
530 downloads
