ComfyUI Node: Qwen3 VQA

Authored by IuvenisSapiens

Created

Updated

566 stars

Run ComfyUI workflows without the setup

No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.

Category

Comfyui_Qwen3-VL-Instruct

Inputs

text STRING
model
  • Qwen3-VL-4B-Instruct-FP8
  • Qwen3-VL-4B-Thinking-FP8
  • Qwen3-VL-8B-Instruct-FP8
  • Qwen3-VL-8B-Thinking-FP8
  • Qwen3-VL-4B-Instruct
  • Qwen3-VL-4B-Thinking
  • Qwen3-VL-8B-Instruct
  • Qwen3-VL-8B-Thinking
quantization
  • none
  • 4bit
  • 8bit
keep_model_loaded BOOLEAN
temperature FLOAT
max_new_tokens INT
min_pixels INT
max_pixels INT
seed INT
attention
  • eager
  • sdpa
  • flash_attention_2
source_path PATH
image IMAGE

Outputs

STRING

Extension: ComfyUI_Qwen3-VL-Instruct

The successful integration of Qwen3-VL-Instruct series into the ComfyUI platform has enabled a smooth operation, supporting (but not limited to) text-based queries, video queries, single-image queries, and multi-image queries for generating captions or responses.

Authored by IuvenisSapiens

Looking for a different node?

Also provided by 1 other extension

Run ComfyUI workflows without the setup

No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.

Learn more