Viggle Turbo v0.2.1 — 6-step Qwen-Image-2.1

Text-to-image and image editing in 6 steps: about 5× faster than the 40-step Qwen-Image-2.1 and very competitive with it in quality — see the Comparison tab. Leave the references empty for text-to-image, or add up to 6 to edit, compose or transfer style.

Model card, weights and ComfyUI workflows: Viggle/Qwen-Image-2.1-viggle-turbo

A DMD-distilled student of Qwen-Image-2.1, sampled with no classifier-free guidance. On most prompts it is hard to tell apart from the base model; small, dense text (8 steps narrows the gap) and complicated edits (multi-reference composition, face swaps, identity-preserving edits) can still fall short of it. The Comparison tab has 32 examples of the official Qwen Space side by side, turbo vs base, same seed.

v0.2.1 (2026-09-24): the step-700 checkpoint of the v0.2 run on the 6-step schedule — intra-prompt diversity 0.98× the base model, 0% composition drift; earlier versions and the numbers are on the model card. Weights: v0.2.1 LoRA r256 from Viggle/Qwen-Image-2.1-viggle-turbo.

Output size

The menu switches to the editing sizes as soon as a reference is attached. Custom sizes above 2048² area (1536² when editing) are scaled down, keeping the ratio.

Enhance prompt

Rewrite the prompt with DeepSeek V4.1 Flash (+~2 s). Auto rewrites prompts under 30 tokens and sends longer ones as written. When it runs, the prompt and the reference images (downscaled to at most 0.5 MP) are sent through OpenRouter to a zero-retention host of that model; Off sends nothing. The text actually sent to the image model is shown under the result.

3 8
Examples · results pre-rendered by this model — click a row to load it with its settings; Generate reproduces it
Prompt References (optional, up to 6) Result

Model: Viggle/Qwen-Image-2.1-viggle-turbo · Built with Qwen — distilled from Qwen/Qwen-Image-2.1, which is released under the Qwen RESEARCH LICENSE AGREEMENT (non-commercial: research or evaluation purposes only). This demo inherits that restriction.