The open model that can actually spell.
Qwen Image is Alibaba's open image model, and its distinguishing strength is rendering text inside a picture — the failure mode that gives most open models away instantly. It handles both Latin and Chinese script with far fewer of the invented glyphs and dropped letters that SDXL-family models produce.
That makes it the value option for anything with words in it: signage, mockups, posters, UI screenshots, labels. It costs 168 credits against Grok Imagine 2.0's 420, so when the typography merely has to be right rather than designer-grade, this is the cheaper road to it. It accepts LoRAs too.
General image quality sits below Flux Dev. Grok Imagine 2.0 still handles typography better when the layout is complex — this is the value pick, not the best one.
| Price | 168 credits — about 15c |
|---|---|
| Type | Image |
| Provider | fal |
| Aspect ratios (13) | 1:1, 4:3, 16:9, 3:2, 19.5:9, 20:9, 2:1, 3:4, 9:16, 2:3, 9:19.5, 9:20, 1:2 |
| Resolutions | 1k, 2k |
| LoRA | Yes — Hugging Face or CivitAI adapter URLs |
Credit prices are the live studio rates. Cash equivalents are approximate at starter-pack rates and improve on the larger packs — see pricing.
Every AI model on comfyarts — image models, prices and measured capabilities side by side.