local/qwen-image-2.1-7b-nf4
LocalQwen-Image-2.1 is Alibaba Qwen's open-weight text-to-image and image-edit model, released 2026-09-20 under the non-commercial Qwen Research License. It pairs a 32-layer single-stream 7B DiT with a Qwen3-VL 8B text encoder and a 64-channel RGBA autoencoder, so transparent cutouts need no matting pass. This entry benchmarks the NF4 configuration: the 7B diffusion transformer runs at full BF16 precision and only the 8B text encoder is quantised to NF4 (bitsandbytes). Quantising just the encoder cuts the pipeline from roughly 31 GiB to 20.5 GiB, which fits a 32 GB consumer card without CPU offload. Image quality therefore comes from an unquantised denoiser; NF4 affects only how the prompt is encoded.
- Maker
- Alibaba / Qwen
- Family
- Qwen-Image
- Model Size
- 7B DiT (BF16) + 8B text encoder (NF4)high
- Cost
- local run; no API pricenot_applicable
- Run Target
- 5090/Qwen/Qwen-Image-2.1
- Effective Request
- response_format: b64_json · size: 1024x1024 · num_inference_steps: 40 · true_cfg_scale: 1 · seed: 7