Update model card with benchmarks, badges, and workflow instructions
Browse files
README.md
CHANGED
|
@@ -2,45 +2,95 @@
|
|
| 2 |
license: other
|
| 3 |
license_name: qwen-research
|
| 4 |
license_link: https://huggingface.co/Qwen/Qwen-Image-2.1/blob/main/LICENSE
|
|
|
|
|
|
|
|
|
|
| 5 |
tags:
|
| 6 |
- comfyui
|
| 7 |
- gguf
|
| 8 |
-
-
|
|
|
|
|
|
|
|
|
|
| 9 |
- text-to-image
|
| 10 |
-
|
| 11 |
-
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 12 |
---
|
| 13 |
|
| 14 |
-
# Qwen-Image
|
| 15 |
|
| 16 |
-
|
| 17 |
|
| 18 |
-
|
|
|
|
|
|
|
|
|
|
| 19 |
|
| 20 |
-
|
| 21 |
|
| 22 |
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 23 |
| :--- | :--- | :--- | :--- |
|
| 24 |
-
| `qwen_image_2.1_Q8_0.gguf` |
|
|
|
|
|
|
|
| 25 |
|
| 26 |
-
|
| 27 |
|
| 28 |
-
|
| 29 |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 30 |
```
|
| 31 |
-
π ComfyUI/
|
| 32 |
-
βββ π models/
|
| 33 |
-
βββ π diffusion_models/ (or models/unet/)
|
| 34 |
-
β βββ qwen_image_2.1_Q8_0.gguf
|
| 35 |
-
βββ π text_encoders/
|
| 36 |
-
β βββ qwen3vl_8b_bf16.safetensors (or quantized variant)
|
| 37 |
-
βββ π vae/
|
| 38 |
-
βββ qwen_image_2.1_vae_bf16.safetensors
|
| 39 |
-
```
|
| 40 |
|
| 41 |
-
##
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 42 |
|
| 43 |
-
|
| 44 |
-
|
| 45 |
-
|
| 46 |
-
|
|
|
|
|
|
| 2 |
license: other
|
| 3 |
license_name: qwen-research
|
| 4 |
license_link: https://huggingface.co/Qwen/Qwen-Image-2.1/blob/main/LICENSE
|
| 5 |
+
language:
|
| 6 |
+
- en
|
| 7 |
+
- zh
|
| 8 |
tags:
|
| 9 |
- comfyui
|
| 10 |
- gguf
|
| 11 |
+
- qwen
|
| 12 |
+
- qwen2
|
| 13 |
+
- qwen-image
|
| 14 |
+
- qwen-image-2.1
|
| 15 |
- text-to-image
|
| 16 |
+
- image-to-image
|
| 17 |
+
- image-generation
|
| 18 |
+
- quantized
|
| 19 |
+
- diffusion
|
| 20 |
+
pipeline_tag: text-to-image
|
| 21 |
+
base_model: Qwen/Qwen-Image-2.1
|
| 22 |
+
base_model_relation: quantized
|
| 23 |
---
|
| 24 |
|
| 25 |
+
# Qwen-Image 2.1 GGUF (ComfyUI)
|
| 26 |
|
| 27 |
+
<div align="center">
|
| 28 |
|
| 29 |
+
[](https://huggingface.co/AlperKTS/Qwen-Image-2.1-GGUF)
|
| 30 |
+
[](https://huggingface.co/AlperKTS/Qwen-Image-2.1-GGUF)
|
| 31 |
+
[](https://github.com/city96/ComfyUI-GGUF)
|
| 32 |
+
[](https://huggingface.co/Qwen/Qwen-Image-2.1/blob/main/LICENSE)
|
| 33 |
|
| 34 |
+
</div>
|
| 35 |
|
| 36 |
+
Quantized GGUF weights of the state-of-the-art **Qwen-Image 2.1** diffusion model for **ComfyUI** using [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF).
|
| 37 |
+
|
| 38 |
+
These models allow running Qwen-Image 2.1 on consumer GPUs with significantly reduced VRAM footprint while preserving maximum visual quality, text-rendering capabilities, and composition fidelity.
|
| 39 |
+
|
| 40 |
+
---
|
| 41 |
+
|
| 42 |
+
## π¦ Quantization Versions
|
| 43 |
+
|
| 44 |
+
| File | Size | VRAM Recommended | Description |
|
| 45 |
| :--- | :--- | :--- | :--- |
|
| 46 |
+
| **`qwen_image_2.1_Q8_0.gguf`** | ~7.1 GB | 12GB+ | Near-lossless precision. Best quality & fidelity. |
|
| 47 |
+
| **`qwen_image_2.1_Q6_K.gguf`** | ~5.6 GB | 8GB - 12GB | Optimal balance of high quality and memory efficiency. |
|
| 48 |
+
| **`qwen_image_2.1_Q5_K_M.gguf`** | ~4.8 GB | 8GB | Fast inference, lowest VRAM consumption, excellent quality. |
|
| 49 |
|
| 50 |
+
*Note: The original BF16 diffusion model is ~14.2 GB.*
|
| 51 |
|
| 52 |
+
---
|
| 53 |
|
| 54 |
+
## π How to Use in ComfyUI
|
| 55 |
+
|
| 56 |
+
### 1. Install Prerequisites
|
| 57 |
+
Make sure you have [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF) installed in your `custom_nodes` folder:
|
| 58 |
+
```bash
|
| 59 |
+
cd ComfyUI/custom_nodes
|
| 60 |
+
git clone https://github.com/city96/ComfyUI-GGUF
|
| 61 |
```
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 62 |
|
| 63 |
+
### 2. Download Required Files
|
| 64 |
+
|
| 65 |
+
1. **Diffusion Model (GGUF):**
|
| 66 |
+
- Download any of the `.gguf` files from this repo (e.g. `qwen_image_2.1_Q8_0.gguf`).
|
| 67 |
+
- Place it in: `ComfyUI/models/diffusion_models/`
|
| 68 |
+
|
| 69 |
+
2. **VAE:**
|
| 70 |
+
- Download `qwen_image_2.1_vae_bf16.safetensors` from [Comfy-Org/Qwen-Image-2.1](https://huggingface.co/Comfy-Org/Qwen-Image-2.1/tree/main).
|
| 71 |
+
- Place it in: `ComfyUI/models/vae/`
|
| 72 |
+
|
| 73 |
+
3. **Text Encoder / CLIP:**
|
| 74 |
+
- Download `qwen_2.5_vl_7b_instruct_fp8_scaled.safetensors` (or compatible Qwen2.5-VL CLIP).
|
| 75 |
+
- Place it in: `ComfyUI/models/clip/`
|
| 76 |
+
|
| 77 |
+
---
|
| 78 |
+
|
| 79 |
+
## π¨ Workflows
|
| 80 |
+
|
| 81 |
+
Pre-configured workflows are included in the [`workflows/`](./workflows) folder:
|
| 82 |
+
- **`Qwen_Image_2.1_GGUF_Text_to_Image.json`**: Ready-to-run ComfyUI workflow using `UnetLoaderGGUF`.
|
| 83 |
+
- Drag and drop the workflow JSON directly into ComfyUI to start generating immediately.
|
| 84 |
+
|
| 85 |
+
### Key Node Configuration
|
| 86 |
+
- Node: **`UnetLoaderGGUF`** (from ComfyUI-GGUF)
|
| 87 |
+
- `unet_name`: Select `qwen_image_2.1_Q8_0.gguf` (or `Q6_K` / `Q5_K_M`)
|
| 88 |
+
- Connect to standard KSampler / ModelSamplingDiscrete nodes.
|
| 89 |
+
|
| 90 |
+
---
|
| 91 |
|
| 92 |
+
## π Credits & License
|
| 93 |
+
- Original model created by the **Qwen Team (Alibaba)**.
|
| 94 |
+
- GGUF conversion & quantization by **AlperKTS**.
|
| 95 |
+
- Compatible with [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF) by [city96](https://github.com/city96).
|
| 96 |
+
- Licensed under the original Qwen Research License.
|