AlperKTS commited on
Commit
64c30d8
Β·
verified Β·
1 Parent(s): 79700a6

Update model card with benchmarks, badges, and workflow instructions

Browse files
Files changed (1) hide show
  1. README.md +75 -25
README.md CHANGED
@@ -2,45 +2,95 @@
2
  license: other
3
  license_name: qwen-research
4
  license_link: https://huggingface.co/Qwen/Qwen-Image-2.1/blob/main/LICENSE
 
 
 
5
  tags:
6
  - comfyui
7
  - gguf
8
- - diffusion-single-file
 
 
 
9
  - text-to-image
10
- base_model:
11
- - Qwen/Qwen-Image-2.1
 
 
 
 
 
12
  ---
13
 
14
- # Qwen-Image-2.1-GGUF
15
 
16
- Direct GGUF conversion of [Qwen/Qwen-Image-2.1](https://huggingface.co/Qwen/Qwen-Image-2.1) ([Comfy-Org repackaged weights](https://huggingface.co/Comfy-Org/Qwen-Image-2.1)).
17
 
18
- These model files are optimized for use in [ComfyUI](https://github.com/comfyanonymous/ComfyUI) using the [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF) custom node ().
 
 
 
19
 
20
- ## Models
21
 
22
- | File Name | Quantization | Size | Description |
 
 
 
 
 
 
 
 
23
  | :--- | :--- | :--- | :--- |
24
- | `qwen_image_2.1_Q8_0.gguf` | Q8_0 | ~7.10 GB | 8-bit quantized model preserving 99.9% quality, with F32 1D norms and BF16 sensitive I/O projection layers |
 
 
25
 
26
- ## Setup & File Placement in ComfyUI
27
 
28
- Place the downloaded files in your ComfyUI directory:
29
 
 
 
 
 
 
 
 
30
  ```
31
- πŸ“‚ ComfyUI/
32
- └── πŸ“‚ models/
33
- β”œβ”€β”€ πŸ“‚ diffusion_models/ (or models/unet/)
34
- β”‚ └── qwen_image_2.1_Q8_0.gguf
35
- β”œβ”€β”€ πŸ“‚ text_encoders/
36
- β”‚ └── qwen3vl_8b_bf16.safetensors (or quantized variant)
37
- └── πŸ“‚ vae/
38
- └── qwen_image_2.1_vae_bf16.safetensors
39
- ```
40
 
41
- ## Usage in ComfyUI
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
42
 
43
- 1. Install [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF).
44
- 2. Use the **`Unet Loader (GGUF)`** node instead of the standard `UNETLoader` or `DiffusionModelLoader`.
45
- 3. Select `qwen_image_2.1_Q8_0.gguf`.
46
- 4. Connect the MODEL output to your KSampler / Guider pipeline.
 
 
2
  license: other
3
  license_name: qwen-research
4
  license_link: https://huggingface.co/Qwen/Qwen-Image-2.1/blob/main/LICENSE
5
+ language:
6
+ - en
7
+ - zh
8
  tags:
9
  - comfyui
10
  - gguf
11
+ - qwen
12
+ - qwen2
13
+ - qwen-image
14
+ - qwen-image-2.1
15
  - text-to-image
16
+ - image-to-image
17
+ - image-generation
18
+ - quantized
19
+ - diffusion
20
+ pipeline_tag: text-to-image
21
+ base_model: Qwen/Qwen-Image-2.1
22
+ base_model_relation: quantized
23
  ---
24
 
25
+ # Qwen-Image 2.1 GGUF (ComfyUI)
26
 
27
+ <div align="center">
28
 
29
+ [![Downloads](https://img.shields.io/badge/dynamic/json?url=https%3A%2F%2Fhuggingface.co%2Fapi%2Fmodels%2FAlperKTS%2FQwen-Image-2.1-GGUF&query=%24.downloads&label=Downloads&color=blue&logo=huggingface)](https://huggingface.co/AlperKTS/Qwen-Image-2.1-GGUF)
30
+ [![Likes](https://img.shields.io/badge/dynamic/json?url=https%3A%2F%2Fhuggingface.co%2Fapi%2Fmodels%2FAlperKTS%2FQwen-Image-2.1-GGUF&query=%24.likes&label=Likes&color=red&logo=huggingface)](https://huggingface.co/AlperKTS/Qwen-Image-2.1-GGUF)
31
+ [![ComfyUI](https://img.shields.io/badge/ComfyUI-GGUF-green.svg)](https://github.com/city96/ComfyUI-GGUF)
32
+ [![License](https://img.shields.io/badge/License-Qwen_Research-blue.svg)](https://huggingface.co/Qwen/Qwen-Image-2.1/blob/main/LICENSE)
33
 
34
+ </div>
35
 
36
+ Quantized GGUF weights of the state-of-the-art **Qwen-Image 2.1** diffusion model for **ComfyUI** using [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF).
37
+
38
+ These models allow running Qwen-Image 2.1 on consumer GPUs with significantly reduced VRAM footprint while preserving maximum visual quality, text-rendering capabilities, and composition fidelity.
39
+
40
+ ---
41
+
42
+ ## πŸ“¦ Quantization Versions
43
+
44
+ | File | Size | VRAM Recommended | Description |
45
  | :--- | :--- | :--- | :--- |
46
+ | **`qwen_image_2.1_Q8_0.gguf`** | ~7.1 GB | 12GB+ | Near-lossless precision. Best quality & fidelity. |
47
+ | **`qwen_image_2.1_Q6_K.gguf`** | ~5.6 GB | 8GB - 12GB | Optimal balance of high quality and memory efficiency. |
48
+ | **`qwen_image_2.1_Q5_K_M.gguf`** | ~4.8 GB | 8GB | Fast inference, lowest VRAM consumption, excellent quality. |
49
 
50
+ *Note: The original BF16 diffusion model is ~14.2 GB.*
51
 
52
+ ---
53
 
54
+ ## πŸš€ How to Use in ComfyUI
55
+
56
+ ### 1. Install Prerequisites
57
+ Make sure you have [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF) installed in your `custom_nodes` folder:
58
+ ```bash
59
+ cd ComfyUI/custom_nodes
60
+ git clone https://github.com/city96/ComfyUI-GGUF
61
  ```
 
 
 
 
 
 
 
 
 
62
 
63
+ ### 2. Download Required Files
64
+
65
+ 1. **Diffusion Model (GGUF):**
66
+ - Download any of the `.gguf` files from this repo (e.g. `qwen_image_2.1_Q8_0.gguf`).
67
+ - Place it in: `ComfyUI/models/diffusion_models/`
68
+
69
+ 2. **VAE:**
70
+ - Download `qwen_image_2.1_vae_bf16.safetensors` from [Comfy-Org/Qwen-Image-2.1](https://huggingface.co/Comfy-Org/Qwen-Image-2.1/tree/main).
71
+ - Place it in: `ComfyUI/models/vae/`
72
+
73
+ 3. **Text Encoder / CLIP:**
74
+ - Download `qwen_2.5_vl_7b_instruct_fp8_scaled.safetensors` (or compatible Qwen2.5-VL CLIP).
75
+ - Place it in: `ComfyUI/models/clip/`
76
+
77
+ ---
78
+
79
+ ## 🎨 Workflows
80
+
81
+ Pre-configured workflows are included in the [`workflows/`](./workflows) folder:
82
+ - **`Qwen_Image_2.1_GGUF_Text_to_Image.json`**: Ready-to-run ComfyUI workflow using `UnetLoaderGGUF`.
83
+ - Drag and drop the workflow JSON directly into ComfyUI to start generating immediately.
84
+
85
+ ### Key Node Configuration
86
+ - Node: **`UnetLoaderGGUF`** (from ComfyUI-GGUF)
87
+ - `unet_name`: Select `qwen_image_2.1_Q8_0.gguf` (or `Q6_K` / `Q5_K_M`)
88
+ - Connect to standard KSampler / ModelSamplingDiscrete nodes.
89
+
90
+ ---
91
 
92
+ ## πŸ“œ Credits & License
93
+ - Original model created by the **Qwen Team (Alibaba)**.
94
+ - GGUF conversion & quantization by **AlperKTS**.
95
+ - Compatible with [ComfyUI-GGUF](https://github.com/city96/ComfyUI-GGUF) by [city96](https://github.com/city96).
96
+ - Licensed under the original Qwen Research License.