tooltd commited on
Commit
4bac3f0
·
verified ·
1 Parent(s): cf73508

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +2 -2
README.md CHANGED
@@ -51,7 +51,7 @@ All models were evaluated against the **BF16 baseline** (`Mean PPL = 6.950493`)
51
  | ZB4.36-STD-v4-IQ4_XS | ZB-STD| 13.88 | 0.021552| 93.886% | **6.993211**|
52
  | Q4_0-AutoRound-Code | webhie | 14.64 | 0.026586 | 92.970% | 7.067142 |
53
  | ZB4.14-MIN-IQ4_XS | ZB-MIN | 13.19 | 0.029334 | 92.799% | 7.045689 |
54
- | **⭐[ZB4.00-MIN-v5.1-IQ4_XS](https://huggingface.co/tooltd/Qwen3.8-27B-IQ4-XS-16GB-VRAM-GGUF/blob/main/Qwen3.8-27B-ZB4.00-MIN-v5.1-IQ4_XS.gguf)** | ZB-MIN | **12.79** | **0.033810**| **92.309%** | 7.106519|
55
  | ZB4.00-MIN-v5-IQ4_XS | ZB-MIN | 12.79 | 0.034577| 92.277% | **7.090583**|
56
  | **⭐[ZB3.88-MIN-v5-IQ3_M_XL](https://huggingface.co/tooltd/Qwen3.8-27B-IQ4-XS-16GB-VRAM-GGUF/blob/main/Qwen3.8-27B-ZB3.88-MIN-v5-IQ3_M_XL.gguf)** | ZB-MIN | **12.34** | **0.042164**| **91.452%** | **7.132594**|
57
  | **⭐[ZB3.70-MIN-v4-IQ3_M_L](https://huggingface.co/tooltd/Qwen3.8-27B-IQ4-XS-16GB-VRAM-GGUF/blob/main/Qwen3.8-27B-ZB3.70-MIN-v4-IQ3_M_L.gguf)** | ZB-MIN | 11.82 | 0.052972 | 90.363% | 7.160063 |
@@ -84,7 +84,7 @@ All models were evaluated against the **BF16 baseline** (`Mean PPL = 6.950493`)
84
  - **3.0BPW-IQ3_XXS**: New release size only 9.62 GB 🙀 with metric comparable to Ridge—3.7 bpw.
85
 
86
  **Update: Sep 5, 2026**
87
- - **ZB4.00-MIN-v5.1-IQ4_XS** This version has been updated so that all tensors are IQ3_XXS or higher, previous version contained some IQ2_S tensors. Quality is slightly improved, file size remains unchanged.
88
 
89
  **Recommended Settings:** Based on hands-on testing, set `reasoning_effort` to **medium**.
90
  At this BPW level, it delivers much more stable outputs and fits well in agentic workflows. Leaving it unrestricted makes the model overthink, burning through tokens and slowing things down to an annoying crawl.
 
51
  | ZB4.36-STD-v4-IQ4_XS | ZB-STD| 13.88 | 0.021552| 93.886% | **6.993211**|
52
  | Q4_0-AutoRound-Code | webhie | 14.64 | 0.026586 | 92.970% | 7.067142 |
53
  | ZB4.14-MIN-IQ4_XS | ZB-MIN | 13.19 | 0.029334 | 92.799% | 7.045689 |
54
+ | **⭐[ZB4.00-MIN-v5.1-IQ4_XS](https://huggingface.co/tooltd/Qwen3.8-27B-IQ4-XS-16GB-VRAM-GGUF/blob/main/Qwen3.8-27B-ZB4.00-MIN-v5.1-IQ4_XS.gguf)** | ZB-MIN | **12.74** | **0.033810**| **92.309%** | 7.106519|
55
  | ZB4.00-MIN-v5-IQ4_XS | ZB-MIN | 12.79 | 0.034577| 92.277% | **7.090583**|
56
  | **⭐[ZB3.88-MIN-v5-IQ3_M_XL](https://huggingface.co/tooltd/Qwen3.8-27B-IQ4-XS-16GB-VRAM-GGUF/blob/main/Qwen3.8-27B-ZB3.88-MIN-v5-IQ3_M_XL.gguf)** | ZB-MIN | **12.34** | **0.042164**| **91.452%** | **7.132594**|
57
  | **⭐[ZB3.70-MIN-v4-IQ3_M_L](https://huggingface.co/tooltd/Qwen3.8-27B-IQ4-XS-16GB-VRAM-GGUF/blob/main/Qwen3.8-27B-ZB3.70-MIN-v4-IQ3_M_L.gguf)** | ZB-MIN | 11.82 | 0.052972 | 90.363% | 7.160063 |
 
84
  - **3.0BPW-IQ3_XXS**: New release size only 9.62 GB 🙀 with metric comparable to Ridge—3.7 bpw.
85
 
86
  **Update: Sep 5, 2026**
87
+ - **ZB4.00-MIN-v5.1-IQ4_XS** This version has been updated so that all tensors are IQ3_XXS or higher, previous version contained some IQ2_S tensors. Quality is slightly improved, new file size saves 50MB.
88
 
89
  **Recommended Settings:** Based on hands-on testing, set `reasoning_effort` to **medium**.
90
  At this BPW level, it delivers much more stable outputs and fits well in agentic workflows. Leaving it unrestricted makes the model overthink, burning through tokens and slowing things down to an annoying crawl.