tooltd commited on
Commit
66d867f
·
verified ·
1 Parent(s): 1b50a3d

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +9 -4
README.md CHANGED
@@ -64,10 +64,15 @@ All models were evaluated against the **BF16 baseline** (`Mean PPL = 6.950493`)
64
  - The newly released Unsloth Dynamic v3 is truly the best value for performance right now.
65
  - My ZB is just an experiment, feel free to check it out for fun :)
66
 
67
- **Update: Aug 23, 2026**
68
- - Quant release **⭐24** | ZBv2-4.00 BPW runs cleanly on *16GB VRAM with MTP support and up to 95K context length.*
69
- - **⭐25** | **ZBv2-3.7BPW** ⚔️ **empero-ai/Qwen3.8-27B-Ridge.** 🤣🤣🤣
70
- - **Recommended Settings:** Based on hands-on testing, set `reasoning_effort` to **medium**.
 
 
 
 
 
71
  At this BPW level, it delivers much more stable outputs and fits well in agentic workflows. Leaving it unrestricted makes the model overthink, burning through tokens and slowing things down to an annoying crawl.
72
 
73
  `
 
64
  - The newly released Unsloth Dynamic v3 is truly the best value for performance right now.
65
  - My ZB is just an experiment, feel free to check it out for fun :)
66
 
67
+ **Update: Aug 22, 2026**
68
+ - Quant release **⭐24** | ZB-4.00 BPW runs cleanly on *16GB VRAM with MTP support and up to 95K context length.*
69
+
70
+ **Update: Aug 24, 2026**
71
+ - **⭐23** | **ZBv2-4.00BPW** update maintaining size with KLD and Same top-p performs slightly better.
72
+ - **⭐25** | **ZBv2-3.7BPW** ⚔️ **empero-ai/Qwen3.8-27B-Ridge.** 🤣🤣🤣 (Uploading...)
73
+
74
+
75
+ **Recommended Settings:** Based on hands-on testing, set `reasoning_effort` to **medium**.
76
  At this BPW level, it delivers much more stable outputs and fits well in agentic workflows. Leaving it unrestricted makes the model overthink, burning through tokens and slowing things down to an annoying crawl.
77
 
78
  `