anlord commited on
Commit
fcf980a
Β·
verified Β·
1 Parent(s): 8066131

Upload folder using huggingface_hub

Browse files
.gitattributes CHANGED
@@ -33,3 +33,11 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
 
 
 
 
 
 
 
 
 
33
  *.zip filter=lfs diff=lfs merge=lfs -text
34
  *.zst filter=lfs diff=lfs merge=lfs -text
35
  *tfevents* filter=lfs diff=lfs merge=lfs -text
36
+ qwen3.5-0.8B-abliterated-bf16.gguf filter=lfs diff=lfs merge=lfs -text
37
+ qwen3.5-0.8B-abliterated-f16.gguf filter=lfs diff=lfs merge=lfs -text
38
+ qwen3.5-0.8B-abliterated-q4_0.gguf filter=lfs diff=lfs merge=lfs -text
39
+ qwen3.5-0.8B-abliterated-q4_k_m.gguf filter=lfs diff=lfs merge=lfs -text
40
+ qwen3.5-0.8B-abliterated-q5_0.gguf filter=lfs diff=lfs merge=lfs -text
41
+ qwen3.5-0.8B-abliterated-q5_k_m.gguf filter=lfs diff=lfs merge=lfs -text
42
+ qwen3.5-0.8B-abliterated-q6_k.gguf filter=lfs diff=lfs merge=lfs -text
43
+ qwen3.5-0.8B-abliterated-q8_0.gguf filter=lfs diff=lfs merge=lfs -text
README.md ADDED
@@ -0,0 +1,105 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ base_model: Qwen/Qwen3.5-0.8B
4
+ tags:
5
+ - qwen3.5
6
+ - qwen
7
+ - gguf
8
+ - abliterated
9
+ - llama.cpp
10
+ - quantized
11
+ pipeline_tag: text-generation
12
+ ---
13
+
14
+ # Qwen3.5-0.8B-Abliterated-GGUF-V2
15
+
16
+ GGUF quantizations of **Qwen3.5-0.8B-Abliterated-V2**.
17
+
18
+ The base model was abliterated using **[AnlordAbliterator 1.3.0](https://github.com/justbedwarsplay/AnlordAbliterator)** and then converted to GGUF and quantized into multiple formats.
19
+
20
+ ## Available Quantizations
21
+
22
+ | Quantization | File |
23
+ | ------------ | -------------------------------------- |
24
+ | BF16 | `qwen3.5-0.8B-abliterated-bf16.gguf` |
25
+ | F16 | `qwen3.5-0.8B-abliterated-f16.gguf` |
26
+ | Q8_0 | `qwen3.5-0.8B-abliterated-q8_0.gguf` |
27
+ | Q6_K | `qwen3.5-0.8B-abliterated-q6_k.gguf` |
28
+ | Q5_K_M | `qwen3.5-0.8B-abliterated-q5_k_m.gguf` |
29
+ | Q5_0 | `qwen3.5-0.8B-abliterated-q5_0.gguf` |
30
+ | Q4_K_M | `qwen3.5-0.8B-abliterated-q4_k_m.gguf` |
31
+ | Q4_0 | `qwen3.5-0.8B-abliterated-q4_0.gguf` |
32
+
33
+ ## Which Quantization Should I Use?
34
+
35
+ A simple rule of thumb:
36
+
37
+ | Quantization | Quality | Size | Recommended for |
38
+ | ------------ | ------- | ---------- | ---------------------- |
39
+ | BF16 | β˜…β˜…β˜…β˜…β˜… | Very large | Maximum precision |
40
+ | F16 | β˜…β˜…β˜…β˜…β˜… | Large | Maximum precision |
41
+ | Q8_0 | β˜…β˜…β˜…β˜…β˜… | Large | Near-original quality |
42
+ | Q6_K | β˜…β˜…β˜…β˜…β˜… | Medium | High quality |
43
+ | Q5_K_M | β˜…β˜…β˜…β˜…β˜† | Medium | Quality / size balance |
44
+ | Q5_0 | β˜…β˜…β˜…β˜…β˜† | Medium | General use |
45
+ | Q4_K_M | β˜…β˜…β˜…β˜…β˜† | Small | Recommended default |
46
+ | Q4_0 | β˜…β˜…β˜…β˜†β˜† | Smallest | Maximum memory savings |
47
+
48
+ **Q4_K_M** is the recommended starting point for most users who want a good balance between quality and memory usage.
49
+
50
+ ## Base Model
51
+
52
+ **Qwen/Qwen3.5-0.8B**
53
+
54
+ Original model:
55
+
56
+ https://huggingface.co/Qwen/Qwen3.5-0.8B
57
+
58
+ Abliterated Transformers version (V2):
59
+
60
+ https://huggingface.co/anlord/Qwen3.5-0.8B-Abliterated-V2
61
+
62
+ ## Abliteration
63
+
64
+ The base model was processed with **AnlordAbliterator 1.3.0**. This is the **V2** ablation of the model.
65
+
66
+ ### Results
67
+
68
+ ```text
69
+ Model: Qwen/Qwen3.5-0.8B
70
+
71
+ Initial refusals: 97 / 100
72
+ Final refusals: 2 / 100
73
+
74
+ KL divergence: 0.04527735710144043
75
+
76
+ Abliteration time: ~5050 seconds (200 optimization trials)
77
+ ```
78
+
79
+ ### Tool
80
+
81
+ [AnlordAbliterator](https://github.com/justbedwarsplay/AnlordAbliterator)
82
+
83
+ ## Running with llama.cpp
84
+
85
+ Example:
86
+
87
+ ```bash
88
+ llama-cli -m qwen3.5-0.8B-abliterated-q4_k_m.gguf
89
+ ```
90
+
91
+ The GGUF files are intended for use with GGUF-compatible software such as llama.cpp and other compatible inference applications.
92
+
93
+ ## License
94
+
95
+ This repository contains derivative model files based on **Qwen/Qwen3.5-0.8B**.
96
+
97
+ The original Qwen3.5-0.8B model is licensed under the **Apache License 2.0**.
98
+
99
+ See the included `LICENSE` file and the original model repository for the applicable license terms.
100
+
101
+ ## Disclaimer
102
+
103
+ These quantizations are derived from an abliterated version of Qwen3.5-0.8B.
104
+
105
+ Quantization may introduce small differences in model behavior and output quality compared with the original Safetensors model.
qwen3.5-0.8B-abliterated-bf16.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:eda24c765c19d3c6ae5f2d21c0fd33e338aaf31795a06579e05a84b5e63e54bf
3
+ size 1516744192
qwen3.5-0.8B-abliterated-f16.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:6f4e2dfa48bf05ba0ccf19a567c150fc71f21bfbd75a4f4853b779397d96d3dc
3
+ size 1516744192
qwen3.5-0.8B-abliterated-q4_0.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:9f2fcd473b7a7b787d164275dd0794a602d65f057bfdd1e17d77968f2a926d33
3
+ size 501452288
qwen3.5-0.8B-abliterated-q4_k_m.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:977d9299e899da18e6cda3416b4a7312870f2ca6e2c325815754731c513041b5
3
+ size 529296896
qwen3.5-0.8B-abliterated-q5_0.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:3afaaff53b16b174e0ea74b53bfdc44fa3cb3e3b005b9149382fd0dfd8de7bba
3
+ size 563654144
qwen3.5-0.8B-abliterated-q5_k_m.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:c69bcdd27d9c3888ff7aaae89329d6c6384a5bc187824938ab0434147ddf6dcb
3
+ size 577998336
qwen3.5-0.8B-abliterated-q6_k.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:3dc06b715661e21187cd6f6f01d220dab8b57800d45f0a757c4ddb248f82ae03
3
+ size 629743616
qwen3.5-0.8B-abliterated-q8_0.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:279f32e2cbbed344028529917da1f49c92a3713d89c988502fe1713cf949af79
3
+ size 811843072