HauhauCS commited on
Commit
72cb051
·
verified ·
1 Parent(s): 80c389b

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +72 -0
README.md ADDED
@@ -0,0 +1,72 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ tags:
4
+ - uncensored
5
+ - qwen3
6
+ language:
7
+ - en
8
+ - zh
9
+ base_model: Qwen/Qwen3-4B-Instruct-2507
10
+ ---
11
+
12
+ # Qwen3-4B-2507-Instruct-Uncensored-HauhauCS-Aggressive
13
+
14
+ Qwen3 4B 2507 Instruct uncensored by HauhauCS.
15
+
16
+ ## About
17
+
18
+ No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended - just without the refusals.
19
+
20
+ These are meant to be the best lossless uncensored models out there.
21
+
22
+ ## Aggressive vs Balanced
23
+
24
+ **Aggressive** applies stronger uncensoring. Use this when you need fewer refusals and don't mind occasional quirks.
25
+
26
+ For agentic coding or reliability-critical tasks, use the Balanced variant instead (coming soon).
27
+
28
+ ## Downloads
29
+
30
+ | File | Quant | Size |
31
+ |------|-------|------|
32
+ | Qwen3-4B-2507-Instruct-Uncensored-HauhauCS-Aggressive-FP16.gguf | FP16 | 7.5 GB |
33
+ | Qwen3-4B-2507-Instruct-Uncensored-HauhauCS-Aggressive-Q8_0.gguf | Q8_0 | 4.0 GB |
34
+ | Qwen3-4B-2507-Instruct-Uncensored-HauhauCS-Aggressive-Q6_K.gguf | Q6_K | 3.1 GB |
35
+ | Qwen3-4B-2507-Instruct-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf | Q4_K_M | 2.4 GB |
36
+
37
+ ## Specs
38
+
39
+ - 4B parameters (dense)
40
+ - 262K context
41
+ - Based on [Qwen/Qwen3-4B-Instruct-2507](https://huggingface.co/Qwen/Qwen3-4B-Instruct-2507)
42
+
43
+ ## Recommended Settings
44
+
45
+ From the Qwen team:
46
+
47
+ **Thinking mode (default):**
48
+ - `temperature=0.6`
49
+ - `top_p=0.95`
50
+ - `top_k=20`
51
+ - `min_p=0`
52
+
53
+ **Non-thinking mode:**
54
+ - Add `/no_think` at the end of your prompt, or
55
+ - `temperature=0.7`
56
+ - `top_p=0.8`
57
+ - `top_k=20`
58
+ - `min_p=0`
59
+
60
+ **Important:**
61
+ - Use `--jinja` flag for proper chat template handling
62
+ - Thinking mode produces `<think>...</think>` tags before responses
63
+
64
+ ## Usage
65
+
66
+ Works with llama.cpp, LM Studio, Jan, koboldcpp, Ollama, etc.
67
+
68
+ ```bash
69
+ # llama.cpp example
70
+ ./llama-cli -m Qwen3-4B-2507-Instruct-Uncensored-HauhauCS-Aggressive-Q4_K_M.gguf \
71
+ -p "Hello" --jinja -c 8192
72
+ ```