Image-Text-to-Text
Transformers
GGUF
llama.cpp
vision
multimodal
text-generation-inference
unsloth
conversational
qwen3_5
reasoning
chain-of-thought
lora
sft
agent
tool-use
function-calling
coder
luxuansang Jackrong commited on
Commit
f16022e
Β·
0 Parent(s):

Duplicate from Jackrong/Qwopus3.5-9B-Coder-GGUF

Browse files

Co-authored-by: Jackrong <Jackrong@users.noreply.huggingface.co>

.gitattributes ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ *.7z filter=lfs diff=lfs merge=lfs -text
2
+ *.arrow filter=lfs diff=lfs merge=lfs -text
3
+ *.bin filter=lfs diff=lfs merge=lfs -text
4
+ *.bz2 filter=lfs diff=lfs merge=lfs -text
5
+ *.ckpt filter=lfs diff=lfs merge=lfs -text
6
+ *.ftz filter=lfs diff=lfs merge=lfs -text
7
+ *.gz filter=lfs diff=lfs merge=lfs -text
8
+ *.h5 filter=lfs diff=lfs merge=lfs -text
9
+ *.joblib filter=lfs diff=lfs merge=lfs -text
10
+ *.lfs.* filter=lfs diff=lfs merge=lfs -text
11
+ *.mlmodel filter=lfs diff=lfs merge=lfs -text
12
+ *.model filter=lfs diff=lfs merge=lfs -text
13
+ *.msgpack filter=lfs diff=lfs merge=lfs -text
14
+ *.npy filter=lfs diff=lfs merge=lfs -text
15
+ *.npz filter=lfs diff=lfs merge=lfs -text
16
+ *.onnx filter=lfs diff=lfs merge=lfs -text
17
+ *.ot filter=lfs diff=lfs merge=lfs -text
18
+ *.parquet filter=lfs diff=lfs merge=lfs -text
19
+ *.pb filter=lfs diff=lfs merge=lfs -text
20
+ *.pickle filter=lfs diff=lfs merge=lfs -text
21
+ *.pkl filter=lfs diff=lfs merge=lfs -text
22
+ *.pt filter=lfs diff=lfs merge=lfs -text
23
+ *.pth filter=lfs diff=lfs merge=lfs -text
24
+ *.rar filter=lfs diff=lfs merge=lfs -text
25
+ *.safetensors filter=lfs diff=lfs merge=lfs -text
26
+ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
27
+ *.tar.* filter=lfs diff=lfs merge=lfs -text
28
+ *.tar filter=lfs diff=lfs merge=lfs -text
29
+ *.tflite filter=lfs diff=lfs merge=lfs -text
30
+ *.tgz filter=lfs diff=lfs merge=lfs -text
31
+ *.wasm filter=lfs diff=lfs merge=lfs -text
32
+ *.xz filter=lfs diff=lfs merge=lfs -text
33
+ *.zip filter=lfs diff=lfs merge=lfs -text
34
+ *.zst filter=lfs diff=lfs merge=lfs -text
35
+ *tfevents* filter=lfs diff=lfs merge=lfs -text
36
+ Qwopus3.5-9B-coder-Exp-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
37
+ Qwopus3.5-9B-coder-Exp-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
38
+ Qwopus3.5-9B-coder-Exp-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
39
+ Qwopus3.5-9B-coder-Exp-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
40
+ Qwopus3.5-9B-coder-Exp-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
41
+ Qwopus3.5-9B-coder-Exp-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
42
+ Qwopus3.5-9B-coder-Exp-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
43
+ Qwopus3.5-9B-coder-Exp-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
44
+ Qwopus3.5-9B-coder-Exp-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
45
+ Qwopus3.5-9B-coder-Exp-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
46
+ Qwopus3.5-9B-coder-Exp-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
47
+ Qwopus3.5-9B-coder-Exp-BF16.gguf filter=lfs diff=lfs merge=lfs -text
48
+ mmproj.gguf filter=lfs diff=lfs merge=lfs -text
49
+ mmproj-F32.gguf filter=lfs diff=lfs merge=lfs -text
Qwopus3.5-9B-coder-Exp-Q3_K_M.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:d652b6a26842ead8ebe8f27b9b77a1c66e400e096615900d5efa849eea546862
3
+ size 4623526880
Qwopus3.5-9B-coder-Exp-Q4_K_M.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:bf7d30e6c973538c422038cb6fce23665f52f444b32c6f44c9b668812515fcb0
3
+ size 5629111264
Qwopus3.5-9B-coder-Exp-Q5_K_M.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:30b0da5b73d3ff860ac4355e7d6b28f402b279857783b66429eddccf8b726fff
3
+ size 6467972064
Qwopus3.5-9B-coder-Exp-Q6_K.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:dc7533a36c4de511287f58d34cbc10da8691293c648f7958f543e6310f128bf1
3
+ size 7359261664
Qwopus3.5-9B-coder-Exp-Q8_0.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:1fb4a8bd2b5f0680c0ddc866d1bb1e9b70c0c43132480bc14b0d1612210d8f51
3
+ size 9527503840
README.md ADDED
@@ -0,0 +1,471 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ library_name: transformers
3
+ base_model:
4
+ - Jackrong/Qwopus3.5-9B-v3.5
5
+ tags:
6
+ - gguf
7
+ - llama.cpp
8
+ - image-text-to-text
9
+ - vision
10
+ - multimodal
11
+ - text-generation-inference
12
+ - transformers
13
+ - unsloth
14
+ - conversational
15
+ - qwen3_5
16
+ - reasoning
17
+ - chain-of-thought
18
+ - lora
19
+ - sft
20
+ - agent
21
+ - tool-use
22
+ - function-calling
23
+ - coder
24
+ license: apache-2.0
25
+ language:
26
+ - en
27
+ - zh
28
+ - es
29
+ - ru
30
+ - ja
31
+ pipeline_tag: image-text-to-text
32
+ datasets:
33
+ - lambda/hermes-agent-reasoning-traces
34
+ - Jackrong/Claude-opus-4.7-TraceInversion-5000x
35
+ - Jackrong/Claude-opus-4.6-TraceInversion-9000x
36
+ ---
37
+ # 🌟 Qwopus3.5-9B-coder
38
+
39
+ ## πŸš€ Model Fine-Tuning and Logical Alignment (Qwopus3.5-9B-coder)
40
+
41
+ As the base model of this model, **Qwopus3.5-9B-v3.5** is already a model with powerful capabilities. On this foundation, **Qwopus3.5-9B-coder** is specially optimized and fine-tuned for high-performance **πŸ€– Agentic Coding, complex Tool Calling, and logical reasoning.**
42
+
43
+ > πŸ’‘ **Why the 9B Dense Model?**
44
+ > We believe that the 9B dense architecture represents the perfect **"sweet spot"** for large language models. It runs seamlessly at 8-bit precision on entry-level 16GB RAM devicesβ€”such as standard laptops and the Mac miniβ€”making it exceptionally lightweight yet highly versatile. Without requiring expensive hardware, it allows you to achieve excellent performance paired with impressive inference speeds. Simply put, **Qwen3.5-9B is currently the best open-source model in its class.**
45
+
46
+
47
+ ![image](https://cdn-uploads.huggingface.co/production/uploads/66309bd090589b7c65950665/8qFQVuCxbgkWqKa2B_Vph.jpeg)
48
+
49
+ > [!TIP]
50
+ >**Vision & Tool Calling Support**: This model supports visual capabilities and tool calling. To enable vision, please place the `mmproj.gguf` file from the [GGUF repository](https://huggingface.co/Jackrong/Qwopus3.5-9B-coder-GGUF) into the same directory as the main `.gguf` file.
51
+
52
+ ---
53
+
54
+ ### πŸ›  Training Strategy
55
+
56
+ The fine-tuning process of this model deeply integrates **Trace Inversion** data augmentation technology with high-quality **Agent Traces**. This systematic approach not only strengthens the model's ability to solve complex programming tasks, but also greatly improves its logical coherence and accuracy when using various tools.
57
+
58
+ This model is designed specifically for the following goals:
59
+
60
+ - 🧩 More structured and stronger logical reasoning capabilities, reducing repetitive thinking
61
+ - πŸ’» More powerful capabilities in code writing, debugging, and repository-level task processing
62
+ - πŸ›  More stable and accurate Tool Calling capabilities for terminal commands, file operations, and browsers
63
+ - πŸ” Better cross-data source distillation alignment
64
+
65
+
66
+ > [!WARNING]
67
+ > - **Community Release Notice**: Qwopus3.5-9B-coder is released purely as an experimental community version, aiming to explore the combination of Agent capabilities and deep reasoning, and is only for research and exploration use.
68
+ > - **Warning**: Because this model is vertically fine-tuned for programming agents and deep reasoning, and has not undergone comprehensive general performance evaluation, its capabilities in general domains or specific non-programming tasks may suffer from Capability Decay. Users are advised to be aware of its limitations in other scenarios while exploring its core capabilities.
69
+
70
+ ---
71
+
72
+
73
+ ## πŸ“Š Baseline Performance Comparison
74
+
75
+ To verify the execution efficiency and logical robustness of **Qwopus3.5-9B-coder** in actual agent scenarios, we adopted the open-source testing framework [benchlocal](https://github.com/stevibe/benchlocal).
76
+
77
+ ### Test Configuration
78
+ - **Hardware Environment**: Apple Silicon (Mac)
79
+ - **Inference Backend**: LM Studio / MLX / GGUF
80
+ - **Testing Platform**: [benchlocal](https://github.com/stevibe/benchlocal) - An evaluation suite focusing on local model agent capabilities.
81
+ - 🍎 You can see the actual inference speeds of different model formats on the same device.
82
+
83
+ ### πŸ§ͺ Benchmark Results
84
+
85
+ <div style="display: inline-block; padding: 6px 16px; background: #e0f2fe; color: #0369a1; border: 1px solid #bae6fd; border-radius: 8px; font-weight: 700; font-size: 16px; margin-bottom: 12px;">1. Complex Agent Performance - HermesAgent-20</div>
86
+ The following is the comparative performance under the HermesAgent-20 task set:
87
+
88
+ <table style="width: 100%; border-collapse: collapse; font-family: -apple-system, BlinkMacSystemFont, 'Segoe UI', Roboto, Helvetica, Arial, sans-serif;">
89
+ <thead>
90
+ <tr>
91
+ <td colspan="4" style="padding: 8px 12px; font-weight: 600; color: #7c3aed; border-bottom: 1px solid rgba(124, 58, 237, 0.2); background: rgba(124, 58, 237, 0.05);">HermesAgent-20 Performance Metrics</td>
92
+ </tr>
93
+ <tr style="background: rgba(128, 128, 128, 0.02);">
94
+ <th style="padding: 7px 7px; padding-left: 20px; text-align: left; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Model</th>
95
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Test Set</th>
96
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Comprehensive Score</th>
97
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Core Dimensions (M/O/S/S/B)</th>
98
+ </tr>
99
+ </thead>
100
+ <tbody>
101
+ <tr>
102
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><b><a href="https://huggingface.co/Jackrong/Qwopus3.5-9B-coder-GGUF" style="color: #7c3aed; text-decoration: none;">Qwopus3.5-9B-coder</a></b></td>
103
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">HermesAgent-20</td>
104
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); color: #7c3aed; font-weight: bold;">85</td>
105
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">84 / 93 / 88 / 75 / 84</td>
106
+ </tr>
107
+ <tr>
108
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/Qwen/Qwen3.5-9B" style="color: #666; text-decoration: none;">Qwen/Qwen3.5-9B</a></td>
109
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">HermesAgent-20</td>
110
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">71</td>
111
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">75 / 58 / 100 / 53 / 69</td>
112
+ </tr>
113
+ <tr>
114
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/armand0e/Qwen3.5-9B-Agent" style="color: #666; text-decoration: none;">armand0e/Qwen3.5-9B-Agent</a></td>
115
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">HermesAgent-20</td>
116
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">68</td>
117
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">71 / 83 / 43 / 61 / 80</td>
118
+ </tr>
119
+ <tr>
120
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/DJLougen/Harmonic-Hermes-9B" style="color: #666; text-decoration: none;">DJLougen/Harmonic-Hermes-9B</a></td>
121
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">HermesAgent-20</td>
122
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">47</td>
123
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">60 / 45 / 23 / 69 / 38</td>
124
+ </tr>
125
+ </tbody>
126
+ </table>
127
+
128
+ <div style="display: inline-block; padding: 6px 16px; background: #e0f2fe; color: #0369a1; border: 1px solid #bae6fd; border-radius: 8px; font-weight: 700; font-size: 16px; margin-bottom: 12px;">2. Tool Call Stability - ToolCall-15</div>
129
+ This is a ToolCall-15 test set targeting the stability of tool calls, aiming to test the stability of the model in tool calling:
130
+
131
+ <table style="width: 100%; border-collapse: collapse; font-family: -apple-system, BlinkMacSystemFont, 'Segoe UI', Roboto, Helvetica, Arial, sans-serif;">
132
+ <thead>
133
+ <tr>
134
+ <td colspan="4" style="padding: 8px 12px; font-weight: 600; color: #7c3aed; border-bottom: 1px solid rgba(124, 58, 237, 0.2); background: rgba(124, 58, 237, 0.05);">ToolCall-15 Stability Metrics</td>
135
+ </tr>
136
+ <tr style="background: rgba(128, 128, 128, 0.02);">
137
+ <th style="padding: 7px 7px; padding-left: 20px; text-align: left; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Model</th>
138
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Test Set</th>
139
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Comprehensive Score</th>
140
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Dimension Scores (A/B/C/D/E)</th>
141
+ </tr>
142
+ </thead>
143
+ <tbody>
144
+ <tr>
145
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><b><a href="https://huggingface.co/Jackrong/Qwopus3.5-9B-coder-GGUF" style="color: #7c3aed; text-decoration: none;">Qwopus3.5-9B-coder</a></b></td>
146
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">ToolCall-15</td>
147
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); color: #7c3aed; font-weight: bold;">100</td>
148
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">100 / 100 / 100 / 100 / 100</td>
149
+ </tr>
150
+ <tr>
151
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/Qwen/Qwen3.5-9B" style="color: #666; text-decoration: none;">Qwen/Qwen3.5-9B</a></td>
152
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">ToolCall-15</td>
153
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); color: #7c3aed; font-weight: bold;">100</td>
154
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">100 / 100 / 100 / 100 / 100</td>
155
+ </tr>
156
+ <tr>
157
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/armand0e/Qwen3.5-9B-Agent" style="color: #666; text-decoration: none;">armand0e/Qwen3.5-9B-Agent</a></td>
158
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">ToolCall-15</td>
159
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">93</td>
160
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">100 / 100 / 100 / 67 / 100</td>
161
+ </tr>
162
+ </tbody>
163
+ </table>
164
+
165
+ <div style="display: inline-block; padding: 6px 16px; background: #e0f2fe; color: #0369a1; border: 1px solid #bae6fd; border-radius: 8px; font-weight: 700; font-size: 16px; margin-bottom: 12px;">3. Code Debugging & Bug Fixing - BugFind-15</div>
166
+ BugFind-15 is a test set containing 15 scenarios from shallow to deep, aiming to evaluate the real debugging capabilities of the model in discovering and fixing syntax, logical errors, and "trap" code in multiple programming languages through deterministic environment runtime verification.
167
+
168
+ <table style="width: 100%; border-collapse: collapse; font-family: -apple-system, BlinkMacSystemFont, 'Segoe UI', Roboto, Helvetica, Arial, sans-serif;">
169
+ <thead>
170
+ <tr>
171
+ <td colspan="4" style="padding: 8px 12px; font-weight: 600; color: #7c3aed; border-bottom: 1px solid rgba(124, 58, 237, 0.2); background: rgba(124, 58, 237, 0.05);">BugFind-15 Performance Metrics</td>
172
+ </tr>
173
+ <tr style="background: rgba(128, 128, 128, 0.02);">
174
+ <th style="padding: 7px 7px; padding-left: 20px; text-align: left; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Model</th>
175
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Test Set</th>
176
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Comprehensive Score</th>
177
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Dimension Scores (A/B/C/D/E)</th>
178
+ </tr>
179
+ </thead>
180
+ <tbody>
181
+ <tr>
182
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><b><a href="https://huggingface.co/Jackrong/Qwopus3.5-9B-coder-GGUF" style="color: #7c3aed; text-decoration: none;">Qwopus3.5-9B-coder</a></b></td>
183
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">BugFind-15</td>
184
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); color: #7c3aed; font-weight: bold;">79</td>
185
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">67 / 87 / 100 / 77 / 43</td>
186
+ </tr>
187
+ <tr>
188
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/Jackrong/MLX-Qwen3.5-9B-DeepSeek-V4-Flash-8bit" style="color: #666; text-decoration: none;">Jackrong/MLX-Qwen3.5-9B-DeepSeek-V4-Flash</a></td>
189
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">BugFind-15</td>
190
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">75</td>
191
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">67 / 100 / 67 / 57 / 80</td>
192
+ </tr>
193
+ <tr>
194
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/armand0e/Qwen3.5-9B-Agent" style="color: #666; text-decoration: none;">armand0e/Qwen3.5-9B-Agent</a></td>
195
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">BugFind-15</td>
196
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">58</td>
197
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">29 / 87 / 73 / 20 / 67</td>
198
+ </tr>
199
+ </tbody>
200
+ </table>
201
+
202
+ ### πŸͺ SWE-bench Verified Performance (Repository-level Coding Capability)
203
+ The following shows the comparative performance on **SWE-bench Verified**, which evaluates language models on resolving software engineering issues in real-world open-source repositories:
204
+
205
+ <table style="width: 100%; border-collapse: collapse; font-family: -apple-system, BlinkMacSystemFont, 'Segoe UI', Roboto, Helvetica, Arial, sans-serif;">
206
+ <thead>
207
+ <tr>
208
+ <td colspan="3" style="padding: 8px 12px; font-weight: 600; color: #7c3aed; border-bottom: 1px solid rgba(124, 58, 237, 0.2); background: rgba(124, 58, 237, 0.05);">SWE-bench Verified Performance Metrics</td>
209
+ </tr>
210
+ <tr style="background: rgba(128, 128, 128, 0.02);">
211
+ <th style="padding: 7px 7px; padding-left: 20px; text-align: left; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Model</th>
212
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Test Set</th>
213
+ <th style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); font-size: 13px; color: #666;">Comprehensive Score (%)</th>
214
+ </tr>
215
+ </thead>
216
+ <tbody>
217
+ <tr>
218
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><span style="color: #666;">Claude 4.5 Opus</span></td>
219
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">SWE-bench Verified</td>
220
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">80.9</td>
221
+ </tr>
222
+ <tr>
223
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/Qwen/Qwen3.5-27B" style="color: #666; text-decoration: none;">Qwen/Qwen3.5-27B</a></td>
224
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">SWE-bench Verified</td>
225
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">75.0</td>
226
+ </tr>
227
+ <tr>
228
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/Qwen/Qwen3.6-35B-A3B" style="color: #666; text-decoration: none;">Qwen/Qwen3.6-35B-A3B</a></td>
229
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">SWE-bench Verified</td>
230
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">73.4</td>
231
+ </tr>
232
+ <tr>
233
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><b><a href="https://huggingface.co/Jackrong/Qwopus3.5-9B-coder-GGUF" style="color: #7c3aed; text-decoration: none;">Qwopus3.5-9B-coder</a></b></td>
234
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">SWE-bench Verified</td>
235
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15); color: #7c3aed; font-weight: bold;">53.89</td>
236
+ </tr>
237
+ <tr>
238
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><a href="https://huggingface.co/google/gemma-4-31B-it" style="color: #666; text-decoration: none;">google/gemma-4-31B-it</a></td>
239
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">SWE-bench Verified</td>
240
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">52.0</td>
241
+ </tr>
242
+ <tr>
243
+ <td style="padding: 7px 7px; padding-left: 20px; border-bottom: 1px solid rgba(128, 128, 128, 0.15);"><span style="color: #666;">google/gemma-4-26B-A4B</span></td>
244
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">SWE-bench Verified</td>
245
+ <td style="padding: 7px 7px; text-align: center; border-bottom: 1px solid rgba(128, 128, 128, 0.15);">45.0 - 48.0</td>
246
+ </tr>
247
+ </tbody>
248
+ </table>
249
+
250
+ ---
251
+
252
+ #### Community-Reported Tool Calling Evaluation
253
+
254
+ > [!TIP]
255
+ > A community user independently evaluated Qwopus3.5-9B-coder on a tool-recall test with up to 31 available tools and an adversarial tool-selection phase containing semantic overlaps and decoy tools.
256
+
257
+ | Model | Phase 1: Tool Recall | Phase 2: Adversarial Tool Selection |
258
+ |---|---:|---:|
259
+ | Qwopus3.5-9B-coder | 100% | 27 / 28 (96%) |
260
+ | Claude Opus 4.6 | 100% | 27 / 28 (96%) |
261
+ | Qwen3.5-9B | 100% | 26 / 28 (93%) |
262
+
263
+ ---
264
+
265
+ > [!IMPORTANT]
266
+ > - βš™οΈ All tests were conducted with a temperature of 1 as officially recommended by qwen3.5. All errors and model issues were attempted to be regenerated twice after a test failure. If both attempts fail, it is considered a failure.
267
+ > - 🍎 All screenshots of the test interfaces have been uploaded to the image folder in the repository. Click the link below to view and verify:
268
+ > - πŸ”— [View Test Screenshots](https://huggingface.co/Jackrong/Qwopus3.5-9B-coder/tree/main/test_images)
269
+ > - ❀️ **Kyle Hessling** for his generous hardware and equipment support. You can follow him for more updates on X / Twitter: [@KyleHessling1](https://x.com/KyleHessling1).
270
+
271
+
272
+ ---
273
+
274
+ ### πŸ§ͺ Core Dataset Usage: Trace Inversion and High-Quality Agent Traces
275
+
276
+ In order to break through the "reasoning bubble" limitation of the model in actual programming and tool usage, and to endow it with real Agent behavioral capabilities, this model introduced core augmented datasets during training:
277
+
278
+ #### 1. Reasoning Synthetic Data Combining Trace Inversion
279
+ **Currently, based on public information, commercial models such as OpenAI's GPT series and Anthropic's Claude series have very clearly hidden the true internal reasoning chains of their models. For these models, what we can ultimately see in the API or front-end interface can often only be considered a highly compressed "Reasoning Bubble".**
280
+
281
+ To break through this limitation, we adopted the **Trace Inversion** technology. This technology utilizes an external "surrogate model" to reconstruct a complete and logically coherent deep reasoning chain based on the "question + final answer + compressed reasoning summary" published by commercial models. The "reasoning bubble", which originally consisted of only a few sentences and logical leaps, is expanded into a high-quality deep learning trace with complete derivation, calculation, and logical verification, providing step-by-step logical learning signals for the model.
282
+
283
+
284
+ ![a_high_resolution_infographic_slide_style_figure](https://cdn-uploads.huggingface.co/production/uploads/66309bd090589b7c65950665/Jo2bm_rUJQmfK3Na4Uja2.png)
285
+
286
+
287
+
288
+
289
+ #### 2. GLM-5.1 Agent Real Trace Data: lambda/hermes-agent-reasoning-traces
290
+ To significantly enhance the model's execution and coding capabilities in real environments, this model additionally introduced the **`lambda/hermes-agent-reasoning-traces`** dataset.
291
+
292
+ ![Screenshot 2026-05-16 at 5.34.59β€―PM](https://cdn-uploads.huggingface.co/production/uploads/66309bd090589b7c65950665/BTusWFqYaOS5GmRYvBuPq.png)
293
+
294
+
295
+ - **Data Source and Scale**: This data subset contains approximately 10,000 high-quality multi-turn Tool Calling Trajectories generated based on the ZhipuAI GLM-5.1 and kimi-4.6 models.
296
+ - **Real Agent Behavior**: Unlike traditional synthetic data, these samples represent real Agent conversations. Each sample not only contains the step-by-step reasoning process in the `<think>` tags, but also includes actual tool execution results (rather than fabricated outputs out of thin air).
297
+ - **Extensive Domain Coverage**:
298
+ - **Terminal & Coding**: Script writing, code debugging, environment configuration, and data processing.
299
+ - **Repository Tasks**: Involving real code repository work, such as bug fixes, refactoring, and code review.
300
+ - **Browser Automation**: Web navigation, scraping, and form filling.
301
+ - **Agent Tools**: Memory persistence, task delegation, skill management, etc.
302
+
303
+ By learning these Agent trajectories that contain real feedback and thoughtful processes, Qwopus3.5-9B-coder can exhibit thinking and operational modes closer to human experts when facing complex programming and system operations tasks.
304
+
305
+ ---
306
+
307
+ ## πŸ—ΊοΈ Training Pipeline Overview
308
+
309
+ The training of this model integrates a phased learning pipeline of **Trace Inversion** data augmentation technology and **high-quality Agent Trajectories data**. Its core logic lies in restoring the highly compressed "reasoning bubble" of commercial models into a deep path for learning, and combining it with real agent operational traces to comprehensively improve the model's logical reasoning and code execution capabilities.
310
+
311
+ ```text
312
+ [ πŸ—ΊοΈ Trace Inversion: Full Process of Data Inversion and "Attack" Distillation ]
313
+
314
+ A. Surrogate Model Training
315
+ Open Source Model (GLM-5.1 / DS-V4) ──► Complete Reasoning Chain ──► [ Qwen3-235B Compression ] ──► Reasoning Bubbles
316
+ β”‚ β”‚
317
+ └──────────► [ Training ] β—„β”€β”€β”€β”€β”€β”€β”€β”€β”€β”˜
318
+ (Base: Qwen3-4B-Instruct)
319
+ (Result: Trace-Inverter-4B)
320
+
321
+ B. Inversion Phase: "Attacking" Claude-4.7-Max
322
+ _______________________________________________________
323
+ | |
324
+ | Claude-4.7-Max API ──► Compressed Bubbles + Final Answer |
325
+ |_______________________________________________________|
326
+ β”‚
327
+ β–Ό
328
+ [ 🧠 Trace-Inverter-4B (Logical Reconstructor) ] ────► Synthetic CoT
329
+ β”‚
330
+ β–Ό
331
+ [ 🧩 Data Splicing ] ◄────────── (Original Prompt + Response)
332
+ (Embed the inverted chain of thought into <think> tags, and splice with the original Q&A pair for restoration)
333
+ β”‚
334
+ β–Ό
335
+ (Result: claude-opus-4.6/4.7 Inversion Set)
336
+
337
+ C. Final SFT Pipeline
338
+ ___________________________________________
339
+ | |
340
+ | Base Model (Qwopus3.5-9B-v3.5) |
341
+ |___________________________________________|
342
+ β”‚
343
+ β–Ό
344
+ [ πŸ“¦ Stage 1: Format Establishment and Logic Injection ] ───────► [ πŸ› οΈ Stage 2: Agent Trajectories and Programming Reinforcement ]
345
+ (Integrate inverted reasoning data, stabilize thinking format) (Introduce GLM-5.1 Agent Trajectories, reinforce interaction and execution)
346
+ β”‚ β”‚
347
+ β”‚ β–Ό
348
+ β”‚ __________________________________________________
349
+ β”‚ | πŸ” Hermes Agent Trace Sample Structure Breakdown (GLM-5.1) |
350
+ β”‚ | 1. [πŸ› οΈ System] -> JSON Tool Definition |
351
+ β”‚ | 2. [πŸ‘€ Human] -> Initial Task Instruction |
352
+ β”‚ | β”Œβ”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β” |
353
+ β”‚ | β”‚ πŸ” Multi-turn Loop: β”‚ |
354
+ β”‚ | β”‚ 3. [🧠 GPT] -> <think> Logical Reasoning/Reflection β”‚ |
355
+ β”‚ | β”‚ 4. [πŸ€– GPT] -> Tool Call Execution Action β”‚ |
356
+ β”‚ | β”‚ 5. [βš™οΈ Tool] -> Real Feedback β”‚ |
357
+ β”‚ | β””β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”˜ |
358
+ β”‚ |__________________________________________________|
359
+ β”‚ β”‚
360
+ β””β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”¬β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”˜
361
+ β–Ό
362
+ ___________________________________
363
+ | |
364
+ | 🌟 Final Model: Qwopus3.5-9B-coder |
365
+ |___________________________________|
366
+ ```
367
+
368
+
369
+ > [!NOTE]
370
+ > Because agent trajectory datasets are complex and diverse. The datasets have undergone rigorous cleaning and formatting.
371
+
372
+ ## 🎯 Three-Stage Curriculum Learning
373
+
374
+ **Qwopus3.5-9B-coder** adopts a phased reasoning data mixture strategy similar to Curriculum Learning, gradually increasing the difficulty and complexity of training signals:
375
+
376
+ 1. **Early Stage (Format Establishment):** Focuses on short-to-medium length reasoning samples with stable formats. The primary goal of this stage is to establish a reliable, structured new reasoning format while avoiding overwhelming the model with extreme complexity.
377
+
378
+ 2. **Middle Stage (Complexity Scaling & Multi-Teacher Distillation):** Gradually increases the proportion of complex reasoning samples from multiple teacher models.
379
+ - The distillation data is sourced from more powerful models whose style distribution closely matches the base model, ensuring that the capability gap is not too wide, thereby achieving efficient learning.
380
+
381
+ 3. **Late Stage (Long-Context Reinforcement & Drift Prevention):** Reinforces reasoning capabilities in long contexts. Crucially, this stage retains **short-sample replay** to ensure the model maintains its short-context instruction-following capability and minimizes capability drift.
382
+
383
+ ---
384
+
385
+ ## πŸš€ Context Length and Long-Context Usage
386
+
387
+ During fine-tuning, this model was trained with a maximum sequence length of **32K tokens**. The training data mixture was also constructed around samples up to **32K tokens**, so the "Context Length Distribution" shown in this model card reflects the fine-tuning data distribution rather than a hard architectural limit.
388
+
389
+ The model still inherits the native long-context capability of the Qwen3.6 base model. Therefore, longer context windows such as **128K** or **256K** may be available in compatible inference runtimes, depending on the backend and configuration.
390
+
391
+ For practical long-context inference beyond 32K, especially when using **llama.cpp / GGUF**, it is recommended to enable **RoPE/YaRN scaling** instead of only increasing `n_ctx` / `--ctx-size`. Directly setting a larger context window without RoPE scaling may work in some cases, but it can be less stable and may not achieve the expected long-context performance.
392
+
393
+ This is consistent with Qwen community guidance for long-context GGUF usage: **128K context generally requires YaRN/RoPE scaling**, and it is not necessarily enabled by default in llama.cpp. For example, Qwen maintainers have noted that "128K context length needs YaRN" and that it should be explicitly enabled when supported by the runtime.
394
+ Reference: https://huggingface.co/Qwen/Qwen2.5-72B-Instruct-GGUF/discussions/2
395
+
396
+ Community feedback also suggests that RoPE/YaRN scaling can improve long-context stability for this model family. One user reported that, on **HermesAgent-20**, `Qwopus3.6-35B-A3B-v1` performed better when extending from **32K to 128K via RoPE scaling** than when directly setting a **128K context window** without scaling, with scores of **83 vs. 72** in their setup. This result may vary depending on the backend, quantization type, KV cache settings, hardware, and benchmark configuration, but it is consistent with the recommendation to use RoPE/YaRN scaling for contexts beyond 32K.
397
+
398
+ Example llama.cpp configuration for extending from 32K to 128K:
399
+
400
+ ```bash
401
+ ./llama-server \
402
+ -m model.gguf \
403
+ --ctx-size 131072 \
404
+ --rope-scaling yarn \
405
+ --rope-scale 4 \
406
+ --yarn-orig-ctx 32768
407
+ ```
408
+
409
+ For 256K context, users may need to adjust the scaling factor accordingly and validate the result in their own workload:
410
+
411
+ ```bash
412
+ ./llama-server \
413
+ -m model.gguf \
414
+ --ctx-size 262144 \
415
+ --rope-scaling yarn \
416
+ --rope-scale 8 \
417
+ --yarn-orig-ctx 32768
418
+ ```
419
+
420
+ Please note that long-context behavior may vary depending on the inference backend, quantization type, KV cache settings, available memory, and task type. For best results, users should benchmark their own target workload when using contexts beyond 32K.
421
+
422
+ ---
423
+
424
+ ## 🀝 Collaboration & Training Details
425
+
426
+ This model is the result of continuous exploration in Agentic AI and reasoning capabilities.
427
+
428
+ **Training Infrastructure & Configuration:**
429
+ - πŸ–₯️ **Hardware:** Local compute devices / Cloud GPUs (e.g. GB10 / H100 / RTX 5090 / A100)
430
+ - βš™οΈ **Framework:** Unsloth for efficient fine-tuning
431
+
432
+ ---
433
+
434
+ ## ⚠️ IMPORTANT
435
+
436
+ > [!CAUTION]
437
+ > **Compatibility and Deployment Notice**
438
+ > - **Tool Calling Format**: When using this model for tool calling, please ensure that you use a Prompt format and System Prompt that match the training data to activate its Agent capabilities.
439
+ > - **Reasoning Output Extraction**: The model's thinking process is typically wrapped in `<think>` and `</think>` tags. When deploying to front-end applications, these tags may need to be parsed and hidden.
440
+
441
+ ---
442
+
443
+ ## πŸ“š Resources & Guides
444
+
445
+ πŸ‘‰ **[GitHub Repository: Jackrong-llm-finetuning-guide](https://github.com/R6410418/Jackrong-llm-finetuning-guide.git)**
446
+ Visit the repository to dive into our fine-tuning codebase and guides.
447
+
448
+ ---
449
+
450
+ ## πŸ™ Acknowledgements
451
+
452
+ Special thanks to:
453
+ - The Qwen team for the strong Qwen3.6 MoE base model.
454
+ - Unsloth for efficient fine-tuning frameworks.
455
+ - Open-source datasets and community contributors.
456
+ - **Kyle Hessling** for his generous hardware and equipment support. You can follow him for more updates on X / Twitter: [@KyleHessling1](https://x.com/KyleHessling1).
457
+
458
+
459
+
460
+ ---
461
+
462
+ ## πŸ“– Citation
463
+
464
+ ```bibtex
465
+ @misc{jackrong_qwopus35_9b_coder,
466
+ title = {Qwopus3.5-9B-coder},
467
+ author = {Jackrong},
468
+ year = {2026},
469
+ publisher = {Hugging Face}
470
+ }
471
+ ```
mmproj-F32.gguf ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:5c769161b31697b6a2d83d8a806f37ee8ee7104bca15313c608dc53359fa0ef2
3
+ size 921704480