Atomic-Germ commited on
Commit
d64ebda
·
verified ·
1 Parent(s): b9c5ca5

Upload q4nx-build.log with huggingface_hub

Browse files
Files changed (1) hide show
  1. q4nx-build.log +443 -0
q4nx-build.log ADDED
@@ -0,0 +1,443 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ [INFO] Base chain: empero-ai/Qwen3.8-2B-Distill-GGUF -> empero-ai/Qwen3.8-2B -> Qwen/Qwen3.5-2B -> Qwen/Qwen3.5-2B-Base
2
+ [INFO] Skeleton source: Atomic-Germ/Qwen3.5-2B-NPU2
3
+ [INFO] Weights type: language
4
+ [INFO] Found GGUF in empero-ai/Qwen3.8-2B-Distill-GGUF: Qwen3.8-2B-Q8_0.gguf
5
+ [INFO] Converting /home/atomic-germ/.cache/huggingface/hub/models--empero-ai--Qwen3.8-2B-Distill-GGUF/snapshots/f4f73582d0b149595450c719b9a7521a03894f9c/Qwen3.8-2B-Q8_0.gguf to /home/atomic-germ/Projects/q4nx-build-v0.2.0/Qwen3.8-2B-Distill-NPU2...
6
+ [INFO] Using Qwen35 converter (variant: ['qwen35-2B', 'qwen3.5-2B'])
7
+ [INFO] Loading Q4NX config from /home/atomic-germ/Projects/q4nx-build-v0.2.0/configs/qwen3.5_2b.json
8
+ [INFO] Creating name maps...
9
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.ssm_beta.weight'
10
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.attn_q_norm.weight'
11
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.attn_q.weight'
12
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.ssm_dt.bias'
13
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.ffn_down.weight'
14
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.attn_norm.weight'
15
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.ffn_up.weight'
16
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.attn_k_norm.weight'
17
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.post_attention_norm.weight'
18
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.attn_output.weight'
19
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.ssm_a'
20
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.ssm_norm.weight'
21
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.attn_v.weight'
22
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.attn_gate.weight'
23
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.attn_qkv.weight'
24
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.ffn_gate.weight'
25
+ [INFO] Detected 25 layers for pattern 'blk.{bid}.attn_k.weight'
26
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.ssm_out.weight'
27
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.ssm_conv1d.weight'
28
+ [INFO] Detected 23 layers for pattern 'blk.{bid}.ssm_alpha.weight'
29
+ Converted token_embd.weight to model.embed_tokens.weight
30
+ Converted blk.0.attn_q.weight to model.layers.0.self_attn.q_proj.weight
31
+ Converted blk.0.attn_q_norm.weight to model.layers.0.self_attn.q_norm.weight
32
+ Converted blk.0.attn_k.weight to model.layers.0.self_attn.k_proj.weight
33
+ Converted blk.0.attn_k_norm.weight to model.layers.0.self_attn.k_norm.weight
34
+ Converted blk.0.attn_v.weight to model.layers.0.self_attn.v_proj.weight
35
+ Converted blk.0.attn_output.weight to model.layers.0.self_attn.o_proj.weight
36
+ Converted blk.0.ffn_up.weight to model.layers.0.mlp.up_proj.weight
37
+ Converted blk.0.ffn_gate.weight to model.layers.0.mlp.gate_proj.weight
38
+ Converted blk.0.ffn_down.weight to model.layers.0.mlp.down_proj.weight
39
+ Converted blk.0.attn_norm.weight to model.layers.0.input_layernorm.weight
40
+ Converted blk.0.post_attention_norm.weight to model.layers.0.post_attention_layernorm.weight
41
+ Converted blk.0.attn_gate.weight to model.layers.0.self_attn.gate_proj.weight
42
+ Converted blk.0.attn_qkv.weight to model.layers.0.linear_attn.qkv_proj.weight
43
+ Converted blk.0.ssm_a to model.layers.0.linear_attn.ssm_a
44
+ Converted blk.0.ssm_alpha.weight to model.layers.0.linear_attn.ssm_alpha_proj.weight
45
+ Converted blk.0.ssm_beta.weight to model.layers.0.linear_attn.ssm_beta_proj.weight
46
+ Converted blk.0.ssm_conv1d.weight to model.layers.0.linear_attn.ssm_conv1d.weight
47
+ Converted blk.0.ssm_norm.weight to model.layers.0.linear_attn.ssm_norm.weight
48
+ Converted blk.0.ssm_out.weight to model.layers.0.linear_attn.ssm_out_proj.weight
49
+ Converted blk.0.ssm_dt.bias to model.layers.0.linear_attn.ssm_dt.bias
50
+ Converted output_norm.weight to model.norm.weight
51
+ Converted output.weight to lm_head.weight
52
+ Converted v.position_embd.weight to model.visual.pos_embed.weight
53
+ Converted v.patch_embd.weight to model.visual.patch_embed.proj.weight
54
+ Converted v.patch_embd.weight.1 to model.visual.patch_embed.proj.weight.1
55
+ Converted v.patch_embd.bias to model.visual.patch_embed.proj.bias
56
+ Converted v.post_ln.weight to model.visual.merger.norm.weight
57
+ Converted v.post_ln.bias to model.visual.merger.norm.bias
58
+ Converted mm.0.weight to model.visual.merger.linear_fc1.weight
59
+ Converted mm.0.bias to model.visual.merger.linear_fc1.bias
60
+ Converted mm.2.weight to model.visual.merger.linear_fc2.weight
61
+ Converted mm.2.bias to model.visual.merger.linear_fc2.bias
62
+ [INFO] Model does not have a lm_head, use embedding weights as lm_head
63
+ Processing tensor: output_norm.weight with type F32 -> model.norm.weight with dtype F32
64
+ Processing tensor: token_embd.weight with type Q8_0 -> model.embed_tokens.weight with dtype Q4_1
65
+ Processing tensor: blk.0.attn_gate.weight with type Q8_0 -> model.layers.0.self_attn.gate_proj.weight with dtype Q4_1
66
+ Processing tensor: blk.0.attn_norm.weight with type F32 -> model.layers.0.input_layernorm.weight with dtype F32
67
+ Processing tensor: blk.0.attn_qkv.weight with type Q8_0 -> model.layers.0.linear_attn.qkv_proj.weight with dtype Q4_1
68
+ Processing tensor: blk.0.ffn_down.weight with type Q8_0 -> model.layers.0.mlp.down_proj.weight with dtype Q4_1
69
+ Processing tensor: blk.0.ffn_gate.weight with type Q8_0 -> model.layers.0.mlp.gate_proj.weight with dtype Q4_1
70
+ Processing tensor: blk.0.ffn_up.weight with type Q8_0 -> model.layers.0.mlp.up_proj.weight with dtype Q4_1
71
+ Processing tensor: blk.0.post_attention_norm.weight with type F32 -> model.layers.0.post_attention_layernorm.weight with dtype F32
72
+ Processing tensor: blk.0.ssm_a with type F32 -> model.layers.0.linear_attn.ssm_a with dtype F32
73
+ Processing tensor: blk.0.ssm_alpha.weight with type Q8_0 -> model.layers.0.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
74
+ Processing tensor: blk.0.ssm_beta.weight with type Q8_0 -> model.layers.0.linear_attn.ssm_beta_proj.weight with dtype Q8_0
75
+ Processing tensor: blk.0.ssm_conv1d.weight with type F32 -> model.layers.0.linear_attn.ssm_conv1d.weight with dtype F32
76
+ [INFO] transpose conv1d
77
+ Processing tensor: blk.0.ssm_dt.bias with type F32 -> model.layers.0.linear_attn.ssm_dt.bias with dtype F32
78
+ Processing tensor: blk.0.ssm_norm.weight with type F32 -> model.layers.0.linear_attn.ssm_norm.weight with dtype F32
79
+ Processing tensor: blk.0.ssm_out.weight with type Q8_0 -> model.layers.0.linear_attn.ssm_out_proj.weight with dtype Q8_0
80
+ Processing tensor: blk.1.attn_gate.weight with type Q8_0 -> model.layers.1.self_attn.gate_proj.weight with dtype Q4_1
81
+ Processing tensor: blk.1.attn_norm.weight with type F32 -> model.layers.1.input_layernorm.weight with dtype F32
82
+ Processing tensor: blk.1.attn_qkv.weight with type Q8_0 -> model.layers.1.linear_attn.qkv_proj.weight with dtype Q4_1
83
+ Processing tensor: blk.1.ffn_down.weight with type Q8_0 -> model.layers.1.mlp.down_proj.weight with dtype Q4_1
84
+ Processing tensor: blk.1.ffn_gate.weight with type Q8_0 -> model.layers.1.mlp.gate_proj.weight with dtype Q4_1
85
+ Processing tensor: blk.1.ffn_up.weight with type Q8_0 -> model.layers.1.mlp.up_proj.weight with dtype Q4_1
86
+ Processing tensor: blk.1.post_attention_norm.weight with type F32 -> model.layers.1.post_attention_layernorm.weight with dtype F32
87
+ Processing tensor: blk.1.ssm_a with type F32 -> model.layers.1.linear_attn.ssm_a with dtype F32
88
+ Processing tensor: blk.1.ssm_alpha.weight with type Q8_0 -> model.layers.1.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
89
+ Processing tensor: blk.1.ssm_beta.weight with type Q8_0 -> model.layers.1.linear_attn.ssm_beta_proj.weight with dtype Q8_0
90
+ Processing tensor: blk.1.ssm_conv1d.weight with type F32 -> model.layers.1.linear_attn.ssm_conv1d.weight with dtype F32
91
+ [INFO] transpose conv1d
92
+ Processing tensor: blk.1.ssm_dt.bias with type F32 -> model.layers.1.linear_attn.ssm_dt.bias with dtype F32
93
+ Processing tensor: blk.1.ssm_norm.weight with type F32 -> model.layers.1.linear_attn.ssm_norm.weight with dtype F32
94
+ Processing tensor: blk.1.ssm_out.weight with type Q8_0 -> model.layers.1.linear_attn.ssm_out_proj.weight with dtype Q8_0
95
+ Processing tensor: blk.2.attn_gate.weight with type Q8_0 -> model.layers.2.self_attn.gate_proj.weight with dtype Q4_1
96
+ Processing tensor: blk.2.attn_norm.weight with type F32 -> model.layers.2.input_layernorm.weight with dtype F32
97
+ Processing tensor: blk.2.attn_qkv.weight with type Q8_0 -> model.layers.2.linear_attn.qkv_proj.weight with dtype Q4_1
98
+ Processing tensor: blk.2.ffn_down.weight with type Q8_0 -> model.layers.2.mlp.down_proj.weight with dtype Q4_1
99
+ Processing tensor: blk.2.ffn_gate.weight with type Q8_0 -> model.layers.2.mlp.gate_proj.weight with dtype Q4_1
100
+ Processing tensor: blk.2.ffn_up.weight with type Q8_0 -> model.layers.2.mlp.up_proj.weight with dtype Q4_1
101
+ Processing tensor: blk.2.post_attention_norm.weight with type F32 -> model.layers.2.post_attention_layernorm.weight with dtype F32
102
+ Processing tensor: blk.2.ssm_a with type F32 -> model.layers.2.linear_attn.ssm_a with dtype F32
103
+ Processing tensor: blk.2.ssm_alpha.weight with type Q8_0 -> model.layers.2.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
104
+ Processing tensor: blk.2.ssm_beta.weight with type Q8_0 -> model.layers.2.linear_attn.ssm_beta_proj.weight with dtype Q8_0
105
+ Processing tensor: blk.2.ssm_conv1d.weight with type F32 -> model.layers.2.linear_attn.ssm_conv1d.weight with dtype F32
106
+ [INFO] transpose conv1d
107
+ Processing tensor: blk.2.ssm_dt.bias with type F32 -> model.layers.2.linear_attn.ssm_dt.bias with dtype F32
108
+ Processing tensor: blk.2.ssm_norm.weight with type F32 -> model.layers.2.linear_attn.ssm_norm.weight with dtype F32
109
+ Processing tensor: blk.2.ssm_out.weight with type Q8_0 -> model.layers.2.linear_attn.ssm_out_proj.weight with dtype Q8_0
110
+ Processing tensor: blk.3.attn_k.weight with type Q8_0 -> model.layers.3.self_attn.k_proj.weight with dtype Q4_1
111
+ Processing tensor: blk.3.attn_k_norm.weight with type F32 -> model.layers.3.self_attn.k_norm.weight with dtype F32
112
+ Processing tensor: blk.3.attn_norm.weight with type F32 -> model.layers.3.input_layernorm.weight with dtype F32
113
+ Processing tensor: blk.3.attn_output.weight with type Q8_0 -> model.layers.3.self_attn.o_proj.weight with dtype Q4_1
114
+ Processing tensor: blk.3.attn_q.weight with type Q8_0 -> model.layers.3.self_attn.q_proj.weight with dtype Q4_1
115
+ [INFO] Seperate q, gate for q_proj
116
+ Processing tensor: blk.3.attn_q_norm.weight with type F32 -> model.layers.3.self_attn.q_norm.weight with dtype F32
117
+ Processing tensor: blk.3.attn_v.weight with type Q8_0 -> model.layers.3.self_attn.v_proj.weight with dtype Q4_1
118
+ Processing tensor: blk.3.ffn_down.weight with type Q8_0 -> model.layers.3.mlp.down_proj.weight with dtype Q4_1
119
+ Processing tensor: blk.3.ffn_gate.weight with type Q8_0 -> model.layers.3.mlp.gate_proj.weight with dtype Q4_1
120
+ Processing tensor: blk.3.ffn_up.weight with type Q8_0 -> model.layers.3.mlp.up_proj.weight with dtype Q4_1
121
+ Processing tensor: blk.3.post_attention_norm.weight with type F32 -> model.layers.3.post_attention_layernorm.weight with dtype F32
122
+ Processing tensor: blk.4.attn_gate.weight with type Q8_0 -> model.layers.4.self_attn.gate_proj.weight with dtype Q4_1
123
+ Processing tensor: blk.4.attn_norm.weight with type F32 -> model.layers.4.input_layernorm.weight with dtype F32
124
+ Processing tensor: blk.4.attn_qkv.weight with type Q8_0 -> model.layers.4.linear_attn.qkv_proj.weight with dtype Q4_1
125
+ Processing tensor: blk.4.ffn_down.weight with type Q8_0 -> model.layers.4.mlp.down_proj.weight with dtype Q4_1
126
+ Processing tensor: blk.4.ffn_gate.weight with type Q8_0 -> model.layers.4.mlp.gate_proj.weight with dtype Q4_1
127
+ Processing tensor: blk.4.ffn_up.weight with type Q8_0 -> model.layers.4.mlp.up_proj.weight with dtype Q4_1
128
+ Processing tensor: blk.4.post_attention_norm.weight with type F32 -> model.layers.4.post_attention_layernorm.weight with dtype F32
129
+ Processing tensor: blk.4.ssm_a with type F32 -> model.layers.4.linear_attn.ssm_a with dtype F32
130
+ Processing tensor: blk.4.ssm_alpha.weight with type Q8_0 -> model.layers.4.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
131
+ Processing tensor: blk.4.ssm_beta.weight with type Q8_0 -> model.layers.4.linear_attn.ssm_beta_proj.weight with dtype Q8_0
132
+ Processing tensor: blk.4.ssm_conv1d.weight with type F32 -> model.layers.4.linear_attn.ssm_conv1d.weight with dtype F32
133
+ [INFO] transpose conv1d
134
+ Processing tensor: blk.4.ssm_dt.bias with type F32 -> model.layers.4.linear_attn.ssm_dt.bias with dtype F32
135
+ Processing tensor: blk.4.ssm_norm.weight with type F32 -> model.layers.4.linear_attn.ssm_norm.weight with dtype F32
136
+ Processing tensor: blk.4.ssm_out.weight with type Q8_0 -> model.layers.4.linear_attn.ssm_out_proj.weight with dtype Q8_0
137
+ Processing tensor: blk.5.attn_gate.weight with type Q8_0 -> model.layers.5.self_attn.gate_proj.weight with dtype Q4_1
138
+ Processing tensor: blk.5.attn_norm.weight with type F32 -> model.layers.5.input_layernorm.weight with dtype F32
139
+ Processing tensor: blk.5.attn_qkv.weight with type Q8_0 -> model.layers.5.linear_attn.qkv_proj.weight with dtype Q4_1
140
+ Processing tensor: blk.5.ffn_down.weight with type Q8_0 -> model.layers.5.mlp.down_proj.weight with dtype Q4_1
141
+ Processing tensor: blk.5.ffn_gate.weight with type Q8_0 -> model.layers.5.mlp.gate_proj.weight with dtype Q4_1
142
+ Processing tensor: blk.5.ffn_up.weight with type Q8_0 -> model.layers.5.mlp.up_proj.weight with dtype Q4_1
143
+ Processing tensor: blk.5.post_attention_norm.weight with type F32 -> model.layers.5.post_attention_layernorm.weight with dtype F32
144
+ Processing tensor: blk.5.ssm_a with type F32 -> model.layers.5.linear_attn.ssm_a with dtype F32
145
+ Processing tensor: blk.5.ssm_alpha.weight with type Q8_0 -> model.layers.5.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
146
+ Processing tensor: blk.5.ssm_beta.weight with type Q8_0 -> model.layers.5.linear_attn.ssm_beta_proj.weight with dtype Q8_0
147
+ Processing tensor: blk.5.ssm_conv1d.weight with type F32 -> model.layers.5.linear_attn.ssm_conv1d.weight with dtype F32
148
+ [INFO] transpose conv1d
149
+ Processing tensor: blk.5.ssm_dt.bias with type F32 -> model.layers.5.linear_attn.ssm_dt.bias with dtype F32
150
+ Processing tensor: blk.5.ssm_norm.weight with type F32 -> model.layers.5.linear_attn.ssm_norm.weight with dtype F32
151
+ Processing tensor: blk.5.ssm_out.weight with type Q8_0 -> model.layers.5.linear_attn.ssm_out_proj.weight with dtype Q8_0
152
+ Processing tensor: blk.6.attn_gate.weight with type Q8_0 -> model.layers.6.self_attn.gate_proj.weight with dtype Q4_1
153
+ Processing tensor: blk.6.attn_norm.weight with type F32 -> model.layers.6.input_layernorm.weight with dtype F32
154
+ Processing tensor: blk.6.attn_qkv.weight with type Q8_0 -> model.layers.6.linear_attn.qkv_proj.weight with dtype Q4_1
155
+ Processing tensor: blk.6.ffn_down.weight with type Q8_0 -> model.layers.6.mlp.down_proj.weight with dtype Q4_1
156
+ Processing tensor: blk.6.ffn_gate.weight with type Q8_0 -> model.layers.6.mlp.gate_proj.weight with dtype Q4_1
157
+ Processing tensor: blk.6.ffn_up.weight with type Q8_0 -> model.layers.6.mlp.up_proj.weight with dtype Q4_1
158
+ Processing tensor: blk.6.post_attention_norm.weight with type F32 -> model.layers.6.post_attention_layernorm.weight with dtype F32
159
+ Processing tensor: blk.6.ssm_a with type F32 -> model.layers.6.linear_attn.ssm_a with dtype F32
160
+ Processing tensor: blk.6.ssm_alpha.weight with type Q8_0 -> model.layers.6.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
161
+ Processing tensor: blk.6.ssm_beta.weight with type Q8_0 -> model.layers.6.linear_attn.ssm_beta_proj.weight with dtype Q8_0
162
+ Processing tensor: blk.6.ssm_conv1d.weight with type F32 -> model.layers.6.linear_attn.ssm_conv1d.weight with dtype F32
163
+ [INFO] transpose conv1d
164
+ Processing tensor: blk.6.ssm_dt.bias with type F32 -> model.layers.6.linear_attn.ssm_dt.bias with dtype F32
165
+ Processing tensor: blk.6.ssm_norm.weight with type F32 -> model.layers.6.linear_attn.ssm_norm.weight with dtype F32
166
+ Processing tensor: blk.6.ssm_out.weight with type Q8_0 -> model.layers.6.linear_attn.ssm_out_proj.weight with dtype Q8_0
167
+ Processing tensor: blk.7.attn_k.weight with type Q8_0 -> model.layers.7.self_attn.k_proj.weight with dtype Q4_1
168
+ Processing tensor: blk.7.attn_k_norm.weight with type F32 -> model.layers.7.self_attn.k_norm.weight with dtype F32
169
+ Processing tensor: blk.7.attn_norm.weight with type F32 -> model.layers.7.input_layernorm.weight with dtype F32
170
+ Processing tensor: blk.7.attn_output.weight with type Q8_0 -> model.layers.7.self_attn.o_proj.weight with dtype Q4_1
171
+ Processing tensor: blk.7.attn_q.weight with type Q8_0 -> model.layers.7.self_attn.q_proj.weight with dtype Q4_1
172
+ [INFO] Seperate q, gate for q_proj
173
+ Processing tensor: blk.7.attn_q_norm.weight with type F32 -> model.layers.7.self_attn.q_norm.weight with dtype F32
174
+ Processing tensor: blk.7.attn_v.weight with type Q8_0 -> model.layers.7.self_attn.v_proj.weight with dtype Q4_1
175
+ Processing tensor: blk.7.ffn_down.weight with type Q8_0 -> model.layers.7.mlp.down_proj.weight with dtype Q4_1
176
+ Processing tensor: blk.7.ffn_gate.weight with type Q8_0 -> model.layers.7.mlp.gate_proj.weight with dtype Q4_1
177
+ Processing tensor: blk.7.ffn_up.weight with type Q8_0 -> model.layers.7.mlp.up_proj.weight with dtype Q4_1
178
+ Processing tensor: blk.7.post_attention_norm.weight with type F32 -> model.layers.7.post_attention_layernorm.weight with dtype F32
179
+ Processing tensor: blk.8.attn_gate.weight with type Q8_0 -> model.layers.8.self_attn.gate_proj.weight with dtype Q4_1
180
+ Processing tensor: blk.8.attn_norm.weight with type F32 -> model.layers.8.input_layernorm.weight with dtype F32
181
+ Processing tensor: blk.8.attn_qkv.weight with type Q8_0 -> model.layers.8.linear_attn.qkv_proj.weight with dtype Q4_1
182
+ Processing tensor: blk.8.ffn_down.weight with type Q8_0 -> model.layers.8.mlp.down_proj.weight with dtype Q4_1
183
+ Processing tensor: blk.8.ffn_gate.weight with type Q8_0 -> model.layers.8.mlp.gate_proj.weight with dtype Q4_1
184
+ Processing tensor: blk.8.ffn_up.weight with type Q8_0 -> model.layers.8.mlp.up_proj.weight with dtype Q4_1
185
+ Processing tensor: blk.8.post_attention_norm.weight with type F32 -> model.layers.8.post_attention_layernorm.weight with dtype F32
186
+ Processing tensor: blk.8.ssm_a with type F32 -> model.layers.8.linear_attn.ssm_a with dtype F32
187
+ Processing tensor: blk.8.ssm_alpha.weight with type Q8_0 -> model.layers.8.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
188
+ Processing tensor: blk.8.ssm_beta.weight with type Q8_0 -> model.layers.8.linear_attn.ssm_beta_proj.weight with dtype Q8_0
189
+ Processing tensor: blk.8.ssm_conv1d.weight with type F32 -> model.layers.8.linear_attn.ssm_conv1d.weight with dtype F32
190
+ [INFO] transpose conv1d
191
+ Processing tensor: blk.8.ssm_dt.bias with type F32 -> model.layers.8.linear_attn.ssm_dt.bias with dtype F32
192
+ Processing tensor: blk.8.ssm_norm.weight with type F32 -> model.layers.8.linear_attn.ssm_norm.weight with dtype F32
193
+ Processing tensor: blk.8.ssm_out.weight with type Q8_0 -> model.layers.8.linear_attn.ssm_out_proj.weight with dtype Q8_0
194
+ Processing tensor: blk.9.attn_gate.weight with type Q8_0 -> model.layers.9.self_attn.gate_proj.weight with dtype Q4_1
195
+ Processing tensor: blk.9.attn_norm.weight with type F32 -> model.layers.9.input_layernorm.weight with dtype F32
196
+ Processing tensor: blk.9.attn_qkv.weight with type Q8_0 -> model.layers.9.linear_attn.qkv_proj.weight with dtype Q4_1
197
+ Processing tensor: blk.9.ffn_down.weight with type Q8_0 -> model.layers.9.mlp.down_proj.weight with dtype Q4_1
198
+ Processing tensor: blk.9.ffn_gate.weight with type Q8_0 -> model.layers.9.mlp.gate_proj.weight with dtype Q4_1
199
+ Processing tensor: blk.9.ffn_up.weight with type Q8_0 -> model.layers.9.mlp.up_proj.weight with dtype Q4_1
200
+ Processing tensor: blk.9.post_attention_norm.weight with type F32 -> model.layers.9.post_attention_layernorm.weight with dtype F32
201
+ Processing tensor: blk.9.ssm_a with type F32 -> model.layers.9.linear_attn.ssm_a with dtype F32
202
+ Processing tensor: blk.9.ssm_alpha.weight with type Q8_0 -> model.layers.9.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
203
+ Processing tensor: blk.9.ssm_beta.weight with type Q8_0 -> model.layers.9.linear_attn.ssm_beta_proj.weight with dtype Q8_0
204
+ Processing tensor: blk.9.ssm_conv1d.weight with type F32 -> model.layers.9.linear_attn.ssm_conv1d.weight with dtype F32
205
+ [INFO] transpose conv1d
206
+ Processing tensor: blk.9.ssm_dt.bias with type F32 -> model.layers.9.linear_attn.ssm_dt.bias with dtype F32
207
+ Processing tensor: blk.9.ssm_norm.weight with type F32 -> model.layers.9.linear_attn.ssm_norm.weight with dtype F32
208
+ Processing tensor: blk.9.ssm_out.weight with type Q8_0 -> model.layers.9.linear_attn.ssm_out_proj.weight with dtype Q8_0
209
+ Processing tensor: blk.10.attn_gate.weight with type Q8_0 -> model.layers.10.self_attn.gate_proj.weight with dtype Q4_1
210
+ Processing tensor: blk.10.attn_norm.weight with type F32 -> model.layers.10.input_layernorm.weight with dtype F32
211
+ Processing tensor: blk.10.attn_qkv.weight with type Q8_0 -> model.layers.10.linear_attn.qkv_proj.weight with dtype Q4_1
212
+ Processing tensor: blk.10.ffn_down.weight with type Q8_0 -> model.layers.10.mlp.down_proj.weight with dtype Q4_1
213
+ Processing tensor: blk.10.ffn_gate.weight with type Q8_0 -> model.layers.10.mlp.gate_proj.weight with dtype Q4_1
214
+ Processing tensor: blk.10.ffn_up.weight with type Q8_0 -> model.layers.10.mlp.up_proj.weight with dtype Q4_1
215
+ Processing tensor: blk.10.post_attention_norm.weight with type F32 -> model.layers.10.post_attention_layernorm.weight with dtype F32
216
+ Processing tensor: blk.10.ssm_a with type F32 -> model.layers.10.linear_attn.ssm_a with dtype F32
217
+ Processing tensor: blk.10.ssm_alpha.weight with type Q8_0 -> model.layers.10.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
218
+ Processing tensor: blk.10.ssm_beta.weight with type Q8_0 -> model.layers.10.linear_attn.ssm_beta_proj.weight with dtype Q8_0
219
+ Processing tensor: blk.10.ssm_conv1d.weight with type F32 -> model.layers.10.linear_attn.ssm_conv1d.weight with dtype F32
220
+ [INFO] transpose conv1d
221
+ Processing tensor: blk.10.ssm_dt.bias with type F32 -> model.layers.10.linear_attn.ssm_dt.bias with dtype F32
222
+ Processing tensor: blk.10.ssm_norm.weight with type F32 -> model.layers.10.linear_attn.ssm_norm.weight with dtype F32
223
+ Processing tensor: blk.10.ssm_out.weight with type Q8_0 -> model.layers.10.linear_attn.ssm_out_proj.weight with dtype Q8_0
224
+ Processing tensor: blk.11.attn_k.weight with type Q8_0 -> model.layers.11.self_attn.k_proj.weight with dtype Q4_1
225
+ Processing tensor: blk.11.attn_k_norm.weight with type F32 -> model.layers.11.self_attn.k_norm.weight with dtype F32
226
+ Processing tensor: blk.11.attn_norm.weight with type F32 -> model.layers.11.input_layernorm.weight with dtype F32
227
+ Processing tensor: blk.11.attn_output.weight with type Q8_0 -> model.layers.11.self_attn.o_proj.weight with dtype Q4_1
228
+ Processing tensor: blk.11.attn_q.weight with type Q8_0 -> model.layers.11.self_attn.q_proj.weight with dtype Q4_1
229
+ [INFO] Seperate q, gate for q_proj
230
+ Processing tensor: blk.11.attn_q_norm.weight with type F32 -> model.layers.11.self_attn.q_norm.weight with dtype F32
231
+ Processing tensor: blk.11.attn_v.weight with type Q8_0 -> model.layers.11.self_attn.v_proj.weight with dtype Q4_1
232
+ Processing tensor: blk.11.ffn_down.weight with type Q8_0 -> model.layers.11.mlp.down_proj.weight with dtype Q4_1
233
+ Processing tensor: blk.11.ffn_gate.weight with type Q8_0 -> model.layers.11.mlp.gate_proj.weight with dtype Q4_1
234
+ Processing tensor: blk.11.ffn_up.weight with type Q8_0 -> model.layers.11.mlp.up_proj.weight with dtype Q4_1
235
+ Processing tensor: blk.11.post_attention_norm.weight with type F32 -> model.layers.11.post_attention_layernorm.weight with dtype F32
236
+ Processing tensor: blk.12.attn_gate.weight with type Q8_0 -> model.layers.12.self_attn.gate_proj.weight with dtype Q4_1
237
+ Processing tensor: blk.12.attn_norm.weight with type F32 -> model.layers.12.input_layernorm.weight with dtype F32
238
+ Processing tensor: blk.12.attn_qkv.weight with type Q8_0 -> model.layers.12.linear_attn.qkv_proj.weight with dtype Q4_1
239
+ Processing tensor: blk.12.ffn_down.weight with type Q8_0 -> model.layers.12.mlp.down_proj.weight with dtype Q4_1
240
+ Processing tensor: blk.12.ffn_gate.weight with type Q8_0 -> model.layers.12.mlp.gate_proj.weight with dtype Q4_1
241
+ Processing tensor: blk.12.ffn_up.weight with type Q8_0 -> model.layers.12.mlp.up_proj.weight with dtype Q4_1
242
+ Processing tensor: blk.12.post_attention_norm.weight with type F32 -> model.layers.12.post_attention_layernorm.weight with dtype F32
243
+ Processing tensor: blk.12.ssm_a with type F32 -> model.layers.12.linear_attn.ssm_a with dtype F32
244
+ Processing tensor: blk.12.ssm_alpha.weight with type Q8_0 -> model.layers.12.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
245
+ Processing tensor: blk.12.ssm_beta.weight with type Q8_0 -> model.layers.12.linear_attn.ssm_beta_proj.weight with dtype Q8_0
246
+ Processing tensor: blk.12.ssm_conv1d.weight with type F32 -> model.layers.12.linear_attn.ssm_conv1d.weight with dtype F32
247
+ [INFO] transpose conv1d
248
+ Processing tensor: blk.12.ssm_dt.bias with type F32 -> model.layers.12.linear_attn.ssm_dt.bias with dtype F32
249
+ Processing tensor: blk.12.ssm_norm.weight with type F32 -> model.layers.12.linear_attn.ssm_norm.weight with dtype F32
250
+ Processing tensor: blk.12.ssm_out.weight with type Q8_0 -> model.layers.12.linear_attn.ssm_out_proj.weight with dtype Q8_0
251
+ Processing tensor: blk.13.attn_gate.weight with type Q8_0 -> model.layers.13.self_attn.gate_proj.weight with dtype Q4_1
252
+ Processing tensor: blk.13.attn_norm.weight with type F32 -> model.layers.13.input_layernorm.weight with dtype F32
253
+ Processing tensor: blk.13.attn_qkv.weight with type Q8_0 -> model.layers.13.linear_attn.qkv_proj.weight with dtype Q4_1
254
+ Processing tensor: blk.13.ffn_down.weight with type Q8_0 -> model.layers.13.mlp.down_proj.weight with dtype Q4_1
255
+ Processing tensor: blk.13.ffn_gate.weight with type Q8_0 -> model.layers.13.mlp.gate_proj.weight with dtype Q4_1
256
+ Processing tensor: blk.13.ffn_up.weight with type Q8_0 -> model.layers.13.mlp.up_proj.weight with dtype Q4_1
257
+ Processing tensor: blk.13.post_attention_norm.weight with type F32 -> model.layers.13.post_attention_layernorm.weight with dtype F32
258
+ Processing tensor: blk.13.ssm_a with type F32 -> model.layers.13.linear_attn.ssm_a with dtype F32
259
+ Processing tensor: blk.13.ssm_alpha.weight with type Q8_0 -> model.layers.13.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
260
+ Processing tensor: blk.13.ssm_beta.weight with type Q8_0 -> model.layers.13.linear_attn.ssm_beta_proj.weight with dtype Q8_0
261
+ Processing tensor: blk.13.ssm_conv1d.weight with type F32 -> model.layers.13.linear_attn.ssm_conv1d.weight with dtype F32
262
+ [INFO] transpose conv1d
263
+ Processing tensor: blk.13.ssm_dt.bias with type F32 -> model.layers.13.linear_attn.ssm_dt.bias with dtype F32
264
+ Processing tensor: blk.13.ssm_norm.weight with type F32 -> model.layers.13.linear_attn.ssm_norm.weight with dtype F32
265
+ Processing tensor: blk.13.ssm_out.weight with type Q8_0 -> model.layers.13.linear_attn.ssm_out_proj.weight with dtype Q8_0
266
+ Processing tensor: blk.14.attn_gate.weight with type Q8_0 -> model.layers.14.self_attn.gate_proj.weight with dtype Q4_1
267
+ Processing tensor: blk.14.attn_norm.weight with type F32 -> model.layers.14.input_layernorm.weight with dtype F32
268
+ Processing tensor: blk.14.attn_qkv.weight with type Q8_0 -> model.layers.14.linear_attn.qkv_proj.weight with dtype Q4_1
269
+ Processing tensor: blk.14.ffn_down.weight with type Q8_0 -> model.layers.14.mlp.down_proj.weight with dtype Q4_1
270
+ Processing tensor: blk.14.ffn_gate.weight with type Q8_0 -> model.layers.14.mlp.gate_proj.weight with dtype Q4_1
271
+ Processing tensor: blk.14.ffn_up.weight with type Q8_0 -> model.layers.14.mlp.up_proj.weight with dtype Q4_1
272
+ Processing tensor: blk.14.post_attention_norm.weight with type F32 -> model.layers.14.post_attention_layernorm.weight with dtype F32
273
+ Processing tensor: blk.14.ssm_a with type F32 -> model.layers.14.linear_attn.ssm_a with dtype F32
274
+ Processing tensor: blk.14.ssm_alpha.weight with type Q8_0 -> model.layers.14.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
275
+ Processing tensor: blk.14.ssm_beta.weight with type Q8_0 -> model.layers.14.linear_attn.ssm_beta_proj.weight with dtype Q8_0
276
+ Processing tensor: blk.14.ssm_conv1d.weight with type F32 -> model.layers.14.linear_attn.ssm_conv1d.weight with dtype F32
277
+ [INFO] transpose conv1d
278
+ Processing tensor: blk.14.ssm_dt.bias with type F32 -> model.layers.14.linear_attn.ssm_dt.bias with dtype F32
279
+ Processing tensor: blk.14.ssm_norm.weight with type F32 -> model.layers.14.linear_attn.ssm_norm.weight with dtype F32
280
+ Processing tensor: blk.14.ssm_out.weight with type Q8_0 -> model.layers.14.linear_attn.ssm_out_proj.weight with dtype Q8_0
281
+ Processing tensor: blk.15.attn_k.weight with type Q8_0 -> model.layers.15.self_attn.k_proj.weight with dtype Q4_1
282
+ Processing tensor: blk.15.attn_k_norm.weight with type F32 -> model.layers.15.self_attn.k_norm.weight with dtype F32
283
+ Processing tensor: blk.15.attn_norm.weight with type F32 -> model.layers.15.input_layernorm.weight with dtype F32
284
+ Processing tensor: blk.15.attn_output.weight with type Q8_0 -> model.layers.15.self_attn.o_proj.weight with dtype Q4_1
285
+ Processing tensor: blk.15.attn_q.weight with type Q8_0 -> model.layers.15.self_attn.q_proj.weight with dtype Q4_1
286
+ [INFO] Seperate q, gate for q_proj
287
+ Processing tensor: blk.15.attn_q_norm.weight with type F32 -> model.layers.15.self_attn.q_norm.weight with dtype F32
288
+ Processing tensor: blk.15.attn_v.weight with type Q8_0 -> model.layers.15.self_attn.v_proj.weight with dtype Q4_1
289
+ Processing tensor: blk.15.ffn_down.weight with type Q8_0 -> model.layers.15.mlp.down_proj.weight with dtype Q4_1
290
+ Processing tensor: blk.15.ffn_gate.weight with type Q8_0 -> model.layers.15.mlp.gate_proj.weight with dtype Q4_1
291
+ Processing tensor: blk.15.ffn_up.weight with type Q8_0 -> model.layers.15.mlp.up_proj.weight with dtype Q4_1
292
+ Processing tensor: blk.15.post_attention_norm.weight with type F32 -> model.layers.15.post_attention_layernorm.weight with dtype F32
293
+ Processing tensor: blk.16.attn_gate.weight with type Q8_0 -> model.layers.16.self_attn.gate_proj.weight with dtype Q4_1
294
+ Processing tensor: blk.16.attn_norm.weight with type F32 -> model.layers.16.input_layernorm.weight with dtype F32
295
+ Processing tensor: blk.16.attn_qkv.weight with type Q8_0 -> model.layers.16.linear_attn.qkv_proj.weight with dtype Q4_1
296
+ Processing tensor: blk.16.ffn_down.weight with type Q8_0 -> model.layers.16.mlp.down_proj.weight with dtype Q4_1
297
+ Processing tensor: blk.16.ffn_gate.weight with type Q8_0 -> model.layers.16.mlp.gate_proj.weight with dtype Q4_1
298
+ Processing tensor: blk.16.ffn_up.weight with type Q8_0 -> model.layers.16.mlp.up_proj.weight with dtype Q4_1
299
+ Processing tensor: blk.16.post_attention_norm.weight with type F32 -> model.layers.16.post_attention_layernorm.weight with dtype F32
300
+ Processing tensor: blk.16.ssm_a with type F32 -> model.layers.16.linear_attn.ssm_a with dtype F32
301
+ Processing tensor: blk.16.ssm_alpha.weight with type Q8_0 -> model.layers.16.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
302
+ Processing tensor: blk.16.ssm_beta.weight with type Q8_0 -> model.layers.16.linear_attn.ssm_beta_proj.weight with dtype Q8_0
303
+ Processing tensor: blk.16.ssm_conv1d.weight with type F32 -> model.layers.16.linear_attn.ssm_conv1d.weight with dtype F32
304
+ [INFO] transpose conv1d
305
+ Processing tensor: blk.16.ssm_dt.bias with type F32 -> model.layers.16.linear_attn.ssm_dt.bias with dtype F32
306
+ Processing tensor: blk.16.ssm_norm.weight with type F32 -> model.layers.16.linear_attn.ssm_norm.weight with dtype F32
307
+ Processing tensor: blk.16.ssm_out.weight with type Q8_0 -> model.layers.16.linear_attn.ssm_out_proj.weight with dtype Q8_0
308
+ Processing tensor: blk.17.attn_gate.weight with type Q8_0 -> model.layers.17.self_attn.gate_proj.weight with dtype Q4_1
309
+ Processing tensor: blk.17.attn_norm.weight with type F32 -> model.layers.17.input_layernorm.weight with dtype F32
310
+ Processing tensor: blk.17.attn_qkv.weight with type Q8_0 -> model.layers.17.linear_attn.qkv_proj.weight with dtype Q4_1
311
+ Processing tensor: blk.17.ffn_down.weight with type Q8_0 -> model.layers.17.mlp.down_proj.weight with dtype Q4_1
312
+ Processing tensor: blk.17.ffn_gate.weight with type Q8_0 -> model.layers.17.mlp.gate_proj.weight with dtype Q4_1
313
+ Processing tensor: blk.17.ffn_up.weight with type Q8_0 -> model.layers.17.mlp.up_proj.weight with dtype Q4_1
314
+ Processing tensor: blk.17.post_attention_norm.weight with type F32 -> model.layers.17.post_attention_layernorm.weight with dtype F32
315
+ Processing tensor: blk.17.ssm_a with type F32 -> model.layers.17.linear_attn.ssm_a with dtype F32
316
+ Processing tensor: blk.17.ssm_alpha.weight with type Q8_0 -> model.layers.17.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
317
+ Processing tensor: blk.17.ssm_beta.weight with type Q8_0 -> model.layers.17.linear_attn.ssm_beta_proj.weight with dtype Q8_0
318
+ Processing tensor: blk.17.ssm_conv1d.weight with type F32 -> model.layers.17.linear_attn.ssm_conv1d.weight with dtype F32
319
+ [INFO] transpose conv1d
320
+ Processing tensor: blk.17.ssm_dt.bias with type F32 -> model.layers.17.linear_attn.ssm_dt.bias with dtype F32
321
+ Processing tensor: blk.17.ssm_norm.weight with type F32 -> model.layers.17.linear_attn.ssm_norm.weight with dtype F32
322
+ Processing tensor: blk.17.ssm_out.weight with type Q8_0 -> model.layers.17.linear_attn.ssm_out_proj.weight with dtype Q8_0
323
+ Processing tensor: blk.18.attn_gate.weight with type Q8_0 -> model.layers.18.self_attn.gate_proj.weight with dtype Q4_1
324
+ Processing tensor: blk.18.attn_norm.weight with type F32 -> model.layers.18.input_layernorm.weight with dtype F32
325
+ Processing tensor: blk.18.attn_qkv.weight with type Q8_0 -> model.layers.18.linear_attn.qkv_proj.weight with dtype Q4_1
326
+ Processing tensor: blk.18.ffn_down.weight with type Q8_0 -> model.layers.18.mlp.down_proj.weight with dtype Q4_1
327
+ Processing tensor: blk.18.ffn_gate.weight with type Q8_0 -> model.layers.18.mlp.gate_proj.weight with dtype Q4_1
328
+ Processing tensor: blk.18.ffn_up.weight with type Q8_0 -> model.layers.18.mlp.up_proj.weight with dtype Q4_1
329
+ Processing tensor: blk.18.post_attention_norm.weight with type F32 -> model.layers.18.post_attention_layernorm.weight with dtype F32
330
+ Processing tensor: blk.18.ssm_a with type F32 -> model.layers.18.linear_attn.ssm_a with dtype F32
331
+ Processing tensor: blk.18.ssm_alpha.weight with type Q8_0 -> model.layers.18.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
332
+ Processing tensor: blk.18.ssm_beta.weight with type Q8_0 -> model.layers.18.linear_attn.ssm_beta_proj.weight with dtype Q8_0
333
+ Processing tensor: blk.18.ssm_conv1d.weight with type F32 -> model.layers.18.linear_attn.ssm_conv1d.weight with dtype F32
334
+ [INFO] transpose conv1d
335
+ Processing tensor: blk.18.ssm_dt.bias with type F32 -> model.layers.18.linear_attn.ssm_dt.bias with dtype F32
336
+ Processing tensor: blk.18.ssm_norm.weight with type F32 -> model.layers.18.linear_attn.ssm_norm.weight with dtype F32
337
+ Processing tensor: blk.18.ssm_out.weight with type Q8_0 -> model.layers.18.linear_attn.ssm_out_proj.weight with dtype Q8_0
338
+ Processing tensor: blk.19.attn_k.weight with type Q8_0 -> model.layers.19.self_attn.k_proj.weight with dtype Q4_1
339
+ Processing tensor: blk.19.attn_k_norm.weight with type F32 -> model.layers.19.self_attn.k_norm.weight with dtype F32
340
+ Processing tensor: blk.19.attn_norm.weight with type F32 -> model.layers.19.input_layernorm.weight with dtype F32
341
+ Processing tensor: blk.19.attn_output.weight with type Q8_0 -> model.layers.19.self_attn.o_proj.weight with dtype Q4_1
342
+ Processing tensor: blk.19.attn_q.weight with type Q8_0 -> model.layers.19.self_attn.q_proj.weight with dtype Q4_1
343
+ [INFO] Seperate q, gate for q_proj
344
+ Processing tensor: blk.19.attn_q_norm.weight with type F32 -> model.layers.19.self_attn.q_norm.weight with dtype F32
345
+ Processing tensor: blk.19.attn_v.weight with type Q8_0 -> model.layers.19.self_attn.v_proj.weight with dtype Q4_1
346
+ Processing tensor: blk.19.ffn_down.weight with type Q8_0 -> model.layers.19.mlp.down_proj.weight with dtype Q4_1
347
+ Processing tensor: blk.19.ffn_gate.weight with type Q8_0 -> model.layers.19.mlp.gate_proj.weight with dtype Q4_1
348
+ Processing tensor: blk.19.ffn_up.weight with type Q8_0 -> model.layers.19.mlp.up_proj.weight with dtype Q4_1
349
+ Processing tensor: blk.19.post_attention_norm.weight with type F32 -> model.layers.19.post_attention_layernorm.weight with dtype F32
350
+ Processing tensor: blk.20.attn_gate.weight with type Q8_0 -> model.layers.20.self_attn.gate_proj.weight with dtype Q4_1
351
+ Processing tensor: blk.20.attn_norm.weight with type F32 -> model.layers.20.input_layernorm.weight with dtype F32
352
+ Processing tensor: blk.20.attn_qkv.weight with type Q8_0 -> model.layers.20.linear_attn.qkv_proj.weight with dtype Q4_1
353
+ Processing tensor: blk.20.ffn_down.weight with type Q8_0 -> model.layers.20.mlp.down_proj.weight with dtype Q4_1
354
+ Processing tensor: blk.20.ffn_gate.weight with type Q8_0 -> model.layers.20.mlp.gate_proj.weight with dtype Q4_1
355
+ Processing tensor: blk.20.ffn_up.weight with type Q8_0 -> model.layers.20.mlp.up_proj.weight with dtype Q4_1
356
+ Processing tensor: blk.20.post_attention_norm.weight with type F32 -> model.layers.20.post_attention_layernorm.weight with dtype F32
357
+ Processing tensor: blk.20.ssm_a with type F32 -> model.layers.20.linear_attn.ssm_a with dtype F32
358
+ Processing tensor: blk.20.ssm_alpha.weight with type Q8_0 -> model.layers.20.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
359
+ Processing tensor: blk.20.ssm_beta.weight with type Q8_0 -> model.layers.20.linear_attn.ssm_beta_proj.weight with dtype Q8_0
360
+ Processing tensor: blk.20.ssm_conv1d.weight with type F32 -> model.layers.20.linear_attn.ssm_conv1d.weight with dtype F32
361
+ [INFO] transpose conv1d
362
+ Processing tensor: blk.20.ssm_dt.bias with type F32 -> model.layers.20.linear_attn.ssm_dt.bias with dtype F32
363
+ Processing tensor: blk.20.ssm_norm.weight with type F32 -> model.layers.20.linear_attn.ssm_norm.weight with dtype F32
364
+ Processing tensor: blk.20.ssm_out.weight with type Q8_0 -> model.layers.20.linear_attn.ssm_out_proj.weight with dtype Q8_0
365
+ Processing tensor: blk.21.attn_gate.weight with type Q8_0 -> model.layers.21.self_attn.gate_proj.weight with dtype Q4_1
366
+ Processing tensor: blk.21.attn_norm.weight with type F32 -> model.layers.21.input_layernorm.weight with dtype F32
367
+ Processing tensor: blk.21.attn_qkv.weight with type Q8_0 -> model.layers.21.linear_attn.qkv_proj.weight with dtype Q4_1
368
+ Processing tensor: blk.21.ffn_down.weight with type Q8_0 -> model.layers.21.mlp.down_proj.weight with dtype Q4_1
369
+ Processing tensor: blk.21.ffn_gate.weight with type Q8_0 -> model.layers.21.mlp.gate_proj.weight with dtype Q4_1
370
+ Processing tensor: blk.21.ffn_up.weight with type Q8_0 -> model.layers.21.mlp.up_proj.weight with dtype Q4_1
371
+ Processing tensor: blk.21.post_attention_norm.weight with type F32 -> model.layers.21.post_attention_layernorm.weight with dtype F32
372
+ Processing tensor: blk.21.ssm_a with type F32 -> model.layers.21.linear_attn.ssm_a with dtype F32
373
+ Processing tensor: blk.21.ssm_alpha.weight with type Q8_0 -> model.layers.21.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
374
+ Processing tensor: blk.21.ssm_beta.weight with type Q8_0 -> model.layers.21.linear_attn.ssm_beta_proj.weight with dtype Q8_0
375
+ Processing tensor: blk.21.ssm_conv1d.weight with type F32 -> model.layers.21.linear_attn.ssm_conv1d.weight with dtype F32
376
+ [INFO] transpose conv1d
377
+ Processing tensor: blk.21.ssm_dt.bias with type F32 -> model.layers.21.linear_attn.ssm_dt.bias with dtype F32
378
+ Processing tensor: blk.21.ssm_norm.weight with type F32 -> model.layers.21.linear_attn.ssm_norm.weight with dtype F32
379
+ Processing tensor: blk.21.ssm_out.weight with type Q8_0 -> model.layers.21.linear_attn.ssm_out_proj.weight with dtype Q8_0
380
+ Processing tensor: blk.22.attn_gate.weight with type Q8_0 -> model.layers.22.self_attn.gate_proj.weight with dtype Q4_1
381
+ Processing tensor: blk.22.attn_norm.weight with type F32 -> model.layers.22.input_layernorm.weight with dtype F32
382
+ Processing tensor: blk.22.attn_qkv.weight with type Q8_0 -> model.layers.22.linear_attn.qkv_proj.weight with dtype Q4_1
383
+ Processing tensor: blk.22.ffn_down.weight with type Q8_0 -> model.layers.22.mlp.down_proj.weight with dtype Q4_1
384
+ Processing tensor: blk.22.ffn_gate.weight with type Q8_0 -> model.layers.22.mlp.gate_proj.weight with dtype Q4_1
385
+ Processing tensor: blk.22.ffn_up.weight with type Q8_0 -> model.layers.22.mlp.up_proj.weight with dtype Q4_1
386
+ Processing tensor: blk.22.post_attention_norm.weight with type F32 -> model.layers.22.post_attention_layernorm.weight with dtype F32
387
+ Processing tensor: blk.22.ssm_a with type F32 -> model.layers.22.linear_attn.ssm_a with dtype F32
388
+ Processing tensor: blk.22.ssm_alpha.weight with type Q8_0 -> model.layers.22.linear_attn.ssm_alpha_proj.weight with dtype Q8_0
389
+ Processing tensor: blk.22.ssm_beta.weight with type Q8_0 -> model.layers.22.linear_attn.ssm_beta_proj.weight with dtype Q8_0
390
+ Processing tensor: blk.22.ssm_conv1d.weight with type F32 -> model.layers.22.linear_attn.ssm_conv1d.weight with dtype F32
391
+ [INFO] transpose conv1d
392
+ Processing tensor: blk.22.ssm_dt.bias with type F32 -> model.layers.22.linear_attn.ssm_dt.bias with dtype F32
393
+ Processing tensor: blk.22.ssm_norm.weight with type F32 -> model.layers.22.linear_attn.ssm_norm.weight with dtype F32
394
+ Processing tensor: blk.22.ssm_out.weight with type Q8_0 -> model.layers.22.linear_attn.ssm_out_proj.weight with dtype Q8_0
395
+ Processing tensor: blk.23.attn_k.weight with type Q8_0 -> model.layers.23.self_attn.k_proj.weight with dtype Q4_1
396
+ Processing tensor: blk.23.attn_k_norm.weight with type F32 -> model.layers.23.self_attn.k_norm.weight with dtype F32
397
+ Processing tensor: blk.23.attn_norm.weight with type F32 -> model.layers.23.input_layernorm.weight with dtype F32
398
+ Processing tensor: blk.23.attn_output.weight with type Q8_0 -> model.layers.23.self_attn.o_proj.weight with dtype Q4_1
399
+ Processing tensor: blk.23.attn_q.weight with type Q8_0 -> model.layers.23.self_attn.q_proj.weight with dtype Q4_1
400
+ [INFO] Seperate q, gate for q_proj
401
+ Processing tensor: blk.23.attn_q_norm.weight with type F32 -> model.layers.23.self_attn.q_norm.weight with dtype F32
402
+ Processing tensor: blk.23.attn_v.weight with type Q8_0 -> model.layers.23.self_attn.v_proj.weight with dtype Q4_1
403
+ Processing tensor: blk.23.ffn_down.weight with type Q8_0 -> model.layers.23.mlp.down_proj.weight with dtype Q4_1
404
+ Processing tensor: blk.23.ffn_gate.weight with type Q8_0 -> model.layers.23.mlp.gate_proj.weight with dtype Q4_1
405
+ Processing tensor: blk.23.ffn_up.weight with type Q8_0 -> model.layers.23.mlp.up_proj.weight with dtype Q4_1
406
+ Processing tensor: blk.23.post_attention_norm.weight with type F32 -> model.layers.23.post_attention_layernorm.weight with dtype F32
407
+ Processing tensor: blk.24.attn_k.weight with type Q8_0 -> model.layers.24.self_attn.k_proj.weight with dtype Q4_1
408
+ Processing tensor: blk.24.attn_k_norm.weight with type F32 -> model.layers.24.self_attn.k_norm.weight with dtype F32
409
+ Processing tensor: blk.24.attn_norm.weight with type F32 -> model.layers.24.input_layernorm.weight with dtype F32
410
+ Processing tensor: blk.24.attn_output.weight with type Q8_0 -> model.layers.24.self_attn.o_proj.weight with dtype Q4_1
411
+ Processing tensor: blk.24.attn_q.weight with type Q8_0 -> model.layers.24.self_attn.q_proj.weight with dtype Q4_1
412
+ [INFO] Seperate q, gate for q_proj
413
+ Processing tensor: blk.24.attn_q_norm.weight with type F32 -> model.layers.24.self_attn.q_norm.weight with dtype F32
414
+ Processing tensor: blk.24.attn_v.weight with type Q8_0 -> model.layers.24.self_attn.v_proj.weight with dtype Q4_1
415
+ Processing tensor: blk.24.ffn_down.weight with type Q8_0 -> model.layers.24.mlp.down_proj.weight with dtype Q4_1
416
+ Processing tensor: blk.24.ffn_gate.weight with type Q8_0 -> model.layers.24.mlp.gate_proj.weight with dtype Q4_1
417
+ Processing tensor: blk.24.ffn_up.weight with type Q8_0 -> model.layers.24.mlp.up_proj.weight with dtype Q4_1
418
+ [SKIP] blk.24.nextn.eh_proj.weight (MTP next-token prediction weights, absent from official Q4NX)
419
+ [SKIP] blk.24.nextn.enorm.weight (MTP next-token prediction weights, absent from official Q4NX)
420
+ [SKIP] blk.24.nextn.hnorm.weight (MTP next-token prediction weights, absent from official Q4NX)
421
+ [SKIP] blk.24.nextn.shared_head_norm.weight (MTP next-token prediction weights, absent from official Q4NX)
422
+ Processing tensor: blk.24.post_attention_norm.weight with type F32 -> model.layers.24.post_attention_layernorm.weight with dtype F32
423
+ [INFO] Extracting tokenizer JSON...
424
+ [INFO] EOS token ID: 248046
425
+ [INFO] Padding token ID: 248044
426
+ [INFO] Tokenizer saved to /home/atomic-germ/Projects/q4nx-build-v0.2.0/Qwen3.8-2B-Distill-NPU2/tokenizer.json
427
+ [INFO] Saving Q4NX tensors to /home/atomic-germ/Projects/q4nx-build-v0.2.0/Qwen3.8-2B-Distill-NPU2/model.q4nx...
428
+ [INFO] Copied model assets from local source: /home/atomic-germ/.cache/huggingface/hub/models--Atomic-Germ--Qwen3.5-2B-NPU2/snapshots/92b00a5dc3487149c471c158628f63d9a621bb25
429
+ [INFO] Model assets present: ['vision_weight.q4nx', 'config.json', 'tokenizer.json', 'tokenizer_config.json', 'chat_template.jinja']
430
+ [INFO] Keeping skeleton-provided vision_config (matches shipped vision weights)
431
+ [INFO] Patched tokenizer_config.json with token ids from GGUF metadata
432
+ [INFO] Writing README.md based on Atomic-Germ/Qwen3.5-2B-NPU2's model card
433
+ [INFO] Model directory ready: /home/atomic-germ/Projects/q4nx-build-v0.2.0/Qwen3.8-2B-Distill-NPU2
434
+ [WARN] Overwriting existing model directory: /home/atomic-germ/.config/flm/models/Qwen3.8-Distill-2B-NPU2
435
+ [INFO] Deployed model files to: /home/atomic-germ/.config/flm/models/Qwen3.8-Distill-2B-NPU2
436
+ [INFO] Registered tag 'qwen3.8-distill:2b' in: /home/atomic-germ/.config/flm/model_list.json
437
+ [INFO] Registry defaults copied from official entry: qwen3.5:2b
438
+ [INFO] FLM does not read this registry yet. Add to your shell rc:
439
+ export FLM_CONFIG_PATH=/home/atomic-germ/.config/flm/model_list.json
440
+ [INFO] Linked xclbins for 'Qwen3.8-Distill-2B-NPU2' -> 'Qwen3.5-2B-NPU2'
441
+ [INFO] The runtime needs to find these xclbins. Add to your shell rc:
442
+ export FLM_XCLBIN_PATH=/home/atomic-germ/.config/flm
443
+ [INFO] Conversion complete! Output saved to /home/atomic-germ/Projects/q4nx-build-v0.2.0/Qwen3.8-2B-Distill-NPU2