Update README.md
Browse files
README.md
CHANGED
|
@@ -13,7 +13,7 @@ tags:
|
|
| 13 |
`Qwen/Qwen3.6-35B-A3B`, trained with the open-source NLA pipeline. Companion model:
|
| 14 |
[qwen3.6-35B-A3B-ar-sft](https://huggingface.co/stanleytheli/qwen3.6-35B-A3B-ar-sft).
|
| 15 |
|
| 16 |
-
**THIS IS THE SFT CHECKPOINT. FOR THE RL'D MODEL, SEE** [qwen3.6-35B-A3B-
|
| 17 |
|
| 18 |
Given a layer-29 residual-stream activation (d=2048) injected at a marker token in its
|
| 19 |
prompt, this model generates a natural-language description of the context that produced
|
|
|
|
| 13 |
`Qwen/Qwen3.6-35B-A3B`, trained with the open-source NLA pipeline. Companion model:
|
| 14 |
[qwen3.6-35B-A3B-ar-sft](https://huggingface.co/stanleytheli/qwen3.6-35B-A3B-ar-sft).
|
| 15 |
|
| 16 |
+
**THIS IS THE SFT CHECKPOINT. FOR THE RL'D MODEL, SEE** [qwen3.6-35B-A3B-av-RL1200](https://huggingface.co/stanleytheli/qwen3.6-35B-A3B-av-RL1200).
|
| 17 |
|
| 18 |
Given a layer-29 residual-stream activation (d=2048) injected at a marker token in its
|
| 19 |
prompt, this model generates a natural-language description of the context that produced
|