mili-tan commited on
Commit
7fcc5ab
·
verified ·
1 Parent(s): cd02637

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +48 -0
README.md ADDED
@@ -0,0 +1,48 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ base_model: togethercomputer/Tev1-0.8B-experimental
3
+ language:
4
+ - en
5
+ library_name: llama.cpp
6
+ pipeline_tag: text-generation
7
+ tags:
8
+ - decision-model
9
+ - gguf
10
+ - qwen3.5
11
+ - tev1
12
+ ---
13
+
14
+ # Tev1-0.8B-experimental GGUF
15
+
16
+ Quantized GGUF conversions of
17
+ [togethercomputer/Tev1-0.8B-experimental](https://huggingface.co/togethercomputer/Tev1-0.8B-experimental),
18
+ an experimental Qwen3.5-0.8B decision model from Together AI. It is a supervised
19
+ fine-tune trained to choose one option from a structured state, question, and a
20
+ list of 2-24 labeled choices, returning a single option letter.
21
+
22
+ These files run with [dohnuts.cpp](https://github.com/DreamBlooms/dohnuts.cpp)
23
+ through its `tev1` profile.
24
+
25
+ | File | Quantization |
26
+ | --- | --- |
27
+ | `tev1-f16.gguf` | F16 |
28
+ | `tev1-Q8_0.gguf` | Q8_0 |
29
+
30
+ `tev1.json` (the profile config) is required alongside the GGUF:
31
+
32
+ ```sh
33
+ build/dohnuts-cli --model tev1-0.8b-q8_0.gguf --metadata tev1.json
34
+ ```
35
+
36
+ ## Conversion
37
+
38
+ A full fine-tune, converted directly:
39
+
40
+ ```sh
41
+ scripts/build_tev1_gguf.sh <tev1-0.8b-dir> work/side/tev1-0.8b-q8_0.gguf
42
+ ```
43
+
44
+ ## License
45
+
46
+ The base Qwen3.5-0.8B model is Apache-2.0. The upstream release license for the
47
+ fine-tuned weights is being finalized; see the upstream model card for the
48
+ current terms.