prithivMLmods commited on
Commit
cea2293
·
verified ·
1 Parent(s): ddffd9e

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +2 -9
README.md CHANGED
@@ -11,6 +11,8 @@ tags:
11
  - llm-compressor
12
  - vllm
13
  - clef
 
 
14
  - cloudflare
15
  - systemone
16
  - qwen3.8
@@ -141,15 +143,6 @@ processor.save_pretrained(dst)
141
  Then copy these files from the original Clef repo into `clef-FP8/` unchanged:
142
  `joint_head.safetensors`, `joint_head_config.json`, `joint_schema_model.py`.
143
 
144
- ## Evaluation
145
-
146
- I have **not** run the Decision Index or the workflow evals on this FP8 checkpoint. The numbers
147
- in the [Clef card](https://huggingface.co/Cloudflare/clef) are for the BF16 model and should not be
148
- assumed to carry over exactly. FP8 dynamic quantization is generally close to lossless, but Clef
149
- outputs per-option probabilities, so small shifts in calibration (confidence values) are possible
150
- even when the argmax decisions match. If you depend on calibrated probabilities or thresholds,
151
- validate on your own data before switching.
152
-
153
  ## Notes and limitations
154
 
155
  - **Joint head:** the joint schema head is stored separately and is kept in its original
 
11
  - llm-compressor
12
  - vllm
13
  - clef
14
+ - fp8
15
+ - w8a8
16
  - cloudflare
17
  - systemone
18
  - qwen3.8
 
143
  Then copy these files from the original Clef repo into `clef-FP8/` unchanged:
144
  `joint_head.safetensors`, `joint_head_config.json`, `joint_schema_model.py`.
145
 
 
 
 
 
 
 
 
 
 
146
  ## Notes and limitations
147
 
148
  - **Joint head:** the joint schema head is stored separately and is kept in its original