Akahsizrr commited on
Commit
97cc08f
·
verified ·
1 Parent(s): d2df1b4

Add quantized version links and deployment docs

Browse files
Files changed (1) hide show
  1. README.md +8 -0
README.md CHANGED
@@ -186,6 +186,14 @@ print(response)
186
 
187
  ## Quantization & Deployment
188
 
 
 
 
 
 
 
 
 
189
  ### bitsandbytes 4-bit (NF4) — ~4.5 GB VRAM
190
 
191
  ```python
 
186
 
187
  ## Quantization & Deployment
188
 
189
+ ### Pre-quantized Versions
190
+
191
+ | Version | Repo | VRAM | Format |
192
+ |---------|------|------|--------|
193
+ | **4-bit NF4** | [`Akahsizrr/fuse-1-Lite-4bit`](https://huggingface.co/Akahsizrr/fuse-1-Lite-4bit) | 3.36 GB | bitsandbytes |
194
+ | **8-bit** | [`Akahsizrr/fuse-1-Lite-8bit`](https://huggingface.co/Akahsizrr/fuse-1-Lite-8bit) | 6.00 GB | bitsandbytes |
195
+ | **bfloat16** | This repo | ~12 GB | safetensors |
196
+
197
  ### bitsandbytes 4-bit (NF4) — ~4.5 GB VRAM
198
 
199
  ```python