Transformers
GGUF
English
llama
TheBloke commited on
Commit
e8f6c4c
·
1 Parent(s): defd511

Upload README.md

Browse files
Files changed (1) hide show
  1. README.md +6 -31
README.md CHANGED
@@ -80,15 +80,8 @@ A chat between a curious user and an artificial intelligence assistant. The assi
80
  ```
81
 
82
  <!-- prompt-template end -->
83
- <!-- licensing start -->
84
- ## Licensing
85
 
86
- The creator of the source model has listed its license as `other`, and this quantization has therefore used that same license.
87
 
88
- As this model is based on Llama 2, it is also subject to the Meta Llama 2 license terms, and the license files for that are additionally included. It should therefore be considered as being claimed to be licensed under both licenses. I contacted Hugging Face for clarification on dual licensing but they do not yet have an official position. Should this change, or should Meta provide any feedback on this situation, I will update this section accordingly.
89
-
90
- In the meantime, any questions regarding licensing, and in particular how these two licenses might interact, should be directed to the original model repository: [Eric Hartford's WizardLM-13b-V1.0-Uncensored](https://huggingface.co/ehartford/WizardLM-13b-V1.0-Uncensored).
91
- <!-- licensing end -->
92
  <!-- compatibility_gguf start -->
93
  ## Compatibility
94
 
@@ -147,7 +140,7 @@ The following clients/libraries will automatically download models for you, prov
147
 
148
  ### In `text-generation-webui`
149
 
150
- Under Download Model, you can enter the model repo: TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF and below it, a specific filename to download, such as: wizardlm-13b-v1.0-uncensored.q4_K_M.gguf.
151
 
152
  Then click Download.
153
 
@@ -162,7 +155,7 @@ pip3 install huggingface-hub>=0.17.1
162
  Then you can download any individual model file to the current directory, at high speed, with a command like this:
163
 
164
  ```shell
165
- huggingface-cli download TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF wizardlm-13b-v1.0-uncensored.q4_K_M.gguf --local-dir . --local-dir-use-symlinks False
166
  ```
167
 
168
  <details>
@@ -185,7 +178,7 @@ pip3 install hf_transfer
185
  And set environment variable `HF_HUB_ENABLE_HF_TRANSFER` to `1`:
186
 
187
  ```shell
188
- HUGGINGFACE_HUB_ENABLE_HF_TRANSFER=1 huggingface-cli download TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF wizardlm-13b-v1.0-uncensored.q4_K_M.gguf --local-dir . --local-dir-use-symlinks False
189
  ```
190
 
191
  Windows CLI users: Use `set HUGGINGFACE_HUB_ENABLE_HF_TRANSFER=1` before running the download command.
@@ -198,7 +191,7 @@ Windows CLI users: Use `set HUGGINGFACE_HUB_ENABLE_HF_TRANSFER=1` before running
198
  Make sure you are using `llama.cpp` from commit [d0cee0d36d5be95a0d9088b674dbb27354107221](https://github.com/ggerganov/llama.cpp/commit/d0cee0d36d5be95a0d9088b674dbb27354107221) or later.
199
 
200
  ```shell
201
- ./main -ngl 32 -m wizardlm-13b-v1.0-uncensored.q4_K_M.gguf --color -c 4096 --temp 0.7 --repeat_penalty 1.1 -n -1 -p "A chat between a curious user and an artificial intelligence assistant. The assistant gives helpful, detailed, and polite answers to the user's questions. USER: {prompt} ASSISTANT:"
202
  ```
203
 
204
  Change `-ngl 32` to the number of layers to offload to GPU. Remove it if you don't have GPU acceleration.
@@ -238,7 +231,7 @@ CT_METAL=1 pip install ctransformers>=0.2.24 --no-binary ctransformers
238
  from ctransformers import AutoModelForCausalLM
239
 
240
  # Set gpu_layers to the number of layers to offload to GPU. Set to 0 if no GPU acceleration is available on your system.
241
- llm = AutoModelForCausalLM.from_pretrained("TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF", model_file="wizardlm-13b-v1.0-uncensored.q4_K_M.gguf", model_type="llama", gpu_layers=50)
242
 
243
  print(llm("AI is going to"))
244
  ```
@@ -289,24 +282,6 @@ And thank you again to a16z for their generous grant.
289
  <!-- original-model-card start -->
290
  # Original model card: Eric Hartford's WizardLM-13b-V1.0-Uncensored
291
 
292
-
293
- This is a retraining of https://huggingface.co/WizardLM/WizardLM-13B-V1.0 with a filtered dataset, intended to reduce refusals, avoidance, and bias.
294
-
295
- Note that LLaMA itself has inherent ethical beliefs, so there's no such thing as a "truly uncensored" model. But this model will be more compliant than WizardLM/WizardLM-7B-V1.0.
296
-
297
- Shout out to the open source AI/ML community, and everyone who helped me out.
298
-
299
- Note: An uncensored model has no guardrails. You are responsible for anything you do with the model, just as you are responsible for anything you do with any dangerous object such as a knife, gun, lighter, or car. Publishing anything this model generates is the same as publishing it yourself. You are responsible for the content you publish, and you cannot blame the model any more than you can blame the knife, gun, lighter, or car for what you do with it.
300
-
301
- Like WizardLM/WizardLM-13B-V1.0, this model is trained with Vicuna-1.1 style prompts.
302
-
303
- ```
304
- You are a helpful AI assistant.
305
-
306
- USER: <prompt>
307
- ASSISTANT:
308
- ```
309
-
310
- Thank you [chirper.ai](https://chirper.ai) for sponsoring some of my compute!
311
 
312
  <!-- original-model-card end -->
 
80
  ```
81
 
82
  <!-- prompt-template end -->
 
 
83
 
 
84
 
 
 
 
 
85
  <!-- compatibility_gguf start -->
86
  ## Compatibility
87
 
 
140
 
141
  ### In `text-generation-webui`
142
 
143
+ Under Download Model, you can enter the model repo: TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF and below it, a specific filename to download, such as: wizardlm-13b-v1.0-uncensored.Q4_K_M.gguf.
144
 
145
  Then click Download.
146
 
 
155
  Then you can download any individual model file to the current directory, at high speed, with a command like this:
156
 
157
  ```shell
158
+ huggingface-cli download TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF wizardlm-13b-v1.0-uncensored.Q4_K_M.gguf --local-dir . --local-dir-use-symlinks False
159
  ```
160
 
161
  <details>
 
178
  And set environment variable `HF_HUB_ENABLE_HF_TRANSFER` to `1`:
179
 
180
  ```shell
181
+ HUGGINGFACE_HUB_ENABLE_HF_TRANSFER=1 huggingface-cli download TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF wizardlm-13b-v1.0-uncensored.Q4_K_M.gguf --local-dir . --local-dir-use-symlinks False
182
  ```
183
 
184
  Windows CLI users: Use `set HUGGINGFACE_HUB_ENABLE_HF_TRANSFER=1` before running the download command.
 
191
  Make sure you are using `llama.cpp` from commit [d0cee0d36d5be95a0d9088b674dbb27354107221](https://github.com/ggerganov/llama.cpp/commit/d0cee0d36d5be95a0d9088b674dbb27354107221) or later.
192
 
193
  ```shell
194
+ ./main -ngl 32 -m wizardlm-13b-v1.0-uncensored.Q4_K_M.gguf --color -c 4096 --temp 0.7 --repeat_penalty 1.1 -n -1 -p "A chat between a curious user and an artificial intelligence assistant. The assistant gives helpful, detailed, and polite answers to the user's questions. USER: {prompt} ASSISTANT:"
195
  ```
196
 
197
  Change `-ngl 32` to the number of layers to offload to GPU. Remove it if you don't have GPU acceleration.
 
231
  from ctransformers import AutoModelForCausalLM
232
 
233
  # Set gpu_layers to the number of layers to offload to GPU. Set to 0 if no GPU acceleration is available on your system.
234
+ llm = AutoModelForCausalLM.from_pretrained("TheBloke/WizardLM-13B-V1.0-Uncensored-GGUF", model_file="wizardlm-13b-v1.0-uncensored.Q4_K_M.gguf", model_type="llama", gpu_layers=50)
235
 
236
  print(llm("AI is going to"))
237
  ```
 
282
  <!-- original-model-card start -->
283
  # Original model card: Eric Hartford's WizardLM-13b-V1.0-Uncensored
284
 
285
+ No original model card was available.
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
286
 
287
  <!-- original-model-card end -->