ICEPVP8977/Uncensored_Small_Reasoning
Viewer • Updated • 4.54k • 29 • 7
How to use Mote1001/llama-8b-lora-uncensored-thinking-F16-GGUF with PEFT:
Task type is invalid.
This LoRA adapter was converted to GGUF format from vpakarinen/llama-8b-lora-uncensored-thinking via the ggml.ai's GGUF-my-lora space.
Refer to the original adapter repository for more details.
# with cli
llama-cli -m base_model.gguf --lora llama-8b-lora-uncensored-thinking-f16.gguf (...other args)
# with server
llama-server -m base_model.gguf --lora llama-8b-lora-uncensored-thinking-f16.gguf (...other args)
To know more about LoRA usage with llama.cpp server, refer to the llama.cpp server documentation.
16-bit
Base model
meta-llama/Llama-3.1-8B