rodrigoramosrs commited on
Commit
c363cea
·
verified ·
1 Parent(s): 18b298e

Upload README.md with huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +1 -1
README.md CHANGED
@@ -56,7 +56,7 @@ Quantized by [Rodrigo Ramos](https://github.com/rodrigoramosrs).
56
 
57
  ## Quantization Approach
58
 
59
- All quants were produced with [llama.cpp](https://github.com/ggml-org/llama.cpp) using a **code-specialized importance matrix (imatrix)**. Unlike generic imatrix datasets, this one was curated from software engineering corpora — repository-level code, patches, test suites, and agentic coding traces — ensuring that quantization preserves fidelity on the distributions that matter most for coding tasks.
60
 
61
  The result is a set of GGUF files that retain the original model's strong software-engineering capabilities while being deployable via `llama.cpp`, `llama-cpp-python`, `Ollama`, `LM Studio`, and other GGUF-compatible runtimes.
62
 
 
56
 
57
  ## Quantization Approach
58
 
59
+ All quants were produced with [llama.cpp](https://github.com/ggml-org/llama.cpp) using a **code-specialized importance matrix (imatrix)**. Unlike generic imatrix datasets, this one was curated from software engineering corpora: repository-level code, patches, test suites, and agentic coding traces, ensuring that quantization preserves fidelity on the distributions that matter most for coding tasks.
60
 
61
  The result is a set of GGUF files that retain the original model's strong software-engineering capabilities while being deployable via `llama.cpp`, `llama-cpp-python`, `Ollama`, `LM Studio`, and other GGUF-compatible runtimes.
62