helenk commited on
Commit
41a7b77
·
verified ·
1 Parent(s): d5b2fd3

Upload E4B Q4_K_M GGUF (2026-05-09 retrain) — llama.cpp convert_hf_to_gguf.py + llama-quantize Q4_K_M from peft-merged fp16 safetensors. 5.0 GB; loads in Ollama via Modelfile (FROM gemma-4-E4B-offlineaid-Q4_K_M.gguf, RENDERER gemma4). Tier A held-out: format-OK 70.3% with RAG (vs stock+RAG 53.2%); AR format 21.6%->54.1% (2.5x).

Browse files
Files changed (1) hide show
  1. gemma-4-E4B-offlineaid-Q4_K_M.gguf +2 -2
gemma-4-E4B-offlineaid-Q4_K_M.gguf CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:31e14f6c39eed42a9c1c12a137c45455d47e41bdafaaa8a2cc94e716e0e4d935
3
- size 5302271712
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:5816bb6a011a1e7bdf8404ba1a07673b72296cddaaa114be147c95225718ac5f
3
+ size 5302257888