How to use from
Ollama
ollama run hf.co/tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v3:Q4_K_M
Quick Links

Gemma-4 12B Coder โ€” SFT v3 (deprecated)

โš ๏ธ Deprecated โ€” do not use for new work. intermediate SFT iteration superseded by v5.

Replaced by tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v5-GGUF.

gemma-4 12B coder for local, agentic tool use โ€” GGUF quantizations for llama.cpp / Ollama.

Run it: llama-server -hf tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v3:Q4_K_M --jinja (full commands below).

At a glance

Type GGUF quantizations ยท llama.cpp / Ollama
Techniques โ€”
Tool-calling native --jinja
Status โš ๏ธ Deprecated โ†’ tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v5-GGUF
Use llama-server -hf tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v3:Q4_K_M --jinja

Use it

# llama.cpp (server) โ€” tool-calling needs the recovery shim, see below
llama-server -hf tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v3:Q4_K_M --jinja --ctx-size 16384

# Ollama
ollama run hf.co/tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v3:Q4_K_M

Files

Sizes and a one-click loader are in the file browser / Quantizations widget above; the note says which quant to reach for.

Quant Notes
Q4_K_M good default โ€” fits 12 GB VRAM, best size/quality balance

Intended use & limitations

Built for code generation and agentic tool use; serve locally via llama.cpp / Ollama, or use as a base to fine-tune / merge / quantize. Outputs can be wrong or fabricated โ€” validate tool arguments before executing, and keep a human in the loop for anything consequential.

Where this sits in the family


Provenance & reproduction

How this model was built โ€” technique chain, training mix, and the exact knobs/pins, so the result is reproducible without any of our tooling.

Mechanics applied

Step Technique What it does Provenance

Part of the Gemma-4 12B Coder โ€” archive (superseded) collection.

Something not right, or a request? Open a discussion โ€” happy to help.

Downloads last month
95
Safetensors
Model size
12B params
Tensor type
BF16
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v3

Collection including tpls/gemma-4-12B-coder-fable5-composer2.5-v1-sft-v3