gemma-3-1b-it-Abliterated-GGUF

GGUF quantizations of gemma-3-1b-it-Abliterated.

The base model was abliterated using AnlordAbliterator 1.4.0 and then converted to GGUF and quantized into multiple formats.

Available Quantizations

Quantization File
BF16 gemma-3-1b-it-abliterated-bf16.gguf
F16 gemma-3-1b-it-abliterated-f16.gguf
Q8_0 gemma-3-1b-it-abliterated-q8_0.gguf
Q6_K gemma-3-1b-it-abliterated-q6_k.gguf
Q5_K_M gemma-3-1b-it-abliterated-q5_k_m.gguf
Q5_0 gemma-3-1b-it-abliterated-q5_0.gguf
Q4_K_M gemma-3-1b-it-abliterated-q4_k_m.gguf
Q4_0 gemma-3-1b-it-abliterated-q4_0.gguf

Which Quantization Should I Use?

A simple rule of thumb:

Quantization Quality Size Recommended for
BF16 โ˜…โ˜…โ˜…โ˜…โ˜… Very large Maximum precision
F16 โ˜…โ˜…โ˜…โ˜…โ˜… Large Maximum precision
Q8_0 โ˜…โ˜…โ˜…โ˜…โ˜… Large Near-original quality
Q6_K โ˜…โ˜…โ˜…โ˜…โ˜… Medium High quality
Q5_K_M โ˜…โ˜…โ˜…โ˜…โ˜† Medium Quality / size balance
Q5_0 โ˜…โ˜…โ˜…โ˜…โ˜† Medium General use
Q4_K_M โ˜…โ˜…โ˜…โ˜…โ˜† Small Recommended default
Q4_0 โ˜…โ˜…โ˜…โ˜†โ˜† Smallest Maximum memory savings

Q4_K_M is the recommended starting point for most users who want a good balance between quality and memory usage.

Base Model

google/gemma-3-1b-it

Original model:

https://huggingface.co/google/gemma-3-1b-it

Abliterated Transformers version:

https://huggingface.co/anlord/gemma-3-1b-it-Abliterated

Abliteration

The base model was processed with AnlordAbliterator 1.4.0.

Results

Model: google/gemma-3-1b-it

Initial refusals: 97 / 104
Final refusals:    5 / 104

KL divergence: 0.10103859007358551

200 optimization trials

Tool

AnlordAbliterator

Running with llama.cpp

Example:

llama-cli -m gemma-3-1b-it-abliterated-q4_k_m.gguf

The GGUF files are intended for use with GGUF-compatible software such as llama.cpp and other compatible inference applications.

License

This repository contains derivative model files based on google/gemma-3-1b-it.

The original gemma-3-1b-it model is released under the Gemma Terms of Use and is gated on Hugging Face.

Refer to the original model repository for the applicable license terms.

Disclaimer

These quantizations are derived from an abliterated version of gemma-3-1b-it.

Quantization may introduce small differences in model behavior and output quality compared with the original Safetensors model.

Downloads last month
454
GGUF
Model size
1.0B params
Architecture
gemma3
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for anlord/gemma-3-1b-it-Abliterated-GGUF

Quantized
(477)
this model