This is a decensored version of google/gemma-3-4b-it-qat-q4_0-unquantized, made using Heretic v1.1.0

Abliteration parameters

Parameter Value
direction_index per layer
attn.o_proj.max_weight 1.27
attn.o_proj.max_weight_position 32.22
attn.o_proj.min_weight 1.04
attn.o_proj.min_weight_distance 18.63
mlp.down_proj.max_weight 1.29
mlp.down_proj.max_weight_position 20.41
mlp.down_proj.min_weight 0.92
mlp.down_proj.min_weight_distance 9.75

Performance

Metric This model Original model (google/gemma-3-4b-it-qat-q4_0-unquantized)
KL divergence 0.4659 0 (by definition)
Refusals 3/100 97/100

Gemma 3 model card

Model Page: Gemma

This repository corresponds to the 4B instruction-tuned version of the Gemma 3 model using Quantization Aware Training (QAT).

The checkpoint in this repository is unquantized, please make sure to quantize with Q4_0 with your favorite tool

Thanks to QAT, the model is able to preserve similar quality as bfloat16 while significantly reducing the memory requirements to load the model.

Downloads last month
31
Safetensors
Model size
4B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for addansee2/gemma-3-4b-it-qat-q4_0-unquantized-heretic

Finetuned
(12)
this model
Quantizations
3 models
MiniMax H3 Video Generator 20 free credits · Text & image to video Try Free →