Gemma 4 QAT (Quantization-Aware Training) for 3x less memory use and near original accuracy.
Note GGUFs to run:
Note BF16 Safetensor: