A compact, openly licensed model from Google's Gemma 4 family, quantized to 6-bit precision for efficient local deployment via MLX. The reduced bit-width means it runs well on Apple Silicon hardware with lower memory overhead, though some precision is traded off compared to full-weight versions. It handles text-in, text-out tasks and suits developers who want a capable open-weight model running entirely on-device.