A compact, open-weight model from Google's Gemma 4 family, quantized to 5-bit precision and packaged by lmstudio-community for local MLX inference on Apple Silicon. The quantization keeps memory footprint manageable while preserving much of the original model's capability, though some precision loss is inherent to the format. It handles text-in, text-out tasks and is straightforward to run locally without cloud dependencies.