A heavily compressed open-weight model running at 2-bit quantization via the MLX framework, meaning it trades raw precision for dramatic size reduction. This makes it unusually memory-efficient for a 27B parameter model, though aggressive quantization at this level typically introduces noticeable quality degradation compared to higher-bit variants. It fits a niche where fitting large parameter counts into constrained hardware matters more than peak accuracy.