A compact text model that punches above its weight class given its 2B parameter count, offering a surprisingly long 131K token context window for a model this small. It handles text-in, text-out tasks with the efficiency you'd expect from a miniaturized architecture. The MLX format suggests it's optimized for Apple Silicon environments, making it a practical choice for on-device or local inference workflows.