A lean, efficient text model that punches above its weight through sparse activation — only 3 billion parameters fire at once despite the 30B total parameter count, keeping inference costs low. The NVFP4 quantization further trims memory footprint, making it practical to run on constrained hardware. It handles an unusually large context window of over one million tokens, though its text-only focus keeps its scope narrow.