A coding-focused model built on Qwen3's architecture, distilled down to a 35B parameter count with only 3B active parameters via mixture-of-experts routing — meaning it punches above its weight in compute efficiency. The 'abliterated' variant has had refusal behaviors removed, making it more permissive in what it will generate. It handles a very large context window of 262K tokens, useful for working across large codebases.