A Japanese-language focused model from llm-jp that incorporates a thinking or reasoning step before producing responses. With 32 billion total parameters but only 3 billion active at a time, it uses a sparse mixture-of-experts architecture that keeps inference costs lower than its full parameter count suggests. Its open-weight nature means it can be run locally and inspected freely.