A multimodal model that handles both text and image inputs, built on the Qwen3.8 architecture and quantized to FP8 precision for efficient deployment. The FP8 format reduces memory footprint compared to full-precision weights, which can affect output fidelity at the margins. Published by lribeiro under an open Apache 2.0 license, making it freely usable and modifiable.