A quantized vision-language model that combines image and text understanding in an open-weight package. Built on Qwen3's 27B architecture and compressed to NVFP4 precision by bottlecapai, it trades some numerical fidelity for reduced memory footprint — a common characteristic of FP4 quantization. Specific behavioral traits and capability details beyond its multimodal input support are not well documented.