A multimodal model that accepts both text and image inputs, producing text output. Published by orcarouter as an uncensored variant, it operates without the content restrictions typically found in standard releases, which means it responds more freely to a wider range of prompts. This openness comes with the trade-off that guardrails are removed, so outputs require more careful downstream filtering.