If you're building content moderation or authenticity verification systems, Pangram 4 demonstrates that modern classifiers can reliably detect AI-generated text across diverse domains while handling real-world challenges like mixed authorship and adversarial manipulation.
Pangram 4 is an AI-text detection model that identifies whether text was written by AI or humans. It achieves 99.16% accuracy with very low error rates, and can detect mixed human-AI writing and subtle edits better than previous versions. The model also handles out-of-distribution data and adversarial attacks more robustly.