Perplexity/burstiness detectors (e.g. GPTZero)
Transparent and fast, but the easiest to move with rewriting — and the most prone to false-flagging plain human writing. A signal, not a verdict.
Turnitin
Claims under 1% false positives; a Washington Post test found up to 50% on a small sample, and it misses ~15% of AI text by design. Widely used, widely contested.
Originality.ai / Copyleaks / Winston
Commercial classifiers tuned for high AI-catch rates, which tends to raise false positives. Accuracy claims are self-reported; independent results vary by text type.
Watermarking
Robust in theory, but requires the AI vendor to embed it, and paraphrasing degrades it. Most text people run was never watermarked, so it rarely applies.
Paraphrasers & "humanizers"
Lower scores by raising perplexity/burstiness. Effective today, unstable tomorrow, and increasingly detected as their own category. They also can dent clarity — the honest use is improving writing, not laundering authorship.
Retrieval / provenance defenses
The strongest emerging counter: match text against a store of known AI outputs. Catches paraphrased text well, but only the AI provider can run it — third-party detectors cannot.