Technology

Built to catch AI text.Honest about what it can’t.

The figures below come from our own testing on thousands of human and AI-generated texts, including paraphrased, edited, and disguised content. We also benchmark against the public RAID benchmark, an open dataset built to test detectors against exactly those tricks. Here’s what it actually does.

99.2%
Accuracy, internal testing
<2s
Analysis time

Where you can trust it

Real strengths, in plain language.

Catches modern AI tools

Trained on text from ChatGPT, Claude, Gemini, Llama, Mistral, and other widely used AI assistants. On text from today's top commercial models, it's right close to 100% of the time.

Sees through common dodges

Not fooled by look-alike characters (Cyrillic 'а' swapped for Latin 'a'), deliberate typos, weird spacing, or article deletions. These tricks have near-zero effect on the score.

Honest about how confident it is

When the score says 'very confident', it means it. A 90% score really means about 90% likely AI, not 50% rounded up, not 99% inflated. The percentage you see matches reality.

Strict mode for high-stakes decisions

If wrongly flagging a human is the worst outcome, a tighter setting further reduces the chance of that happening, at the cost of occasionally missing borderline AI text.

When to double-check

What our detector struggles with. Every AI detector has limits, and we'd rather tell you ours.

Heavily paraphrased AI

If someone takes AI output and substantially rewrites it in their own words, we catch about 8 in 10 of those, and miss roughly 1 in 5. This is the hardest case in AI detection, and no tool solves it cleanly today.

Very short text

The minimum is 50 characters, but short text is where we are weakest. Single sentences don't give the model enough to commit to a verdict, and we'll tell you when there's too little signal.

Occasional human false positives

A small share of genuinely human-written texts can still be flagged as AI. For decisions that affect someone's grade, job, or reputation, treat our score as evidence, not a verdict. Always combine with human review.

English only

Trained and benchmarked on English. We don't make claims about other languages until we've tested them properly.

How it works

Our detector reads your entire text in context, not sentence by sentence in isolation, but together, the way a careful reviewer would. It looks at patterns most AI assistants share: word choices, sentence rhythm, transitions, and the small stylistic tells they leave behind.

The output is a single probability from 0 to 100% plus per-sentence highlights, so you can see where the signal is coming from instead of just an opaque verdict.

Our Principles

The values that guide how we build this tool.

Accuracy First

We prioritize not flagging human writing by mistake. Wrongly labelling someone's own work as AI can have serious consequences, so we would rather miss a borderline case than accuse the wrong person.

Privacy by Design

Submitted texts are not stored permanently or used for training. We process content for analysis and return results. Your writing stays yours.

Honest Scoring

AI detection is probabilistic, not absolute. We calibrate scores so the percentage you see matches the real probability, and we publish what the tool can and can't do.

Continuously Validated

We regularly benchmark our models against established datasets and real-world adversarial examples. As AI writing evolves, our detection keeps pace.

About the RAID benchmark

RAID (Robust AI Detection) is an independent public benchmark that pits AI detectors against thousands of real human and AI-generated texts, plus a dozen-plus ways people try to evade detection: paraphrasing, character swaps, deliberate typos, and more.

We run our model on RAID regularly and publish what it scores. We don’t grade our own homework.

See it in action

Paste any text and get an instant AI detection analysis. Free, no account needed.

Sentence-level highlights included.