Guide

Turnitin AI Detection: What the Score Actually Means

The number in a Turnitin AI Writing Report is widely misread. Students confuse it with the similarity score, and institutions treat it as proof. This guide explains what the percentage measures, what Turnitin itself publishes about its error rates, and where the tool is documented not to work.

The Score Is Not What Most People Think

Four things about the AI Writing Report that account for most of the confusion around it.

It Is Not the Similarity Score

The AI writing percentage is separate from and independent of the similarity score, and AI highlights do not appear in the Similarity Report at all. Most questions along the lines of 'is 25% bad?' are really about the similarity score, which measures matched sources and is a completely different thing.

Scores Under 20% Are Hidden

Since July 2024, Turnitin does not surface any score above 0% and below 20%. It shows an asterisk instead. Their stated reason is that testing found a higher incidence of false positives in that band, so the number was withheld rather than shown.

Only 'Qualifying Text' Is Counted

The percentage covers prose sentences in long-form writing. Turnitin states the model does not reliably detect AI text in poetry, scripts, code, bullet points, tables or annotated bibliographies, so a mixed-format document can show a percentage that does not match its highlights.

It Needs at Least 300 Words

A submission must contain at least 300 words of prose and no more than 30,000 to generate a report at all. Short assignments, discussion posts and problem sets fall outside what the tool is built to assess.

What Turnitin Publishes About Its Own Error Rate

These are Turnitin’s figures, not ours. They are worth reading closely, because the document-level number and the sentence-level number are very different.

<1%

Document-level false positives

Turnitin's stated rate for incorrectly flagging a fully human-written document, but only for documents scored at 20% AI writing or above.

~4%

Sentence-level false positives

Turnitin's stated likelihood that any individual sentence highlighted as AI-written was in fact written by a person.

54%

Of those sit beside real AI text

Turnitin reports that more than half of falsely highlighted sentences appear directly next to genuine AI writing, most often at the transitions.

Why a small percentage still matters at scale

In August 2023 Vanderbilt University disabled Turnitin’s AI detector and explained the arithmetic behind the decision: the university submitted 75,000 papers in 2022, so a 1% false positive rate would mean roughly 750 papers wrongly flagged in a single year. They also cited a lack of transparency about how the model works, and research showing detectors are more likely to flag writing by non-native English speakers.

Turnitin’s own instructor guidance is consistent with that caution. It states the model “may not always be accurate (it may misidentify human-written, AI-generated, and AI-paraphrased text), so it should not be used as the sole basis for adverse actions against a student,” and advises using a highlighted result “to initiate a conversation, not to draw a conclusion.”

Sources: Turnitin, Using the AI Writing Report and Turnitin, Understanding AI writing detection: false positive rates.

If a Report Flags Work You Believe Is Human

Turnitin's guidance and ours agree on the shape of this: the score opens a conversation, it does not settle one.

1

Check Which Score You Are Looking At

Confirm whether the number is the AI writing percentage or the similarity score. They measure different things, appear in different reports, and a high similarity score usually means quoted or cited sources rather than AI use.

2

Look at the Highlights, Not the Number

Read the specific sentences the report marked. If the flagged passages are formulaic transitions, method descriptions or plain factual summary, that is the writing style most likely to be misread as AI.

3

Ask About the Writing Process

Version history, drafts, notes and outlines are stronger evidence of authorship than any detector score in either direction. Ask the student to talk through how the piece came together.

4

Get an Independent Reading

A second detector built on different models is a useful cross-check. Where two independent tools disagree, that disagreement is itself information, and it is a reason to slow down rather than act.

Common Questions About Turnitin AI Detection

It tells the instructor, not the student. The AI Writing Report is visible to educators at institutions that have the feature enabled; students submitting through Turnitin generally cannot see their own AI writing percentage. This is the single biggest practical difference between Turnitin and a checker you can run yourself.

This question is almost always about the similarity score rather than the AI score, and the two are entirely separate. A similarity score reflects text matching existing sources, so a well-cited essay with long quotations can score high and be perfectly legitimate. Turnitin states plainly that there is no 'right' or 'target' score for either indicator. The number on its own means very little without reading what was actually matched or highlighted.

Turnitin stops showing a figure for any score above 0% and below 20%, displaying an asterisk instead. Their documentation says testing found a higher incidence of false positives in that range, so the score was withheld to reduce misinterpretation. Reports generated before 8 July 2024 may still show a numerical score under 20%.

Yes, and Turnitin says so directly. Their published figures are a document-level false positive rate below 1% for documents scoring 20% or more AI writing, and a sentence-level false positive rate of around 4%. Their instructor guidance states the model may misidentify human-written, AI-generated and AI-paraphrased text, and should not be the sole basis for action against a student.

Turnitin states that its detection covers text that was likely AI-generated and then modified by an AI paraphraser, word spinner or bypasser tool, naming Quillbot as an example. That capability is English-only. Their Spanish and Japanese detectors do not include paraphrasing or bypasser detection.

English, Spanish and Japanese. Work submitted in any other language will not generate an AI Writing Report, and only the English detector includes AI paraphrasing and bypasser detection.

No. Turnitin is licensed by institutions, and the AI writing feature has to be enabled by an administrator, so individual teachers and students cannot simply sign up and run a check. Several unaffiliated websites use the Turnitin name to promote free checkers; those are not operated by Turnitin and will not return the same result an instructor sees.

Not on its own, according to Turnitin. Their guidance says the result requires further scrutiny and human judgment alongside the institution's own academic policies, and should not be the sole basis for adverse action. Most institutions treat the score as a prompt to investigate, through drafts, version history and a conversation about process, rather than as a finding in itself.

Different job, different access model. Turnitin is bought by an institution, bundled with plagiarism matching, and shows results to instructors only. Ours is free to run with no account, returns a score per sentence rather than one document figure, and publishes where it fails. We are not a replacement for an institutional integrity system and we do not reproduce Turnitin's result. We are an independent second reading.

Where an Independent Second Opinion Helps

Not a substitute for your institution's process, and not a way to predict its output.

Built on Different Models

Our checker uses its own detection models rather than reproducing Turnitin's. That is the point of a second reading: two tools that fail in the same way tell you nothing, two that fail differently tell you where to look.

A Score You Can Both See

Because the result is not locked behind an institutional licence, a student and a teacher can look at the same output together and talk about the specific sentences it marked.

Sentence-Level by Default

Every sentence carries its own score, so partial AI use in an otherwise original piece is visible rather than averaged away into a single document percentage.

Known Limits, Written Down

Heavily rewritten AI is the case we handle worst, and the checker is English-only. Both are documented on our technology page rather than left for you to discover.

Run an Independent Check

Paste any text and get a verdict, a confidence score, and the specific sentences behind it. Useful as a second reading alongside an institutional report, not as a prediction of one.

Free to try. No account, no card.