Detect tower · floor

AI detector false positive: what to do next

An ai detector false positive is a machine saying a person wrote something they did not write — except backwards: it says a person did not write something they did. If that has just happened to you, the most useful thing to know is that the companies selling these tools agree with you about what their number proves.

Turnitin
“Turnitin does not make a determination of misconduct… rather, we provide data for educators to make an informed decision.”

And to instructors directly: “use the information to initiate a conversation, not to draw a conclusion.” Read 21 August 2026.

GPTZero
“There always exist edge cases with both instances where AI is classified as human, and human is classified as AI.”

Its own advice to teachers is to ask students to demonstrate understanding in a controlled environment, “or through an editor that can track their edit history”. Read 21 August 2026.

Originality AI
Sells a “Writing Replay” offering “peace of mind for writers and students to show the authorship of writing”.

A separate product, sold alongside the detector, for the job the detector cannot do. Read 21 August 2026.

None of those quotations comes from a critic. All three are from the companies that build and sell detection, read on their own sites and dated. Between them they say: the score is data, not a determination; errors happen in both directions; and the way to establish authorship is to show the process.

What “less than 1%” looks like from the other side

Essays marked across a department in one term2,000
Written honestly, with no AI involved2,000
Flagged anyway, at the rate the vendors publishup to 20
And the chance it is right about your essay in particularunknown

The words people type when this happens are unusually specific, and they say everything about the situation. Falsely accused of using ai. How to prove you didn’t use ai. Turnitin false positive. How accurate are ai detectors — asked after the accusation, not before. Nobody searches those phrases idly; they are typed by somebody with an email open in another tab.

The arithmetic is the part nobody does out loud. A false positive rate of one per cent sounds like precision. Across two thousand honest essays it means up to twenty people are asked to defend themselves — and it says nothing at all about which twenty.

Where to start

Five ways in. The first two are for right now.

“I have been accused and I wrote it myself.”
Start at what to do now
“I have to reply to an email today.”
Go to what to do now
“What evidence actually helps?”
That is the evidence that works
“Why me? I wrote it normally.”
There are documented patterns — who gets flagged
“I am the instructor and I am not sure.”
Read for the person deciding

What to do now

In order, on the day it happens. Nothing here requires a lawyer, and the first three steps cost nothing but care.

Do not panic-rewriteChanging the file now destroys the strongest evidence you have. Leave it alone.Being built
Gather the processVersion history, drafts, notes, search history, the library book — assembled before you reply.Being built
The replyWhat to write, what to leave out, and why a calm factual message beats an indignant one.Being built
Ask what the evidence isYou are entitled to know which tool, which score, and what it is being used for.Being built
If it escalatesInstitutional procedures, who to bring, and the point at which you stop replying informally.Being built

The evidence that works

Every detector company points at the same thing, and it is not their own product. This wing is that answer, in practical form.

Version historyWhat Docs and Word keep automatically, how to show it, and why it is nearly unanswerable.Being built
Writing where the process is recordedThe habit that settles this before it starts, recommended by the detector companies themselves.Being built
Explaining your own textBeing able to talk about the argument, the sources and the choices — the oldest test there is.Being built
What does not helpRunning the text through another detector, which produces a different number and no more proof.Being built

Who gets flagged

The patterns are documented and they are not random. Knowing them explains a lot, and matters for anybody setting policy.

Plain, careful proseClear structure and unadorned sentences look statistically like generated text.Being built
Second-language writersThe documented concern that the same writing is flagged more often, and what follows from it.Being built
Heavily edited workWriting polished with a grammar checker moves toward the same statistical middle.Being built
Technical and formulaic writingMethod sections, legal language and anything that is supposed to read like everybody else’s.Being built

For the person deciding

Written for the instructor or the editor holding the report. They did not build the tool, and the vendor has explicitly declined to make the judgement.

What the vendors ask of youTurnitin’s own guidance: “assume positive intent”, and give “the strong benefit of the doubt”.Being built
Say it in advanceTurnitin advises acknowledging that false positives happen before any submission, because not doing so makes the conversation worse.Being built
A score is not a thresholdWhy an automatic rule above a percentage contradicts every published statement from the vendors.Being built
Better assessment designThe changes that make detection largely unnecessary, which is where this argument actually ends.Being built

What this tower will not do

It will not tell you detectors are useless. They are measuring something real, they are right far more often than they are wrong, and the companies publishing their error rates deserve more credit than the ones that publish nothing.

It will not help anybody who did use AI and is looking for a script. The advice on this floor is the process evidence you either have or do not have, and it cannot be assembled after the fact.

And it will not pretend this is a small thing. A percentage produced by a company that says it is not a determination is being used, in practice, to determine outcomes — and the person on the receiving end usually has no idea that the vendor agrees with them. What holds instead is simple: every statement quoted here comes from a company that sells detection, not from a critic.

Where this page got its facts

  1. Turnitin on false positives — that it does not determine misconduct, and its advice to assume positive intent — www.turnitin.com, read 21 August 2026.
  2. Turnitin on the sentence-level false positive rate of around 4% — www.turnitin.com, read 21 August 2026.
  3. GPTZero’s FAQ — edge cases in both directions, and its advice about edit history — gptzero.me, read 21 August 2026.
  4. Originality AI’s own page — the Writing Replay sold for showing authorship — originality.ai, read 21 August 2026.

Written by Alberto Gulotta

Founder and editor of AI Tools Primer, writing from Palermo, Italy. Thirty-five years of taking computers apart, starting with a Commodore 64 — the long version is on the about page.

Something wrong on this page? Write to aitoolsprimer@gmail.com and it gets fixed.

Independence and limits

No affiliate links and no paid placements anywhere on this site. Nobody pays to appear here, and no company has seen this page before you did.

This is general information, not professional advice. Where a page touches money, health, safety or the law, it names its source and the date it was read — and your situation may still differ. See the privacy page and the cookie policy.