humanflow

You wrote it. A detector says you didn’t.

This page is not reassurance. It is what the score actually is, what evidence rebuts it, what to say in the meeting, and which institutions have already stopped trusting these tools — so that if you are in this position you are arguing from facts rather than from indignation.

Last reviewed 6 August 2026 · The HumanFlow team

First, the thing to do today

Before anything else: preserve your draft history. Do not tidy the document, do not delete old versions, do not start a clean copy, and do not rewrite the submitted text. If you wrote this essay, the record of you writing it already exists, and it is worth more than any argument you can make about detector accuracy.

  • Google Docs — File → Version history → See version history. It keeps a fine-grained record you can scrub through. Note that browsing earlier versions requires edit permission on the file, so make sure you still hold it.
  • Microsoft Word — AutoSave revisions in OneDrive or SharePoint, plus any emailed drafts.
  • Everything around the document — notes, outlines, photographs of handwritten planning, library loan records, browser history for the sources you read, messages where you talked about the assignment.

Timestamps spread over days are the point. They describe a process. A detection score describes a finished artefact and cannot speak to how it came to exist.

What the number actually is

A detector never looks for AI. It measures two surface properties — how predictable each word is given the words before it, and how much your sentence lengths vary — and converts them into a percentage. Language models produce predictable words and even rhythm, so they score one way. So does a careful writer working in a second language. So does anyone taught to write a five-paragraph essay properly. The score cannot tell those apart, because it is not measuring authorship. Our breakdown of the mechanism goes through it in full.

This is not a fringe criticism. In the largest peer-reviewed test of these tools, Weber-Wulff and colleagues examined fourteen detectors and concluded they “are neither accurate nor reliable” — every one scored below 80% accuracy, and on machine-paraphrased text overall accuracy was 26%. Perkins and colleagues measured seven detectors at 39.5% accuracy on unaltered AI text, dropping to 22.2% once the text had been lightly altered.

And the bias is documented and specific. Liang and colleagues found that seven detectors falsely flagged 61.22% of human-written TOEFL essays while judging essays by native- speaking US eighth-graders almost perfectly. If English is not your first language, this is not bad luck. It is a known property of the method, and you may say so.

The vendor’s own numbers help you

Turnitin claims 98% accuracy — but only for documents where more than 20% of the text is flagged, and it hides scores between 1% and 19% behind an asterisk because its own validation showed that range is unreliable. If your flagged percentage is low, the company that built the tool has already said the number should not be printed, let alone acted on.

Turnitin’s chief product officer, Annie Chechitelli, also told BestColleges in April 2023 that the system is deliberately tuned to miss: “we are estimating that we find about 85% of it. We let probably 15% go by in order to reduce our false positives to less than 1 percent.” That cuts both ways in a hearing, and the half that matters to you is this — the vendor designed the tool around the knowledge that false accusations are the worse error.

You are not the first, and institutions have noticed

Vanderbilt disabled Turnitin’s AI detector in August 2023 and published the arithmetic: 75,000 papers submitted in 2022, and at Turnitin’s own claimed error rate, “around 750 student papers could have been incorrectly labeled.” Its conclusion was that it did “not believe that AI detection software is an effective tool that should be used.”

Since then: Yale, Georgetown, the University of Pittsburgh, Johns Hopkins, the University of Alabama and Curtin University have all turned the feature off. Washington State University cancelled its Turnitin AI detection contract outright in February 2026, and its provost’s memo disclosed the reason plainly: between 2023 and 2025, a third of its academic-integrity hearings involving AI allegations ended in a finding of not responsible, because a detector score had been submitted as the only evidence.

Others never adopted it. UC Berkeley ran a pilot and opted out. Syracuse declined to license it. NYU’s provost’s office states that it does not “believe any current AI detectors work well enough to recommend their use.” Cambridge “does not encourage the use of AI detection software given their proven inaccuracies and unreliability.” Monash has approved no detector at all.

Ask which camp your institution is in, and ask for the policy in writing. If your department is running a tool the university has not approved, that is relevant to your case.

What to say, and what not to say

  • Ask for the specific allegation and the evidence behind it. A percentage is not an allegation. What passage, and what supports the claim beyond the score?
  • Offer your process, not your protest. Draft history, notes, timeline. Offer to talk through the argument of the essay — someone who wrote it can.
  • Ask whether the score alone is sufficient under policy. Increasingly it is not, and asking makes the standard explicit.
  • Do not admit to something you did not do to end the meeting. This is the single most costly mistake, and it is common, because these meetings are frightening and an admission feels like the fastest way out. It closes the appeal routes your evidence would have won.
  • Do not alter the submitted text. Rewriting or running the work through a rewriting tool after an accusation reads as tampering, whatever your intention.
  • Ask what support you are entitled to. Most institutions have a student union, advocate or ombudsperson who does this regularly. Use them.

Where you are studying changes the route

The evidence above travels anywhere. The escalation route does not — it is set by national rules, and it is the part most guidance written for a US audience leaves out.

Where we stand, since we sell one of these tools

HumanFlow makes a humanizer and a detector, so treat this section as an interested party speaking. We will not tell you that our detector — or anyone’s — can prove who wrote a document. It cannot. No detector can, and the peer-reviewed record above is the reason we do not make that claim anywhere on this site.

What a detector is genuinely useful for is seeing your own writing the way an institution’s tool might see it, before you submit. That is a different job from proving authorship, and it is the only one we think the technology honestly does. If you are already under investigation, the useful artefact is your draft history, not another score.

Before arguing about the number

Establish what it measures. An asterisk is Turnitin withholding a figure it does not trust, 20% is where it starts displaying one rather than a limit, and a GPTZero probability is confidence in its own verdict rather than a share of your document — what a score actually means.

What the tool will not read

A missing score is often a file limit rather than a verdict. Turnitin’s AI report needs at least 300 words of prose and does not accept slides at all — see PDFs and PowerPoint.

Related reading

Worth having straight before a meeting: academic misconduct is a finding reached through a process, not a score. Being flagged is step one of five, and the vendor itself states the indicator is not intended as the sole basis for a finding.

The institutions that stopped using it, with sources: which universities disabled Turnitin's AI detector?

If a written decision has already been issued, the next page is appealing an AI detection finding — the stages, the evidence pack, and the deadline that ends more appeals than any argument about detectors.

If the report flagged hidden or unusual characters rather than a score, that is a different kind of finding — invisible characters are an observable property of the file, they usually arrive by copy-paste, and the true explanation is normally the best one.

Before the meeting, it helps to know what the document in front of you actually is: an originality report bundles separate checks from unrelated systems onto one page, and three of its rows point at evidence you can open while the AI row does not.

Common questions

Why is my essay flagged as AI when I wrote it?
Because the detector never looked for AI. It measured how predictable your word choices are and how uniform your sentence lengths are, then converted that into a percentage. Writing that is plain, evenly paced and structurally conventional scores the same way machine output does. That describes a great deal of honest student writing — particularly from non-native English speakers and from anyone taught a rigid essay structure.
Can I prove I wrote my essay myself?
Usually yes, and the proof is process rather than text. Version history in Google Docs or Word, saved drafts, notes, outlines, browser history and the timestamps on all of it show the work developing over hours or days. A detector score is a single number produced after the fact; a revision history is a record of the work happening. Bring the second, not an argument about the first.
Is a Turnitin AI score enough to fail me?
It should not be, and a growing number of institutions say so in writing. Michigan State's guidance calls detector outputs "potential indicators—not conclusive evidence" that "should never serve as the sole basis for academic or grading decisions." Washington State University reported that a third of its AI-related integrity hearings between 2023 and 2025 ended in a finding of not responsible precisely because a detector score had been submitted without other evidence.
What should I say when I'm accused?
Ask what the specific allegation is and what evidence supports it beyond the score. Offer your draft history. Do not admit to something you did not do to make the meeting end — an admission closes off the appeal routes that the evidence would otherwise win. Ask for the institution's policy on AI-detection evidence in writing, and ask whether the detector is even approved for use there.
Do universities still use AI detectors?
Many have stopped. Vanderbilt disabled Turnitin's AI indicator in 2023; Yale, Georgetown, Pittsburgh, Johns Hopkins and Curtin have all switched it off; Washington State cancelled its contract outright in February 2026. UC Berkeley piloted the tool and opted out, and NYU's provost's office says it does not believe any current detector works well enough to license. Others still use it. Which camp your institution is in is a fact you are entitled to ask about.
Does running my work through a humanizer help if I've been accused?
No — and it can make things considerably worse. If you are already under investigation, altering the submitted text looks like tampering regardless of why you did it. The thing that rebuts a false accusation is evidence that you wrote it, and that evidence already exists in your draft history. Rewriting after the fact destroys the strongest thing you have.

Sources

  1. 1.Testing of detection tools for AI-generated text Weber-Wulff et al. — International Journal for Educational Integrity 19:26, 2023
  2. 2.Simple techniques to bypass GenAI text detectors: implications for inclusive education Perkins, Roe, Vu, Postma, Hickerson, McGaughran & Khuat — Int. J. of Educational Technology in Higher Education 21:53, 2024
  3. 3.GPT detectors are biased against non-native English writers Liang, Yuksekgonul, Mao, Wu & Zou — Patterns (Cell Press), 2023
  4. 4.Guidance on AI detection and why we're disabling Turnitin's AI detector Vanderbilt University, Brightspace blog, 2023
  5. 5.Cancellation of Turnitin AI Detection Software (memo to instructors) Washington State University, Office of the Provost, 2026
  6. 6.Encouraging academic integrity University of Pittsburgh, Teaching Center, 2026
  7. 7.Turnitin — a note on detection tools Georgetown University, University Information Services, 2023
  8. 8.Availability of Turnitin's artificial intelligence detection UC Berkeley, Research, Teaching & Learning, 2025
  9. 9.Artificial intelligence and education — FAQ University of Cambridge, Blended Learning Service, 2026
  10. 10.We tested Turnitin's new AI detector (Annie Chechitelli on the 85%/15% trade-off) BestColleges, 2023
  11. 11.Turnitin's AI writing detection capabilities FAQs Turnitin Guides, 2026
  12. 12.AI detection tools falsely accuse international students of cheating The Markup, 2023