humanflow

AI detection

AI detectors are widely used and widely misunderstood. These pages explain what the scores actually measure, what the research says about how often they are wrong, and what to do if you have been accused on the strength of one.

We build a humanizer and a detector, so we have an obvious interest here. We have tried to write these pages so that someone who never installs our app still leaves with a correct understanding — and we say plainly where the evidence runs out.

The short version

  • A detection score measures how predictable text is, not who wrote it.
  • False positives are common and fall hardest on people writing in a second language.
  • Different detectors regularly disagree about identical text, because each picks its own threshold.
  • Draft history — Google Docs version history, Word AutoSave — is the strongest response to an accusation.
  • No tool can promise a particular score from a particular detector. Anything claiming otherwise is selling you something.

Pages in this section last reviewed 27 July 2026. Want the tool? See what HumanFlow does.