An AI detector does not know who wrote anything. It has no record of authorship, no browser telemetry and no access to the document's history. It reads the prose and estimates how statistically typical of machine writing it looks — mostly through perplexity, how predictable each next word is, and burstiness, how much sentence length and structure vary.
Human writing tends to be uneven. Model writing tends toward smooth and medium-everything. That difference is the whole mechanism, and it is why the tools struggle with the specific kinds of human writing that happen to be even and conventional.
So a score of 80% does not mean an 80% chance the student used AI. It means the text carries features the model associates with machine writing. Those are not the same claim, and the gap between them is where nearly every unfair accusation lives.