A detection score measures how predictable writing is — not who wrote it. Language models pick the most probable next word over and over, which produces even sentence lengths, borrowed transitions, and smoothed-out specifics. Detectors, ours included, look for exactly that flatness.
That's also why every detector can be wrong in both directions. Careful human writing — an exam essay, a second-language writer, a well-edited memo — often reads as "too even". A number alone can't tell you which case you're in. The sentences can.