The most common reason is the guarantee not surviving contact with an institution. A tool that promised 0% and produced a flag has not just failed — it has left you having relied on it, which is a worse position than never having been promised anything.
The second is the missing per-request cap. Word allowances are stated by the month and the longest single passage is not, so people working on long documents discover the ceiling by hitting it.
The third is scope. It is a rewriting bundle rather than an academic-integrity tool, so there is nothing on the page about disclosure, institutional policy or what to do if you are questioned about your work.
None of that makes GPTinf a bad tool. It makes it a tool with a shape, and the question is whether that shape matches your work. Our full HumanFlow vs GPTinf comparison sets its numbers against ours in detail, including where GPTinf wins.