humanflow
AI detection · The HumanFlow team · 12 min read

QuillBot AI Detector Review: A Good Free Tool From a Company Playing Both Sides

QuillBot sells a paraphraser and an AI detector. Our review of its free limits, 99% RAID claim, and what that dual position means for you.

QuillBot's AI detector is one of the stronger free options available: 1,200 words per scan, six scans a day, sentence-level analysis, 20+ languages, and a 99% detection claim tied to the independent RAID benchmark. The complication is that QuillBot also sells a paraphraser and humanizer — the exact tools detectors exist to catch.

That complication is the honest heart of this review, and we're unusually qualified — or unusually compromised, depending on your view — to write about it. HumanFlow builds both a detector and an AI humanizer too. Same dual position, same questions to answer. So there will be no stones thrown here; instead we'll explain, as plainly as we can, what it means for you when any vendor — QuillBot, us, anyone — sits on both sides of the detection line. Our editorial policy covers how we handle that conflict; read it and judge accordingly. Everything below comes from QuillBot's live pages fetched during writing, published third-party research, or is marked for verification.

What the detector is and where it came from

QuillBot made its name as a paraphrasing tool. Long before ChatGPT, students and non-native English writers used it to rework sentences — nine rewriting modes in the current premium version — and it grew into a full writing suite: grammar checker, summarizer, plagiarism checker, AI chat, and, since the generative-AI wave, both an AI detector and an explicit "Humanizer" with what the pricing page calls a "human score." QuillBot is part of Learneo, the company formerly known as Course Hero — which also owns Scribbr, a fact worth holding onto for the section on Scribbr's comparison study below.

The detector itself does what the modern category standard demands. Paste text of at least 80 words and it returns an AI-likelihood score from 0 to 100%, with sentence-level highlighting of the passages that drove the score. It attempts something subtler than a binary verdict, too: the tool identifies mixed content — human writing that's been AI-refined, and AI output that's been human-edited — which is where most real-world documents now live and where most detectors quietly fall apart. QuillBot says the model is trained on output from current systems including GPT-5, Claude, and Gemini, and supports more than 20 languages including Spanish, French, and Portuguese. For a free tool, that language coverage is notable; most competitors handle four or five at best, and detector reliability outside English is the category's least-examined corner.

Under the hood it's statistical detection like everyone else's: the tool measures how machine-typical your text is against a threshold the vendor chose. If that sentence is doing a lot of work you don't fully trust, our explainer on how AI detectors work unpacks perplexity, burstiness, and why the threshold is a business decision rather than a scientific constant.

Free limits, premium limits, and price

Fetched from quillbot.com at review time:

FreePremium
Words per detection scan1,200Unlimited
Scans per day6Unlimited
Explainer cards (why text was flagged)LimitedFull-text explainers
Rewrite cards1Unlimited
Downloadable reportsNoYes
Paraphraser125 words, 2 modesUnlimited, 9 modes
Humanizer125 words, 6 uses/dayUnlimited, with "human score"
Price$0$8.33/mo billed annually (~$100/yr)

Two observations. First, the free detector tier is genuinely generous — 7,200 words a day of sentence-level analysis costs nothing, which beats Grammarly's free single-percentage readout and matches Scribbr's per-scan cap while adding more granular highlighting. Second, look at what Premium actually bundles: the detector's unlimited tier is sold in the same package as the unlimited humanizer. One subscription, both sides of the arms race. Keep that thought; we'll come back to it.

The 99% claim, read carefully

QuillBot's page claims a "99% detection rate, according to independent evaluations from RAID." RAID — a shared benchmark for robust evaluation of machine-generated text detectors, presented at ACL in 2024 — is the most serious independent public benchmark this category has, built specifically to test detectors against adversarial tricks like paraphrasing. Citing it is the right kind of claim: external, checkable, methodologically grounded. It clears the bar we set in our guide to reading any vendor's accuracy claim far better than the naked percentages most detector marketing prints.

Now the fine print you should supply even where the vendor doesn't. A detection rate is not an error rate: it tells you how much AI text the tool catches, not how often it wrongly flags humans, and QuillBot publishes no false-positive figure anywhere on the detector page. Benchmark conditions are not your conditions — RAID scores depend on which configuration, which generators, and which false-positive budget you evaluate at, and QuillBot's page does not say which of those produced its 99%. Grammarly cites the same benchmark for a 99% claim and a first-place ranking. One benchmark cannot crown two winners, and the paper behind it concluded that detectors claiming "99% or more" are "easily fooled by adversarial attacks". And you may notice that Grammarly's detector page also claims 99% and a #1 RAID ranking. Both claims can be simultaneously true under different benchmark dates and settings, which tells you exactly how much weight a single headline number can bear.

To QuillBot's credit, its own disclaimers are unambiguous: "No AI Detector can provide 100% accuracy," and — this one deserves framing on the wall of every academic-integrity office — "Never rely on AI detection alone to make decisions that could impact someone's career or academic standing." The page further concedes that heavily paraphrased or formulaic writing produces unreliable scores. That is honest, and honesty in this category is rarer than 99% anything.

Independent evidence: what third parties found

Here's something genuinely unusual: the best independent data point on QuillBot's detector comes from a direct competitor. Scribbr — whose own detector we review separately — publishes comparative research on twelve detectors, and in the version live at review time (last revised July 2026, per the page) QuillBot's free detector scored 78% overall accuracy: tied with Scribbr's own free tool, above Originality.ai (76%), Copyleaks (66%), ZeroGPT (64%), and GPTZero (52%), and behind only Scribbr's premium tool (84%). A competitor ranking you level with itself and ahead of most of the field is about as credible as third-party evidence gets in this business.

Scribbr's methodology is small — 30 texts across six categories — so treat the ranking as indicative rather than definitive. But it triangulates with the RAID citation to support a fair summary: QuillBot's detector is a legitimately capable tool, plausibly top-tier among free options, and nothing in the published record contradicts that.

There's a delicious wrinkle in Scribbr's test worth naming. Two of its six text categories were built by running text through QuillBot's paraphraser — and those paraphrased categories were the ones nearly every detector, including the premium winner, handled worst (Scribbr's top tool caught only 60% of paraphrased or mixed texts). The same company's product line supplied both a strong detector and the adversarial attack that defeats detectors. Which brings us to the elephant.

The paradox: selling the sword and the shield

QuillBot sells a paraphraser and humanizer that make AI text harder to detect, and a detector that tries to catch AI text — including, presumably, text processed by tools like its own. If that strikes you as a conflict of interest, you're paying attention.

We are not the ones to throw stones. HumanFlow occupies the identical dual position: we build a humanizer and a detector under one roof. So instead of performative outrage, here is what we think users of any dual-sided vendor — QuillBot, us, whoever enters next — should actually understand.

First, the structural tension is real and doesn't wash out with good intentions. A company earning subscription revenue from making text less detectable has, at minimum, a complicated incentive when tuning its detector's aggressiveness. Does the detector flag its own humanizer's output? If it catches it too well, one product undermines the other's pitch; if it misses it, the detector has a blind spot shaped exactly like the company's flagship product. We haven't tested QuillBot's detector against QuillBot's humanizer — see the hands-on section below for how we intend to — and we'd apply the same test to ourselves.

Second, there's an honest version of the dual position, and it's worth stating because it's the position we try to hold. Detection and humanization are two views of the same underlying question — what makes text read as machine-typical? — and a vendor that studies both sides genuinely understands the problem better than one that studies either alone. The legitimate use of a humanizer is revision toward your own voice, not disguise of banned work; the legitimate use of a detector is screening and conversation-starting, not verdicts. A vendor can serve both honestly if it refuses to promise evasion and refuses to promise certainty. QuillBot's disclaimers suggest it understands at least the second half. Whether any vendor lives up to that standard — including us — is exactly what published methodology and editorial policies exist to let you check.

Third, the practical takeaway for you: never treat a dual-sided vendor's detector as a neutral referee of that same vendor's humanizer, and ideally cross-check any consequential score against a detector with no rewriting product attached. Disagreement between tools is information. Our hub on detector accuracy explains why the same text scores differently across tools even without conflicts of interest — vendor-chosen thresholds guarantee it.

What we have not done

We have not run our own hands-on test of QuillBot's detector. Everything above rests on vendor claims and published research, and that is exactly how you should weight it. We would rather say so than publish a table of numbers we did not measure — and it is why the next section hands you the method instead of asking you to trust ours.

Who should use it — and how

Students and writers wanting a capable free pre-submission check are well served: six 1,200-word scans a day covers most needs, and the sentence-level highlighting tells you where the machine-typical passages are, which a bare percentage can't. Non-English writers get rare multi-language coverage, with the caveat that independent evidence on non-English detection accuracy is thin everywhere. Editors triaging freelance submissions get a fast, free screen — as one signal among several, never a verdict, per QuillBot's own warning.

Teachers and institutions should hear QuillBot's disclaimer at full volume: never as sole evidence for decisions affecting careers or academic standing. The base-rate arithmetic makes the point without any appeal to authority. Screen 10,000 honest documents with a 99%-specific detector and you still flag roughly 100 innocent writers; if only one submission in ten is actually AI, even a 95%-sensitive, 99%-specific tool produces about one false accusation for every ten true catches. The best detector in the world can't outrun that math, and the ones bearing the cost are disproportionately non-native speakers and rigid-structure writers — the pattern Liang et al. documented in Patterns (2023), where seven detectors falsely flagged an average of 61.22% of human-written TOEFL essays.

And if you're using QuillBot's humanizer side: know what it's for. If AI use is banned in your course or contract, laundering banned text through a paraphraser is a violation with extra steps — QuillBot's, ours, anyone's. The legitimate middle is drafting with permitted assistance and revising into your own voice, and no tool substitutes for the revising.

Limits, stated plainly

The free scan cap (1,200 words, six a day) is generous but real; long documents need slicing or a subscription. The explainer cards — the why behind flags — are mostly paywalled, so free users see what was flagged but less of the reasoning. No published false-positive rate appears anywhere on the product pages. The 99% figure inherits every caveat that applies to benchmark scores. Non-English accuracy is claimed but not independently documented. The 80-word minimum means short texts are out of scope — properly so, since short-text detection is unreliable everywhere. And the structural conflict of interest never fully resolves: a detector and a humanizer under one roof requires your informed skepticism, whichever roof it is. Including ours.

For contrast, Grammarly's detector offers a thinner free tool with no humanizer conflict and the category's most cautious framing, while Scribbr's pairs its free tool with the most transparent published research. QuillBot's is arguably the strongest free detector of the three — bought from a vendor whose other flagship product is the reason detectors struggle.

FAQ

Is QuillBot's AI detector free? Yes, up to 1,200 words per scan and six scans per day, with sentence-level highlighting and a 0–100% AI-likelihood score. Premium ($8.33/month billed annually at review time) removes the word and scan limits and adds full-text explainers and downloadable reports.

How accurate is QuillBot's AI detector? QuillBot claims a 99% detection rate citing the independent RAID benchmark, and Scribbr's third-party comparison scored its free tool 78% — among the best free results in that test. Both figures come with caveats: detection rate isn't false-positive rate, and accuracy drops sharply on paraphrased and mixed text by every published account, including QuillBot's own disclaimer.

Isn't it a conflict of interest that QuillBot sells both a humanizer and a detector? It's a real structural tension — the same one we carry at HumanFlow, which is why this review discloses it up front. The practical response isn't to boycott dual-sided vendors but to refuse to treat any vendor's detector as a neutral referee of its own rewriting tools, and to cross-check consequential scores across unrelated detectors.

Can QuillBot's detector detect text from QuillBot's own paraphraser? QuillBot markets mixed-content detection, and Scribbr's research found QuillBot-paraphrased text was the hardest category for every detector tested. We haven't run the self-detection test ourselves yet; our hands-on section above specifies exactly how we will.

What's the minimum text length for QuillBot's detector? 80 words. That's a floor, not a sweet spot — statistical detection needs material to measure, and every detector's reliability improves with longer continuous prose. Treat scores on very short texts skeptically regardless of tool.

Which languages does QuillBot's detector support? QuillBot claims support for more than 20 languages, including English, Spanish, French, and Portuguese. Independent accuracy evidence for non-English detection is scarce across the entire category, so apply extra caution to non-English scores.

Should a teacher act on a QuillBot score alone? QuillBot's own page answers this: "Never rely on AI detection alone to make decisions that could impact someone's career or academic standing." A score justifies a conversation and a look at drafts and version history — not a misconduct finding.

Key facts

  • Free tier: 1,200 words per scan, 6 scans/day; Premium removes limits at $8.33/month billed annually (quillbot.com, fetched 12 August 2026).
  • QuillBot claims a 99% detection rate citing the RAID benchmark (ACL 2024), without stating the configuration behind it; minimum scan length 80 words; 20+ languages claimed.
  • Scribbr's independent comparison (last revised July 2026) scored QuillBot's free detector 78% — tied with Scribbr free, behind only Scribbr premium (84%) among 12 tools tested.
  • The same Scribbr test found detectors performed worst on text paraphrased by QuillBot's own paraphraser; the top tool caught only 60% of paraphrased/mixed texts.
  • QuillBot's page states: "Never rely on AI detection alone" for decisions affecting careers or academic standing (quillbot.com/ai-content-detector).
  • QuillBot also sells a humanizer (free: 125 words, 6 uses/day; unlimited on Premium) — the dual detector-plus-humanizer position this review examines.
  • Liang et al. (Patterns, 2023): 61.22% average false-positive rate on human-written TOEFL essays across seven detectors — the category-wide bias context for any detector score.

Sources

  1. QuillBot — AI Content Detector page (quillbot.com/ai-content-detector), fetched August 2026.
  2. QuillBot — Premium pricing page (quillbot.com/premium), fetched August 2026.
  3. Scribbr — "Best AI Detectors" comparative research (scribbr.com/ai-tools/best-ai-detector/), last revised July 2026, fetched 12 August 2026.
  4. Dugan, L., Hwang, A., Trhlík, F., Ludan, J.M., Zhu, A., Xu, H., Ippolito, D. & Callison-Burch, C. (2024). RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors. ACL 2024. arXiv:2405.07940.
  5. Liang, W. et al., "GPT detectors are biased against non-native English writers," Patterns (Cell Press), 2023.
  6. OpenAI — AI text classifier retirement announcement, July 2023.
All postsPublished by The HumanFlow team