HumanizersFor students
Pangram's August 2025 Table Scores QuillBot at 100.0 Percent
A QuillBot rewrite is detectable by at least one detector company's own account. Pangram Labs publishes a self-scored figure for QuillBot: its humanizer page, updated 27 August 2025, has a Quillbot row reading 100.0 percent accuracy on Pangram's own test set [1]. How big was that set? The post doesn't say, and it gives no account of how the passages were made. Past QuillBot's own product page, every other result for this search belongs to a rival humanizer selling its own tool.
Bill Nguyen & HumanUpdated
What Does the One Published Figure Actually Say?
Pangram Labs keeps a page on how well its detector does against humanizers, updated on 27 August 2025. The table has a Humanizer column and an Accuracy column, and the row the table spells Quillbot reads 100.0 percent 1. The post sums it up as Pangram performing above 90 percent on all the notable humanizers it tested 1. So a detector company is grading itself here.
What the page leaves out matters as much as the row. There is no sample size. Nothing on which model drafted the passages or how long they ran, and nothing on which QuillBot setting produced the rewrite 1.
The post calls the column detection accuracy on text run through the tools and stops there, with no statement of what counted as a hit 1. A number with no denominator under it can't be rechecked. On a humanizer benchmark that word usually means the share of rewritten machine-written passages the detector still called machine-written, which is an inference rather than a quotation. Look at the spread. Undetectable AI sits at 90.3 percent, TwainGPT at 92.7 percent, and ten of the twenty rows land on exactly 100.0 percent 1. Grammarly is one of those ten. That says something about rewriting tools and nothing about a student who ran a spell check.
Pangram's 2026 technical report, written with the University of Maryland, gives ranges instead of a row. In it, humanized text was detected as AI 97.67 percent of the time and as Mixed or AI 98.83 percent. Per-system AI recall ran from 92.78 to 99.70 percent across thirteen commercial humanizers, which the report names only as Commercial A to M, so no row can be matched to QuillBot 2. One tool's score doesn't transfer to another. And a number from August 2025 describes the detector and the rewriting tools of that month, not this one.
Who Else Publishes QuillBot Results, and What Do They Sell?
Competing humanizers, mostly, plus detector vendors scoring themselves. A check of this query on 18 September 2026 returned QuillBot's own product page and four rival humanizer sites. None of them was an independent test. One of the four, a competing humanizer, ran QuillBot's output through three detectors, reported it flagged each time, and reported its own output passing a paragraph later. The other three are reviews published by competing humanizers.
The publisher chose the passages and the detectors, and sells the tool that wins. That's an advertisement with a methods section. Originality.ai, another detector company, publishes its own humanizer figure of up to 97 percent 9, one more vendor scoring itself.
Why so many of these pages? Scale. NBC News reported in 2026 that Turnitin keeps a list of 150 tools charging as much as 50 dollars a subscription to adjust text so a detector won't flag it 3. The same report counted 43 humanizer sites and 33.9 million visits in one month, figures that came from an academic integrity company 3.
That's a large market with no referee in it.
An evidence-grade test needs nothing exotic: a document count, the model that wrote the documents, the detector version with its run date, and an author with nothing to sell on either side. QuillBot's own pages returned an error to the automated request made for this article, so nothing here reports how the company describes its own feature.
What the rewrite does to a paragraph is a separate question from who is scoring it, and it has its own answer: what an AI humanizer does to writing.
What Does a Paraphrase Pass Change, and What Does a Detector Read?
The surface, mostly. Pangram Labs defines a humanizer as a tool that rewrites machine-written text to evade detectors, and names synonym swapping and inserted nonsense among the documented techniques 4. Pangram's August 2025 page lists what it looks for in humanized text: tortured phrases, unnatural spacing, repetitive phrases and non-standard characters 1.
Rewriting used to work well. A 2023 study built a paraphraser called DIPPER. It drove DetectGPT's accuracy from 70.3 percent down to 4.6 percent at a fixed 1 percent false positive rate, without appreciably changing what the text meant 5. Those detectors are three years and several model generations old.
Detection of rewritten text has been broken and then repaired since, on the record cited here. That's why an honest answer carries a date. A page that answers with a flat yes or a flat no, and names no month, is describing a moment it hasn't identified. There's also the cost to the writing. A synonym pass trades precise terms for near-synonyms, and inserted filler reads as filler to a marker long before any tool runs 4. A paragraph that survives a detector and loses the argument has already lost the grade.
The artefacts a rewrite leaves behind are visible without software: what humanized text looks like.
Read the Integrity Policy Before Reading the Detector Score
At USC, a rewrite that no detector flags is still covered by the plagiarism rule. The University of Southern California's academic integrity office, on its page as retrieved on 18 September 2026, states that material created by a generative tool and presented as a student's own work counts as plagiarism whether it was copied verbatim, near-verbatim or paraphrased 6. A paraphrase pass is the paraphrased case, so the policy reaches the rewrite along with the draft it came from.
A student at another school has that school's own policy to read.
The vendor pages fold two questions together. Whether a detector flags a passage is a measurement, and it moves with the detector and the month. Whether a submission breaks the rule is a policy question, and people decide it.
That distinction decides what a rewrite is actually worth, which is the argument under a humanizer is not a plagiarism remover and under is QuillBot cheating.
Check the Draft, Keep the History, and Read the Number Carefully
Keep the version history, then check the draft: both are worth more than another pass through a tool. A dated trail of drafts is evidence a detector reading can't produce. Read what a detector says about the draft itself, knowing what that reading is and what it isn't.
No detector score is a finding. Turnitin's guide names QuillBot as an example of the paraphrasing tools behind one of its report categories, and gives no catch rate for it 10. In a June 2023 post, Turnitin put the sentence-level false positive rate of its own detector at around 4 percent, which works out at roughly one highlighted sentence in 25 that may be human-written after all 7.
Length changes the reading more than most people expect. In 2023 Turnitin raised its own minimum from 150 to 300 words, reasoning that accuracy improves with more text 8. Human declines to score anything under 50 words at all, and 50 to 149 words gives it little to read. A reading on a short excerpt is weak in either direction.
The one false-flag figure Human publishes for its own detector is an in-house measurement on its own held-out set, not an independent audit, as of September 2026: 0 wrongly flagged of 600 human-written papers it had never been trained on; the ceiling on that is 0.498%. This is an estimate from our detector. Treat a flag as a reason to look closer, not as a finding. That figure counts human writing misread as machine-written, so it says nothing about how often a QuillBot rewrite slips through.
A dated folder of earlier drafts outranks a screenshot of a vendor's table.
Common questions
Does Turnitin flag text rewritten with QuillBot?
Turnitin names QuillBot as an example of the AI-paraphrasing tools one of its report categories is meant to cover. It publishes no QuillBot-specific catch rate 10. The figures it published in June 2023 describe the detector overall at that time: a document-level false positive rate under 1 percent for documents it marks at 20 percent or more AI writing, and a sentence-level rate of around 4 percent 7. When a vendor page attaches a percentage to QuillBot, that number came from the vendor's own run, not from Turnitin. To learn what an instructor's Turnitin report says, ask the instructor.
Is QuillBot's paraphraser the same thing as a humanizer?
The categories overlap, and vendors draw the line in different places. Pangram Labs defines a humanizer as a tool that rewrites machine-written text to evade detectors 4, and it lists Quillbot in its humanizer benchmark table 1. A paraphraser rewrites text whatever its origin. That's why one feature can serve a reading-level edit and an attempt at evasion alike. QuillBot's own pages couldn't be retrieved for this article, so nothing here reports how the company describes the feature.
Does a 100.0 percent row mean every QuillBot rewrite gets caught?
No. It means one detector company scored itself on its own test set and reported that result for that month 1. The same company's 2026 technical report puts per-system AI recall across thirteen anonymised commercial humanizers at 92.78 to 99.70 percent 2. The answer varies by tool. And the reading moves with more than the tool: a different detector, a different source model, a shorter passage or a heavier hand-edit can each shift it.
Do the humanizer sites that test QuillBot count as evidence?
Not in a disciplinary conversation. When a company selling a competing tool publishes a test, it decides which passages and detectors go in and what gets reported, and it profits from one outcome. The check made for this article on 18 September 2026 found four such pages and no independent one. Faced with a vendor's comparison on one side and a student's own draft history on the other, an integrity office will take the draft history.
Does getting past a detector make the work original?
No. Changing the wording leaves the authorship question exactly where it stood. The University of Southern California's integrity office counts material created by a generative tool and presented as a student's own work as plagiarism whether it was copied verbatim, near-verbatim or paraphrased 6. No percentage settles that.
How long does a passage need to be before a reading means anything?
Longer than most rewritten excerpts. Turnitin raised its own minimum from 150 to 300 words in 2023, saying accuracy improves with more text 8. Human declines to score anything under 50 words, and 50 to 149 words gives it little to read. One paragraph pasted into a free checker is the weakest evidence there is, whichever way the number falls.
References
- 1.How well does Pangram perform on humanizers? (Updated August 2025) Pangram Labs, 2025. pangram.comThe table listing Quillbot at 100.0 percent accuracy, the 27 August 2025 date, the Undetectable AI and TwainGPT rows, and the post's above-90-percent summary line.
- 2.Pangram 4 Technical Report Pangram Labs and University of Maryland (arXiv), 2026. arxiv.orgHumanized text detected as AI 97.67 percent of the time and as Mixed or AI 98.83 percent; per-system AI recall of 92.78 to 99.70 percent across Commercial A to M (Table 14), with the humanizers anonymised.
- 3.To avoid accusations of AI cheating, college students are turning to AI NBC News, 2026. nbcnews.comTurnitin's list of 150 rewriting tools priced up to 50 dollars, and the count of 43 humanizer sites drawing 33.9 million visits in a month.
- 4.What is a humanizer? Pangram Labs, 2025. pangram.comThe definition of a humanizer as a tool that rewrites AI text to evade detectors, and the inserted-nonsense technique.
- 5.Paraphrasing evades detectors of AI-generated text, but retrieval is an effective defense Krishna, Song, Karpinska, Wieting and Iyyer (arXiv), 2023. arxiv.orgDIPPER dropping DetectGPT accuracy from 70.3 percent to 4.6 percent at a 1 percent false positive rate.
- 6.Academic Integrity & Generative AI University of Southern California, Office of Academic Integrity, 2026. academicintegrity.usc.eduGenerative AI material presented as a student's own work counts as plagiarism whether paraphrased or copied. The page shows no publication date; the year is the retrieval date, 18 September 2026.
- 7.Understanding the false positive rate for sentences of our AI writing detection capability Turnitin, 2023. turnitin.comThe document-level false positive rate under 1 percent at 20 percent or more AI writing, and the sentence-level rate of around 4 percent.
- 8.AI writing detection update from Turnitin's Chief Product Officer Turnitin, 2023. turnitin.comThe minimum word requirement raised from 150 to 300 words because accuracy improves with more text.
- 9.AI Content Detection Accuracy Originality.ai, 2026. originality.aiOriginality.ai's own claim of up to 97 percent accuracy on current humanizers and bypassers, a vendor measuring itself.
- 10.AI writing detection model Turnitin Guides, 2026. guides.turnitin.comRetrieved 19 September 2026: the AI-paraphrased category covers text likely revised using an AI-paraphrasing tool or AI word spinner, such as Quillbot.
10 sources, numbered by first appearance.
General guidance for teachers, administrators and students. What holds at one institution, on one assignment, may not transfer to another.
Human reports how much of a document reads as machine-written. It does not report a probability that a person used AI, it does not check for plagiarism, and no number it produces stands for a student's honesty. This is an estimate from our detector. Treat a flag as a reason to look closer, not as a finding.