Skip to content

HumanizersFor teachers

An AI Humanizer Rewrites Text to Dodge Detectors and Tends to Damage the Prose

An AI humanizer rewrites machine-written text so a detector will not flag it. The mildest version swaps a single word; the most aggressive rewrites whole passages. Pangram Labs names three techniques on its own blog: synonym swaps, inserted nonsense phrases, and plain damage to spelling and grammar [1]. A 2025 audit of 19 such tools found that humanizers tend to degrade the writing, to varying degrees, and the paper's tier averages put the original ahead in every tier [2].

Bill Nguyen & HumanUpdated

What Does an AI Humanizer Do to a Sentence?

It rewrites the sentence, one word at a time in the mildest version, and hopes the result reads as a person. Pangram Labs, the detector company, named three techniques in a post dated 27 January 2025. Humanizers swap words for their synonyms. They add nonsensical phrases, on the theory that a detector reads gibberish as unlikely to have come from a machine. The third technique is plain damage to the text quality 1.

The synonym pass is the quiet one. Clause order holds, sentence length holds, and the content words arrive slightly wrong for the register. The same company's January 2025 research paper audited 19 humanizer and paraphrasing tools and found that split running through the whole category: some tools retain low-level sentence structure and replace individual words with synonyms, while others take more liberty, rewording entire groups of sentences and paragraphs, adding sentences that were not in the original or deleting redundant ones 2.

The inserted noise is far easier to see than the synonyms. That audit collected fabricated references and strings of question marks left mid-paragraph. One passage trailed off into the fragment "CGSizeMake pp 18-23" where a citation belonged 2.

Vocabulary level is the tool's decision rather than the paper's. Some humanizers write only in a formal university register. Others drop to elementary or high school level, and the better ones, usually those built on language models, adopt the tone of the document handed to them 2.

What that leaves on the page has its own signature, and a teacher can learn it by eye: what humanized text looks like.

What the Rewrite Costs the Argument Underneath

The rewrite costs quality, on the one measure the auditors ran and in the tiers they drew by hand. The 2025 study asked GPT-4o to pick the more fluent and coherent of two passages, the original or the humanized version, over 25 samples per tool. The best tier of humanizers won 26.0 percent of those comparisons, the middle tier 14.67 percent, and the worst tier 2.67 percent 2.

Read the top figure the right way round. A win rate of 26.0 percent means that about three times in four, GPT-4o picked the original as the more fluent and coherent passage 2. The paper's own summary is flat: all humanizers tend to degrade the quality of the original text, though the degree varies 2.

Its three tiers were set on faithfulness and fluency alone, not on whether a tool beat a detector. At the top, tone, vocabulary level and complexity survived. The middle tier degraded the writing but kept the intent. At the bottom, nonsensical phrases often went in and sentences became uninterpretable. Meaning often shifted with them 2.

A paragraph from that worst tier, carrying an invented reference, a broken subject and verb, and a key term traded for a near-synonym that means something else, fails on its own merits before any detector runs on it.

None of that damage touches the integrity question. At the University of Southern California, the academic integrity office states that material created by a generative tool and presented as a student's own work counts as plagiarism whether it was copied verbatim, near-verbatim or paraphrased 6.

Under that wording a humanizer pass is paraphrase, and changes nothing. Other institutions write their own rule, and local policy governs: a humanizer is not a plagiarism remover.

Does a Humanized Paper Still Come Back Flagged?

Sometimes, and the published results disagree. In Pangram Labs' own January 2025 test, its detector trained with humanized examples held 98.26 percent on humanized academic text 2; an independent 2026 study found fewer than 4 percent still flagged 3. The two figures come from different test sets, different detectors and different years, so neither settles the other.

In that same January 2025 test, at a 5 percent false positive rate, GPTZero as tested then fell from 99.73 percent on machine-written passages to 60.04 percent once a humanizer had passed over them, and Binoculars from 94.15 percent to 28.23 percent 2.

Pangram's July 2026 technical report puts humanized text at 97.67 percent labelled AI-generated, with a per-system true positive rate running from 91.52 percent to 99.39 percent across the commercial tools it evaluated 5. All of these are figures from the company selling the detector, including the ones about its competitors.

The measurement pointing the other way is the 2026 University of Notre Dame study. It built machine rewrites of 642 published abstracts of 25 to 500 words, then ran the rewrites through a commercial humanizer. Fewer than 4 percent of the rewrites a detector had flagged as AI were still flagged afterwards 3. A separate Pangram test, reported by Nature, was of human-written student essays substantially modified with consumer AI, not of humanized machine text; 41 percent of those results were still labelled fully human 4.

Nature's caution about the whole category is the sentence to keep: only the accuracy rates the companies announce from their own internal testing are current, and those are not externally verified 4. Whether these tools work, and for how long, is its own question: do AI humanizers work.

Which Change Shows Up Where in the Paragraph?

Each technique lands in a different part of the paragraph. The synonym pass alters vocabulary and nothing else. Insertion adds material that was never in the draft. Degradation surfaces as misspellings and grammatical errors 12, which on their own look like any hurried draft and are not evidence of a tool.

What the tool doesWhat it looks like on the pageWhat it costs the paper
Synonym replacement 12Clause order and sentence length unchanged, content words slightly off registerPrecision of the key terms
Nonsensical insertion 12Fabricated references, strings of question marks, stray code-like fragmentsCredibility of the whole reference list
Quality degradation 1Misspellings such as "recieves", broken agreement, awkward phrasingMarks, before any detector runs
Whole-paragraph rewording 2Sentences present that were never drafted, redundant ones deletedThe line of argument in the original
Register flattening 2A uniform formal or school-level tone unlike the rest of the submissionThe voice a teacher already recognises

Price is a separate axis, and Human, the detector at human.olive.is, reports one finding on it. Whether a humanized document is caught depends on the humanizer's price tier more than on which humanizer it is. That reading is Human's own in-house measurement on its own held-out sets as of September 2026, not an independent audit. This is an estimate from our detector. Treat a flag as a reason to look closer, not as a finding.

The student side of that transaction is free humanizer versus paid tier.

The market behind these tools is neither small nor hidden. NBC News reported that Turnitin keeps a list of 150 such tools, some charging as much as 50 dollars for a subscription 7. An academic integrity company counted 43 humanizer sites drawing 33.9 million visits in a single month 7.

Grammarly's AI detector guide says the company does not currently offer a dedicated feature to reduce an AI score 8. Its ordinary red and blue underline corrections should not normally move a detection percentage, though generative rewriting of whole sentences and paragraphs will raise one 8. Grammarly also publishes a separate AI humanizer, and the page for that tool states that it is not intended to bypass AI detectors 9. Where that line falls inside a school is Grammarly's humanizer in school.

Check the Whole Submission, Then Ask for the Draft History

Run the full submission rather than the three sentences that read oddly, and write the tool and version beside any score that goes into a file. The Notre Dame cohort was abstracts of 25 to 500 words 3, so its figures describe short passages and may not carry to a full paper.

Then ask what a rewrite cannot answer. The draft history is quick to ask for and settles a question a second detector run cannot: Google Docs version history. A short conversation about the argument is the other, and it works whether or not any tool was involved: an oral follow-up after a flag.

One more finding from the Notre Dame paper belongs in the same conversation. Disclosed, honest AI editing carried a higher risk of sanction than concealing machine drafting behind a humanizer 3. The paper that declared the help reads as machine-edited. The paper that hid it behind a humanizer reads as clean. That asymmetry is an argument about policy, not about any one paper.

Length decides what a reading is worth, whichever tool produces it. At human.olive.is, Human reads college-level academic writing, reports sentence by sentence how much of a document reads as machine-written or machine-edited, and returns a document verdict of Human, Mixed or AI. The public check reads up to 1,000 words, three times a day, with nothing stored and no account, so a longer submission does not fit in one paste. Under 50 words it declines to give a reading at all. This is an estimate from our detector. Treat a flag as a reason to look closer, not as a finding.

Common questions

Is a humanizer the same thing as a paraphrasing tool?

The 2025 audit treated them as two categories and studied both. It listed DIPPER, Grammarly and QuillBot as paraphrasers, and tools such as StealthGPT, HumanizeAI.io, Humbot AI, Phrasly and WriteHuman as humanizers 2. A humanizer, in Pangram Labs' definition, rewrites AI-generated text to evade detectors 1. A paraphraser sells rewording with no such promise. The techniques overlap heavily, which is why a paraphrased passage and a humanized one can look identical on the page.

Does running a paper through a humanizer count as plagiarism?

At the University of Southern California, yes, under wording written for generative tools generally. That university's academic integrity office states that work created by a generative tool and presented as a student's own is plagiarism whether it was copied verbatim, near-verbatim or paraphrased 6. A humanizer produces a paraphrase of machine-written text. The rewrite does not change what is being submitted. Local policy governs, and it should be read before any conversation with a student.

Can a teacher tell by reading that a humanizer was used?

Sometimes, and the worst tools are the most obvious. The 2025 audit found low-quality humanizers inserting fabricated references and runs of question marks, building uninterpretable sentences and distorting meaning 2. The better tools leave much less: sentence structure holds and only word choice shifts 2. A near-synonym used slightly wrong across a paragraph reads as odd rather than as evidence, which is why it belongs in a conversation about the argument, not in a finding.

Does Grammarly work as a humanizer?

Grammarly's AI detector guide says the company does not currently offer a dedicated feature to reduce an AI score 8. Its own guidance draws a line down the middle of the product: ordinary red and blue underline corrections should not normally move a detection percentage, while using its generative assistant to meaningfully rewrite sentences and full paragraphs will raise it 8. Grammarly does publish a separate AI humanizer, and that tool's own page says it is not intended to bypass AI detectors 9.

If humanized text still gets flagged, is the flag enough to act on?

No. A detector reading describes text, and every figure above comes from a test set rather than from one student's paper. The published results disagree sharply depending on which tool, which detector version and which length was measured, from fewer than 4 percent of humanized machine rewrites still flagged in one 2026 study 3 to 97.67 percent in a vendor report the same year 5. Treat a flag as a reason to ask for the draft history.

References

  1. 1.What is a humanizer? Pangram Labs, 2025. pangram.comThe 27 January 2025 definition of a humanizer as a tool that rewrites AI-generated text to evade detectors, plus the three named techniques: synonym replacement, nonsensical phrases added so a detector reads gibberish as unlikely to have come from a machine, and text quality degradation with misspellings such as "recieves" and grammatical errors.
  2. 2.DAMAGE: Detecting Adversarially Modified AI Generated Text arXiv (Masrour, Emi and Spero), 2025. arxiv.orgThe January 2025 audit of 19 humanizer and paraphrasing tools: the paraphraser and humanizer category lists, the L1 to L3 tiers assigned on faithfulness and fluency, fluency win rates of 26.0, 14.67 and 2.67 percent over 25 samples per tool, the nonsensical-insertion examples including the "CGSizeMake pp 18-23" fragment, the structural-continuity and vocabulary-level findings, and the academic-text detection table at a 5 percent false positive rate.
  3. 3.Why AI Detection Fails for Academic Integrity Karr, Khvatskii, Hua and Chawla, University of Notre Dame (arXiv), 2026. arxiv.orgThe cohort of 642 published abstracts of 25 to 500 words, fewer than 4 percent of AI-labelled rewrites still flagged after a commercial humanizer, and the finding that disclosed AI editing carried a higher sanction risk than humanizer-assisted evasion.
  4. 4.AI-detection tools have made huge leaps forward — how good are they? Nature, 2026. nature.comThe report that Pangram's own test of consumer AI editing of student essays still labelled 41 percent of the results fully human, and the assessment that only the vendors' internal accuracy rates are current and those are not externally verified.
  5. 5.Pangram 4 Technical Report Pangram Labs and University of Maryland (arXiv), 2026. arxiv.orgHumanized text labelled AI-generated 97.67 percent of the time, and a per-system true positive rate ranging from 91.52 to 99.39 percent across the commercial humanizers evaluated.
  6. 6.Academic Integrity & Generative AI University of Southern California, Office of Academic Integrity, 2026. academicintegrity.usc.eduThe statement that work created by a generative tool and represented as a student's own is plagiarism whether paraphrased, copied verbatim or near-verbatim.
  7. 7.To avoid accusations of AI cheating, college students are turning to AI NBC News, 2026. nbcnews.comTurnitin's list of 150 tools charging as much as 50 dollars for a subscription, and Cursive's count of 43 humanizer sites with a combined 33.9 million visits in one month.
  8. 8.AI Detector User Guide Grammarly Support, 2026. support.grammarly.comGrammarly's statement that it offers no dedicated feature to reduce an AI score, that traditional red and blue underline corrections should not typically move the percentage, and that generative rewriting of sentences and paragraphs will raise it.
  9. 9.Free AI Humanizer | Humanize AI to Sound Like You, Not AI Grammarly, 2026. grammarly.comGrammarly's own statement that the humanizer is not intended to bypass AI detectors, and its warning that disguising AI use where AI is restricted may constitute cheating or plagiarism.

9 sources, numbered by first appearance.

General guidance for teachers, administrators and students. What holds at one institution, on one assignment, may not transfer to another.

Human reports how much of a document reads as machine-written. It does not report a probability that a person used AI, it does not check for plagiarism, and no number it produces stands for a student's honesty. This is an estimate from our detector. Treat a flag as a reason to look closer, not as a finding.

Human

Check a paper

Paste a paper and see which sentences read as machine-written, with a verdict of Human, Mixed or AI. Nothing is stored and no account is needed.