Skip to content
FreeFree for schools and universities

Know what an AI flag is worth.

We ran 960 documents through our detector and a commercial one. Neither flagged any of the 600 written by people. With a verified school email, ours is free for up to 1,000 papers a month.

Our detector, run v1_8_s2_seed303, standard setting (6%). One test, scored by us between 8 and 15 September 2026. A flag from our detector is a reason to look closer. It is not proof that anyone cheated.

  • Sentence highlights
  • Three settings
  • 50-word minimum
  • No account needed
  • English only
  • API and MCP
Try an example:

Paste 50 to 1,200 words of English. Our detector marks the passages it reads as AI-written and gives a verdict. No account needed. A longer paper needs a free account.

93 words

Comparison

Same 960 documents, two detectors

We ran our detector and Pangram 4 on identical text: 600 documents written by people (180 of them by English learners), 220 written by AI, and 140 papers people wrote that an AI then partly edited. On the first two groups the result is a tie.

  • Written by people

    Wrongly flagged, of 600

    Our detector
    0
    Pangram 4
    0
  • Written by AI

    Missed, of 220

    Our detector
    0
    Pangram 4
    0
  • Partly AI-edited

    Flagged, of 140

    Our detector
    102
    Pangram 4
    88
  • Learner essays

    Wrongly flagged, of 180

    Our detector
    0
    Pangram 4
    0
  • The 220 AI-written documents come from models our detector was trained on.
  • Our default line (6%) sits lower than Pangram 4's own line (about 10%). At the same 10% line the count is 99 to 88. Most of the gap is papers where a model added a paragraph. One test, 140 papers.
  • A flag from our detector is a reason to look closer. It is not proof that anyone cheated.

Our detector, run v1_8_s2_seed303, standard setting (6%), read 2026-09-15. Pangram 4 as it answered on 2026-09-08, scored once by us through its web dashboard. We run and grade our own tests.

See the comparison

Roles

What a result means to you

Choose the tab for your role to see what a result means for you.

When a paper is flagged

Read the passages our detector marks. Then set them beside the student's drafts and earlier work before you decide what to ask.

Students learning English

Our detector flagged none of 1,195 essays by English learners kept out of training as a fairness test. They come from one university program, written 2005 to 2012.

See the teachers page

Settings

What each verdict and setting means

Human means the part of the paper that reads as machine-written is at or under your setting’s line. Mixed means some of it reads as machine-written or machine-edited. AI means 80% or more of it does. When this site says a paper was flagged, it got Mixed or AI.

  • Accusation-safe

    Flags only above 15%

    Flags a paper only when a substantial part reads as machine-written. It called 10 more AI-edited papers Human than standard did, and removed no false flags: standard had none.

  • Standard

    The default, above 6%

    Flags a paper when more than a small part reads as machine-written. It flagged none of 1,928 human-written papers and essays it was never trained on. The 95% ceiling on its false-flag rate is 0.155%.

  • Sensitive

    Flags above 2%

    Flags at the first small sign of machine-written text, and flagged 2 of 1,928 human-written documents. Use it to prompt a second look, not a conversation about misconduct.

Changing the setting never turns an AI verdict into Human. It only moves the line for papers that are partly machine-written. Our detector, run v1_8_s2_seed303, measured 2026-09-15.

Limits

What it misses and will not do

These are the cases where a result from our detector deserves less weight. The counts come from our own tests on run v1_8_s2_seed303.

  • Light AI polish

    When an AI model lightly polishes a human paper, our detector mostly misses it. Of 39 polished papers, it flagged 3.

  • Paid humanizer tools

    Paid humanizer tools can hide machine writing from our detector.

    The test is on the model card
  • Short answers

    Under 50 words you get Not scored rather than a verdict. From 50 to 149 words the reading is weak.

  • English academic writing only

    Our detector reads English academic and student writing. It is not built or tested for other languages or other kinds of writing.

  • Writing before 2022

    Every human-written document we tested predates 2022. Writing done since then has not been tested, and neither has writing done with grammar checkers or translation tools.

  • No outside check yet

    We run and grade our own tests. No outside party has checked these results.

  • Not a plagiarism check

    It does not check plagiarism or images, and it does not fetch web addresses.

Education

Free for schools and universities

Sign up with a verified school or university email and check up to 1,000 papers a month free, in the web app, the API or the MCP server.

Get the school plan
  • Who qualifies

    Addresses ending in .edu, or in .edu, .ac or .sch followed by a country code, such as .ac.uk. United States k12 school domains qualify too. We check at sign-up.

  • What counts as a paper

    One check of up to 25,000 words counts as one paper.

  • Same detector

    The school plan runs the same detector as every other plan, with the same three settings.

  • More than 1,000 a month

    A district or university that needs more can ask about volume at hello@olive.is.

FAQs

Before you trust a result

Talk to us

Yes. With a verified school or university email, you check up to 1,000 papers a month free. Anyone else gets 2,000 words a day free.

The detector is ours. We run it ourselves on a cloud GPU provider, and no other company's detector sees the text. The free school plan is how we get it in front of teachers. The paid plans for heavier use pay for it, once billing opens.

No. A flag from our detector is a reason to look closer. It is not proof that anyone cheated. Read the marked passages, then talk with the student with their drafts and earlier work in front of you.

Not in our tests. Our detector flagged none of 1,195 essays by English learners kept out of training as a fairness test. They come from one university English language program, written 2005 to 2012. It also flagged none of 600 more it had never scored, from collections whose other essays were partly used in training. That is a count, not proof of fairness.

Human scores the text in memory and does not save it. Neither this site nor our detector keeps a copy of the papers you check.

Our detector flagged none of 1,928 human-written papers and essays it was never trained on. The 95% ceiling on its false-flag rate is 0.155%. The ceiling is the highest error rate still believable given what we saw. Our detector called all 481 fully machine-written test documents AI, at every setting. The 95% ceiling on its miss rate is 0.62%. They were written by AI models used in training, or rewritten by our own tools. We don't give a single accuracy percentage. Each kind of error is reported on its own test instead.

In one 960-document test we ran, the two tied on writing by people and on fully AI-written text: neither flagged any of 600 human-written documents or missed any of 220 AI-written ones. On 140 papers an AI had partly edited, ours flagged 102 and Pangram 4 flagged 88. Our default line (6%) sits lower than Pangram 4's own line (about 10%). At the same 10% line the count is 99 to 88. Most of the gap is papers where a model added a paragraph. One test, 140 papers. The 220 AI-written documents come from models our detector was trained on.

Turnitin says its own document-level false positive rate is under 1% for documents it scores at 20% AI or higher, and shows no percentage at all below that line (Turnitin, 2023-06-14). A 2026 study, "Who wrote this? Evaluating the reliability of AI detection tools in higher education" (International Journal for Educational Integrity, 2026-06-23), found Turnitin scored all 40 fully AI-written papers in its set under that 20% line, missing every one. The 40 were master's-level papers one AI tool wrote for that study, not real student submissions. We have not tested Turnitin ourselves. See the studies on the comparison page

Partly. It flagged 102 of 140 papers an AI had partly edited. When an AI model lightly polishes a paper, our detector mostly misses it: of 39 polished papers, it flagged 3.

Yes. Paid humanizer tools can hide machine writing from our detector. The test is on the model card.

Our detector does not score text under 50 words, and readings under 150 words are weak. It tells you that instead of guessing.

Start with standard, the default. It flags a paper when more than 6% reads as machine-written. Accusation-safe (flags only above 15%) misses more AI-edited papers. Sensitive (above 2%) flagged 2 of 1,928 human-written documents. Check whether your school already has a rule.

No. Our detector reads English academic and student writing. It is not built or tested for other languages or other kinds of writing.

Not yet. We run and grade our own tests.

Start with a paper you know.

Paste something you know a person wrote, such as your own writing, and see what our detector marks in it. With a verified school email, you check up to 1,000 papers a month free.

Turnitin is a trademark of Turnitin, LLC. Pangram is a trademark of Pangram Labs, Inc. Human is not affiliated with either.