Skip to content

Research

Reports

The reports behind the numbers on this site, newest first. The reports are internal. This list gives each file’s name, its date and what it measured.

  • 2026-09-15

    b5_run_compare_v1_8_2026-09-15.json

    The live detector and the commercial detector we compared it with, on 51 heavily rewritten documents.

  • 2026-09-15

    unseen_human_1200_v1_8_2026-09-15.md

    Twelve hundred human-written documents nobody had scored before, read by the live detector.

  • 2026-09-14

    sensitivity_deploy_2026-09-14.md

    Shipping the three settings, and checking the default did not move in the deploy.

  • 2026-09-14

    sensitivity_points_2026-09-14.md

    Where the three settings sit, and what each one costs, measured on the live detector.

  • 2026-09-13

    v1_8_ship_decision_2026-09-13.md

    Which of the candidate runs went live, and what the choice cost.

  • 2026-09-13

    v1_8_seed2_score_2026-09-13.md

    Test scores for the candidate runs, side by side.

  • 2026-09-13

    seed_interval_2026-09-13.md

    How much of a one-document difference comes from the training run alone.

  • 2026-09-12

    v1_8_train_2026-09-12.md

    A training pass for the candidate runs, and the margin it left.

  • 2026-09-12

    v1_8_build_2026-09-12.md

    How the training set was put together, and which documents were kept out because they are test evidence.

  • 2026-09-11

    beating_the_humanizer_2026-09-11.md

    What rewriting tools do to a reading, on an earlier run.

  • 2026-09-10

    decoder_vs_data_2026-09-10.md

    Whether a gain on an earlier run came from the scoring arithmetic or from the training documents.

  • 2026-09-10

    whitepaper_evidence_2026-09-10.md

    Claims traced back to per-document rows, for the two runs before the live one.

  • 2026-09-09

    compute_rates_2026-09-09.md

    What a training pass and a scoring pass cost to run.

  • 2026-09-09

    container_ceiling_2026-09-09.md

    A ceiling on the serving setup in its busiest hour, written down and tested.

  • 2026-09-09

    light_edit_signal_2026-09-09.md

    Whether an earlier run could place a light polish edit inside a paper.

  • 2026-09-09

    placement_whole_papers_2026-09-09.md

    Where the marked passages fall, checked on whole papers, for an earlier run.

  • 2026-09-09

    serve_latency_2026-09-09.md

    Response times on one document, from calls that had already happened. No load test.

  • 2026-09-09

    threshold_sweep_oos_2026-09-09.md

    False flags at many lines, repeated on human documents an earlier run had never scored.

  • 2026-09-09

    truncation_artefact_2026-09-09.md

    A measurement error that raised scores on human writing, found and corrected.

  • 2026-09-09

    trust_set_scores_2026-09-09.md

    Scores on documents kept out of training on purpose, for an earlier run.

  • 2026-09-09

    unseen_human_600_2026-09-09.md

    Six hundred human documents nobody had scored before, read by an earlier run.

Paths are relative to the detector project’s root. The live detector is run v1_8_s2_seed303. Reports dated before 2026-09-12 describe earlier runs. Only a report that names the live run carries its figures.

Check a paper yourself.

Free with a verified school email, up to 1,000 papers a month.