Verification

GPTZero Verification for Articled Content

GPTZero is a detector that has been most often used by people. A clean result provides social utility, but the technology behind it is limited. This is how we assess GPTZero relative to the other checks we run and what these numbers represent.

Order content How it works

How GPTZero Verification Fits Our Human-Writing Process

Like the other models in this panel, it relies on a finished file. It also happens to be the first detector result we look at for a very logical reason. It is the tool most likely to be used by your reader, and due to its free interface and school use, it appears far more frequently in disputes regarding authorship than any other paid model.

Perplexity and burstiness were the foundation for GPTZero’s public explanation. Even now, those are the metrics it relies on. Perplexity is how expected a word is as the next output of a language model; burstiness measures the variation in sentence-length rhythm. Most generated text is likely to sit low on both measures. In less technical language, this means output is even, consistent, and is unlikely to trip or stumble.

The output separates into three sections rather than two: entirely human, entirely AI, or mixed. The mixed section is the most honest, and many long, human first drafts fall into this zone, as everybody writes a flat paragraph every now and then. An editor reads the highlighted sentences against the writer’s draft history and research file before anyone reacts to the headline number.

Read the highlights, not the headline

The document-level percentage roughly summarizes the readings of the sentences beneath it. The sentence view is the only feature editors can actually check against a draft, which is why we read that section first.

What GPTZero Verification Can and Cannot Show

This detects “long runs” of the same color when a user pastes the output from unedited chatbot responses. This is a desired failure mode from a customer purchasing a writing service. It is unclear how the text appears after a person edits it, and this ambiguity is usually the case. This is the hybrid case that is most relevant commercially and the case that no detector accounts for.

The known failure works the other way as well. Independent investigations have shown that detectors flag writing by non-native English speakers more often than writing by native speakers, and GPTZero has had to deal with this directly. Simple sentence structure and a small range of idioms look, statistically, very similar to a model playing it safe. This is a bias in the instrument as opposed to a fault in the writer.

There’s a length minimum for classifiers. Below a few hundred words, there isn’t enough variation for a classifier to make sense, so the vendor sets a minimum at that length for that reason. Short copy scored leads to confident nonsense in both directions that we would rather avoid, so we would rather say there is a limitation than sending you a number that isn’t meaningful.

  • Passages under a few hundred words, where there is not enough variation to measure.
  • Writing that reads statistically flat, typically due to second language proficiency.
  • Tightly formatted copy including: specifications, legal boilerplate, procedures, step lists.
  • Human text edited hard for concision, since heavy editing smooths the rhythm out.

When GPTZero Verification Is Used

Every draft goes through scanning after editing and before the delivery email. The result is dated and attached. A revision causes a rescan, as the document received after the second round is not the document scanned in the first round, and an un-updated report is of no use to anyone.

We can run it mid-order if you request it, which can be useful if your own reviewer uses GPTZero and needs a reading for sign-off. Please do this at the order time request rather than after the writing has been delivered. Then the writer will be able to provide the source list and the outline along with the scan, and that provides significantly more details than a percentage.

We will never adjust writing based on a particular scoring system. If a mixed reading passes the review of an editor, the draft will be accompanied by a report and reasoning, and the decision to publish will remain in your control. Satisfying a scoring tool is a slow and painful way to degrade a piece of writing, and is grounds for the immediate termination of a writer’s contract.

  • The dated GPTZero result for the exact file being delivered to you.
  • The writer’s outline and the source list, as requested at the time of order.
  • A new scan is provided after each of the two rounds of revision included in the price.
  • A simple note, where the document was too brief to allow for a reliable assessment.

GPTZero Verification FAQs

A human writer handles every part of this. Ask us on the contact page and a person will answer the same working day.

A human writer handles every part of this. Ask us on the contact page and a person will answer the same working day.

Not by itself. Mixed is the label the tool uses when some sentences read flat and others do not, and long human documents commonly contain both. We choose to show you the writer’s notes and the sentences we’ve highlighted over rewriting a paragraph to reclassify the text.

No detector can. It can claim that the text does not look like the writing it was trained on, but that is much weaker. Provenance evidence connects a document to a writer through things like drafts or author logs.

Where a document is written in a supported editor, the text typed can be replayed. This is better evidence than any style score, and we can provide it upon request. However, it is determined by the writer’s own means, so ask during the time of order.

Differences in training data and thresholds. One vendor sets up their solution to be more sensitive to paraphrased machine text at the cost of more false positives. Another vendor’s solution does the opposite. Disagreement is natural. It is why we opted to use a panel over a single tool.

Both the model and the tier read the same file. Although the paid version does provide some operational conveniences in the form of bulk scanning and exportable records, the free version of the model actually does a better reading of your document. Don’t think of your free result as being worse.

People tend to prefer this tool over any other. That has less to do with accuracy and more to do with how the internet functions. That is why a GPTZero result is included in all deliveries; we let our users know what the range for the number is more likely to be.

Content a person actually wrote

$10 per 100 words, and the writer keeps all of it. Our 1% sits on top, 0.5% goes to trees, and no generated text appears anywhere in the process.