Berk Bayri

An AI detector can tell whether text was written by AI

A common misconception about AI models and evaluation, tested against the evidence.

The myth
An AI-writing detector can establish whether a particular piece of text was written by a human or by AI.
The reality
AI detectors estimate patterns, not authorship. False positives, false negatives, editing, paraphrasing, language background and model changes can all break the inference.

Explanation and evidence

The argument

An AI detector does not observe how a document was created. It observes the document after the fact and estimates whether its statistical features resemble examples it has learned to classify as AI-generated.

That is a very different claim from authorship.

OpenAI withdrew its own text classifier after describing its accuracy as too low for reliable use. Stanford researchers found another failure mode: several detectors disproportionately labeled essays by non-native English writers as AI-generated. More recent evaluations continue to find that detector performance changes with model generation, editing, paraphrasing, language and the mixture of human and AI text.

A detector can classify a text. It cannot witness who wrote it.

Why people believe it

The interface usually produces a percentage, label or colored score. Numbers feel forensic.

The word "detector" reinforces the impression that there is a stable fingerprint being found, like malware or a chemical residue. But modern generated text does not carry one universal visible signature, and humans can edit AI output while AI can imitate many human styles.

What the evidence says

Detector performance can be useful in controlled settings, especially when the model family, text length and domain resemble the detector's evaluation data. That does not turn the result into proof for an individual document.

matter most when the consequence is disciplinary, reputational or professional. A tool can be statistically useful across a dataset and still be unsafe as a verdict on one person.

The better question

Instead of asking, "Did AI write this?", ask:

What evidence do we have about the creation process?

Draft history, version history, source notes, document metadata and a conversation with the author can establish provenance more directly than a detector score alone.

Sources

New AI classifier for indicating AI-written text

OpenAI · 2023-01-31

AI-Detectors Biased Against Non-Native English Writers

Stanford HAI · 2023-05-15

Trusting AI to Detect AI? A Systematic Evaluation of the Reliability and Robustness of Current AIGC Detection Tools for Student Academic Work

Computers & Education · 2026-01-01