Skip to main content
proofmytext

How we test the detector

Most tools publish an accuracy number and stop there. Below is the whole test behind ours — the texts, the threshold and what the result does not prove. Last run: August 2026.

What we ran

What came back

Human texts wrongly flagged as AI0 of 30
AI texts correctly caught22 of 22
Highest score any human text received18.5%
Lowest score any AI text received86.7%

The gap between those last two numbers is the point: on this corpus human and machine writing did not overlap, so the verdict did not depend on where exactly the threshold sits.

What this does not prove

Our own test on public-domain literary English against one AI model — not an independent audit.

We re-run this test as the detector changes and update the numbers here.