For teachers & students

AI essay checker

Check an essay for signs of AI writing without sending student work to a third-party server. Every sign is highlighted and explained — so a result can start a conversation instead of ending one.

Free · no sign-up · your file never leaves your device

0 words0 characters

What detection can and can’t do in a classroom

AI detectors are statistical tools. They estimate how much a text resembles chatbot output, and they are wrong some of the time in both directions. Published research and the experience of many schools show the same pattern: detectors catch a lot of unedited AI text, miss text that has been paraphrased or rewritten by the student, and occasionally flag genuine work — more often for writers who use a formal, careful, formulaic style, which includes many non-native English speakers and students taught to follow strict essay templates.

That does not make detection useless, but it changes how to use it. A detector is good at telling you where to look. GPTTrace is designed for that: rather than a single score, it highlights the exact sentences that match each sign and explains why the sign is associated with AI writing.

Signs that matter most in essays

  • Generic significance claims — “This pivotal moment left an indelible mark on history” with no specific evidence.
  • Summaries that restate — a conclusion that repeats the introduction’s three points in the same order.
  • Vague sources — “historians argue”, “studies have shown”, with no citation, or citations that don’t exist.
  • Perfectly even structure — five paragraphs of nearly identical length, each opening with a transition word.
  • Leftovers — “Certainly! Here is an essay on…”, Markdown symbols, or invisible characters from copy-pasting.

Invented or mismatched citations are often the most concrete evidence of all, and no detector checks them: look up two or three references.

Designing assignments that hold up

The most effective responses to AI in coursework are about the assignment rather than the detector: writing in class, staged submissions with drafts, personal reflection tied to class discussion, oral follow-ups, and topics specific to your course materials. Detection then becomes a backstop rather than the main line of defence — the role it is actually suited for.

How accurate is it? Our measured numbers

We test GPTTrace on labelled text samples and publish the results, including where it does badly. It is tuned to keep false accusations rare, so it misses some AI content rather than flag real work.

standard text check: AUC 0.795 (cross-validated)

2,467 labelled samples (1,258 AI, 1,209 human), run 2026-10-08. At the “Likely AI” line it caught 34% of AI samples and wrongly flagged 5% of human ones.

SourceTruthSamplesResult at “Likely AI”
pd-literatureHuman3061% wrongly flagged
human-pmc-eslHuman1835% wrongly flagged
human-wikipedia-pre2022Human1558% wrongly flagged

text check with deep scan: AUC 0.91 (cross-validated)

2,467 labelled samples (1,258 AI, 1,209 human), run 2026-10-08. At the “Likely AI” line it caught 56% of AI samples and wrongly flagged 5% of human ones.

SourceTruthSamplesResult at “Likely AI”
pd-literatureHuman3060% wrongly flagged
human-pmc-eslHuman1831% wrongly flagged
human-wikipedia-pre2022Human1551% wrongly flagged

Data sources and method: methodology & accuracy.

A fair process when an essay looks AI-written

  1. Run the check and read the highlighted passages, not just the percentage.
  2. Compare with the student’s earlier work written in class or before AI tools were common.
  3. Ask for process evidence: notes, outline, drafts, document version history, sources.
  4. Talk with the student about the argument and the sources. Someone who wrote an essay can explain and extend it.
  5. Decide on all the evidence together. Never on a detector score alone.

Frequently asked questions

How often does it flag human essays as AI?
We tune for rare false positives. On our test set, the deep scan flagged under 2% of academic writing by non-native English speakers and about 1% of pre-2022 Wikipedia text as “Likely AI”. The standard check is similar. That is not zero — in a class of 100 honest students one or two essays could still be flagged, which is why a score must never be the only evidence.
Is it biased against non-native English writers?
Many detectors are: research has shown some flag a large share of essays by non-native speakers, whose careful, formulaic English resembles model output. We include non-native academic writing in our test set specifically to measure this and set the threshold so it stays low.
Is student work uploaded or stored?
No. The check runs in the browser and nothing is sent to us or anyone else, which avoids the data-protection issues of uploading minors’ work to external services.
Can students use it on their own work?
Yes. If you are worried your writing might be misjudged, check it and look at which signs appear. Keeping drafts and version history is the best protection against a false accusation.
What about grammar tools like Grammarly?
Spelling and grammar correction changes little. Tools that rewrite whole sentences or paragraphs with AI can introduce AI-writing signs, and a heavily rewritten essay may score higher. Be clear with students about which tools are allowed.