How teachers detect AI writing, and what a percentage actually shows

Teachers detect AI writing in three ways, and only one of them is a tool. The first is a detector, usually an AI-writing indicator attached to the similarity service your work is already submitted through. The second is the marker reading the work: they have your earlier essays, they wrote the brief, they know what the seminar covered in week six, and mismatches there are what starts most conversations. The third is the record around the submission, meaning timestamps, save counts and version history the platform kept whether or not anyone looked. The detector percentage is the weakest of the three. It is a classifier's confidence about writing style, there is no source document behind it, and two tools can disagree about the same paragraph on the same afternoon. The things that actually settle a case are the citation that does not exist and the conversation about your own argument.

The tools an institution actually runs

Most detection starts inside the system you already submit through. Similarity checking has been built into university submission portals for years, and several of those services now report an AI-writing figure next to the match report. Your instructor usually sees a number they did not calculate and cannot inspect, produced by a vendor that set the threshold somewhere above them.

The second tool is ad hoc. An instructor who suspects something pastes a paragraph into a free web detector between classes. Whether that happens at all depends on the person marking. The third is not a detector at all. It is the record the platform kept anyway.

  • An AI-writing indicator inside the similarity report, visible to staff and often not to you.
  • Standalone web detectors, run at an individual instructor's discretion, usually with no institutional record of what was checked.
  • Submission metadata: when the file arrived, how many saves it took, how long the assignment was open.
  • Document version history, where the work was written or submitted in a platform that keeps it.
  • Assignment design that removes the question entirely: in-class writing, staged drafts, a short oral on your own argument.

What a marker notices before any tool is opened

Most cases start with a person rather than a score. The marker holds your earlier essays, wrote the brief and sat in the seminar. Work that stops sounding like the student who handed in the last three assignments is the strongest signal anyone has, and getting to it costs one folder search, because nobody else holds that comparison.

The rest is the texture of the prose:

All of that is visible in your own draft for free. The registry of named AI writing tells highlights the exact phrases in your text as the page loads, and the AI vocabulary checker and the readability checker run in the browser tab too, unlimited, with nothing to sign up for.

  • Flat rhythm. Sentence after sentence at the same middling length, with no short one anywhere.
  • The standard shape. An opener about the growing importance of the topic, three body paragraphs of near-identical length, a conclusion restating the same three points in the same order.
  • Stock vocabulary. Delve, leverage, seamless, robust, underscore, moreover, furthermore, navigate the, it is worth noting, in today's landscape.
  • Confident generality. Claims that would fit any book on the reading list, with no page number, no quotation and no named example.
  • Stacked hedges. Arguably somewhat effective in most cases is three qualifiers on one claim.
  • Leftovers. The brief restated in its own words as the first line, bold subheadings nobody asked for, a bulleted list dropped into the middle of an essay.

The checks that settle it, and they are not about style

Style raises a question. Content answers it. The fastest way a marker turns a suspicion into something concrete is to pull one citation. A reference that does not exist, or that exists and does not say what the essay claims it says, takes a minute to check and cannot be argued with.

The same goes for course-specific material. An essay that never touches the reading the seminar spent a fortnight on, or that applies a framework the module explicitly set aside, or that answers a question adjacent to the one that was set, is missing the things that were only ever said in the room.

The last check is a conversation. Which part gave you the most trouble, why did you cut the paragraph that used to be third, which source changed your mind. Somebody who wrote the essay answers in a sentence. Each of those checks produces something both sides can examine and discuss. A percentage does not, and in a fair process it carries less weight than students expect.

Why a percentage is not proof

A detector score is a classifier's confidence that a passage looks machine-made, computed from writing style alone. A plagiarism match names the document it matched and quotes the line, so you can go and read it. An AI reading has nothing to show you, because there is no source document to show.

The threshold that turns the number red is a vendor's choice and is usually unpublished, so flagged means the score crossed a line somebody else drew. If a tool returns a similar share of flags in every cohort, that says more about the tool than about any student in it.

Then there is who gets caught by mistake. Detectors key on prose that is clean, even and correct, which is what careful students, heavily taught writers and second-language writers produce. Liang et al. at Stanford HAI found in 2023 that 61% of TOEFL essays written by non-native speakers were falsely flagged as AI across seven detectors. Grammar assistance and dictation both flatten sentence rhythm, and a detector reads the finished sentence without seeing which tool helped shape it. The full version of that argument, with what to say if it happens to you, is at why detectors flag human writing.

What your writing record shows, and what it does not

Where the work was written in a platform that keeps version history, there is a record of how the document was built, and it is more informative than any score. A steady accumulation across four evenings reads differently from one paste at two in the morning.

It is not proof either way. Plenty of people draft in one application and paste the finished text into another, and plenty write in one sitting on paper or on a phone. A missing revision trail is not an admission, and treating it as one penalises a working habit rather than a behaviour. The version of this argument written for markers is at for teachers, and it is a reasonable thing to point them to.

If you want a record you control, write in the drafting record tool. It logs timestamps, how much you added and cut, and whether anything large arrived in a single paste, then exports a report signed in your browser with an ECDSA P-256 key. Anyone you hand it to can check it at verify a report, in their own browser, without contacting us: the signature confirms the report has not been altered, and the hash confirms it belongs to that exact draft. No tool that reads finished text can prove authorship, so what this gives you in a meeting is specifics to point at instead of a general denial.

Checking your own draft before you hand it

The part you control is knowing what your writing reads like. Paste at least 65 words into the checker and you get a 0 to 100 score plus one of three readings: reads as AI-generated, borderline, or reads as human-written. A run costs 1 credit whatever the length, and the free allowance is 10 credits a day per network, with no account, no login, no card and no email address.

A score is a reading of style, not a verdict on who wrote the text. What it is good for is telling you whether your prose has drifted into the shape described above, at which point the free highlighters name the sentences and you fix them. The rewriter is there if you would rather do it in one pass: it rewrites every sentence in one click, from 25 words up to 12,000 characters a run, roughly 2,000 words, at 1 credit per 50 words.

One thing to be plain about. Rewriting a draft you were permitted to generate is ordinary editing. Where AI is not permitted, that rule covers the draft however it is edited. Read the policy that applies to you.

Questions people actually ask

How teachers detect AI writing, and what a percentage actually shows: common questions

Can teachers really tell if you used ChatGPT?

Sometimes, and rarely from a detector alone. The reliable signals are ones a person notices: work that does not sound like your earlier work, sources that do not check out, an essay that answers a nearby question rather than the one set. A tool can raise the question. It cannot settle it.

Does the similarity checker my university uses detect AI?

Several of the services universities submit through now report an AI-writing figure alongside the match report. Whether your institution has it switched on, who sees it, and what threshold turns it into a flag are questions for your department, and they are better asked before a meeting than during one.

Can a teacher tell if I paraphrased AI output?

Rewriting changes the surface statistics a detector reads, so a score can move. It does not change whether the essay engages with the seminar, whether the citations exist, or whether it sounds like you, which is what a marker looks at. And where AI is not permitted, the rule covers the draft however it was edited.

I wrote it myself and got flagged. What do I do?

Ask what tool produced the flag and what threshold it crossed, then bring what your process left behind: drafts, notes, sources, version history, and your ability to talk about the argument. False positives fall hardest on clean, careful and second-language writing, which is worth saying calmly and with the research to hand. There is a full script for that meeting at you wrote it and were accused.

Do teachers see which sentences were flagged?

It depends on the tool. Some highlight sentences, some report only a number. Highlighted sentences are more useful to both sides because they can be read and discussed. A bare percentage cannot be examined by anyone, including the person holding it.

Is Google Docs version history proof I wrote it?

It is a record of how the document was built, which is useful supporting context, but it is not proof of authorship. Gaps and large pastes are normal for people who draft elsewhere. Bring it as one piece of evidence alongside your notes, your sources and your account of the argument, rather than as something that closes the question on its own.

Keep reading