False positive guide

Your essay got flagged as AI—and you actually wrote it.

This happens more often than most students realise, and a detector score is not proof of anything. Here is why genuine human writing gets flagged, how to respond if you are accused, and how to find the specific sentences that set the detector off.

What to do first

If you have been accused, the order of these steps matters. Evidence that you built the draft over time is far more persuasive than arguing about the score itself.

Pull your version history

Google Docs (File → Version history) and Word both record your draft being written over hours or days. This is the single strongest piece of evidence you have, so retrieve it before anything else.

Learn what triggered it

Find out which specific sentences read as machine-like. A single overall percentage tells you nothing useful and gives you nothing to discuss.

Ask for a conversation

Offer to talk through your argument, sources, and choices. A student who wrote the work can explain why they structured it that way; that discussion is usually more convincing than any tool.

Why AI detectors flag human writing

AI detectors do not identify authorship. They estimate how predictable a passage is—roughly, how often each word is the word a language model would have chosen next. Writing that is clear, tidy, and consistent tends to be predictable, which is exactly what a detector reads as machine-generated.

That produces an unfair result: the more carefully you edit and proofread, the more likely you are to be flagged. Students who write in a plain, organised style, and writers whose first language is not English and who lean on standard sentence patterns, are both disproportionately affected. Studies of detector accuracy have repeatedly documented false positives, and several universities have scaled back or dropped automated AI detection for this reason.

The patterns that commonly trigger a flag

  • Sentences that are all roughly the same length, giving the text an even rhythm.
  • No contractions—writing "do not" and "it is" throughout instead of "don't" and "it's".
  • Repeated sentence openings, such as several paragraphs beginning "This demonstrates" or "Furthermore".
  • Formal synonyms in place of ordinary words: "utilise" for "use", "numerous" for "many".
  • Structure so clean that there are no asides, hedges, or personal specifics anywhere.

None of these mean you did anything wrong. They are simply habits that overlap with how models write, and most are things you were probably taught to do in academic writing.

Why a single percentage is not enough

Most detectors return one number for a whole document. If you are told your essay is "78% AI", you cannot tell whether that reflects one formulaic paragraph or the entire piece, and you have no way to respond to the accusation specifically.

A sentence-level view is more useful. Penlify scores each sentence separately and colour-codes them, so you can see that—for example—your introduction and conclusion read as formulaic while the analytical middle section reads as clearly human. That gives you something concrete to point at, and shows you which habits to vary in future drafts.

To be clear about what this does not do: no tool can guarantee a particular detector outcome, and no tool can prove authorship. What a sentence-level breakdown gives you is an explanation instead of a verdict.

If you want to reduce the risk next time

Vary your sentence lengths deliberately. Let some contractions through where the register allows. Include specific evidence, named examples, and your own terminology rather than generic phrasing. Write in a document that keeps version history. None of this is about evading review—it is ordinary advice for writing that sounds like a person, and it happens to make a false positive less likely.

Common questions

Can an AI detector be wrong?
Yes. Detectors return a statistical estimate, not proof. They measure how predictable your word choices are, and clear, well-structured human writing often scores as predictable. False positives are well documented, and are more common for writers whose first language is not English.
How do I prove I wrote my essay myself?
Version history is the strongest evidence—Google Docs and Word both record the draft being built over time. Keep your notes, outlines, and sources, and offer to discuss the argument in person. Someone who wrote the work can explain the decisions behind it.
Should I rewrite my essay to lower the score?
Not if you are already accused—changing the document afterwards can look worse. Focus on evidence of authorship instead. For future drafts, varying sentence rhythm and adding specific detail is reasonable editing advice regardless of detectors.
Does Penlify guarantee a "human" result?
No, and be sceptical of any tool that promises this. Detector outcomes vary between tools and change as they update. Penlify shows you which sentences read as machine-like and why, so you can make an informed edit—it does not promise a score.