False positive guide

Falsely flagged as AI? What to do if you wrote it.

A detector score is not proof of authorship. Here is how to respond fairly if original writing is questioned, which evidence of process to keep, and how to review a result without treating it as a verdict.

What to do first

If you have been accused, the order of these steps matters. Evidence that you built the draft over time is far more persuasive than arguing about the score itself.

Pull your version history

Google Docs (File → Version history) and Word both record your draft being written over hours or days. This is the single strongest piece of evidence you have, so retrieve it before anything else.

Learn what triggered it

Find out which specific sentences read as machine-like. A single overall percentage tells you nothing useful and gives you nothing to discuss.

Ask for a conversation

Offer to talk through your argument, sources, and choices. A student who wrote the work can explain why they structured it that way; that discussion is usually more convincing than any tool.

Why AI detectors flag human writing

AI detectors do not establish authorship. They produce a statistical estimate from text, using methods and thresholds that differ by tool and can change over time. A score can start a conversation or prompt a closer review, but it cannot show how a passage was created.

That limitation matters when an original draft is formal, concise, translated, or heavily edited. A reviewer should consider the writing itself alongside drafts, notes, sources, version history, and the writer's explanation. False positives and false negatives are both possible.

What to review after a flag

  • Sentences that are all roughly the same length, giving the text an even rhythm.
  • No contractions—writing "do not" and "it is" throughout instead of "don't" and "it's".
  • Repeated sentence openings, such as several paragraphs beginning "This demonstrates" or "Furthermore".
  • Formal synonyms in place of ordinary words: "utilise" for "use", "numerous" for "many".
  • Structure so clean that there are no asides, hedges, or personal specifics anywhere.

None of these mean you did anything wrong. They are simply habits that overlap with how models write, and most are things you were probably taught to do in academic writing.

Why a single percentage is not enough

Most detectors return one number for a whole document. If you are told your essay is "78% AI", you cannot tell whether that reflects one formulaic paragraph or the entire piece, and you have no way to respond to the accusation specifically.

A sentence-level view is more useful. Penlify scores each sentence separately and colour-codes them, so you can see that—for example—your introduction and conclusion read as formulaic while the analytical middle section reads as clearly human. That gives you something concrete to point at, and shows you which habits to vary in future drafts.

To be clear about what this does not do: no tool can guarantee a particular detector outcome, and no tool can prove authorship. What a sentence-level breakdown gives you is an explanation instead of a verdict.

Sources

How to prepare for a fair review next time

Keep your outline, notes, sources, and version history. Make your evidence and examples specific because they help readers assess the work on its merits. If a score is questioned, ask for a review of that process and the actual draft rather than treating a percentage as a verdict.

Common questions

Can an AI detector be wrong?
Yes. Detector results are statistical estimates, not proof of authorship. A score should be reviewed with drafts, sources, version history, and the writer's explanation before anyone reaches a conclusion.
How do I prove I wrote my essay myself?
Version history is the strongest evidence—Google Docs and Word both record the draft being built over time. Keep your notes, outlines, and sources, and offer to discuss the argument in person. Someone who wrote the work can explain the decisions behind it.
Should I rewrite my essay to lower the score?
Not if you are already accused—changing the document afterwards can look worse. Focus on evidence of authorship instead. For future drafts, varying sentence rhythm and adding specific detail is reasonable editing advice regardless of detectors.
Does Penlify guarantee a "human" result?
No, and be sceptical of any tool that promises this. Detector outcomes vary between tools and change as they update. Penlify shows you which sentences read as machine-like and why, so you can make an informed edit—it does not promise a score.