GPTZero flagged your essay, but you wrote it yourself — here's what that means
By ESL Humanizer editorial teamUpdated 7 min read
Short answer
GPTZero estimates how likely a text is to be AI-written, partly by measuring perplexity (how predictable each word is) and burstiness (how much that varies between sentences). Careful non-native writing uses common words and even sentence lengths, which scores as predictable — so genuine ESL work can be flagged. A GPTZero result is a probability, not proof.
GPTZero is one of the most widely used AI detectors in schools. If it has marked an essay you wrote yourself as AI-generated, you are not the first — and the reason is usually the way the detector measures text, not anything you did wrong. This guide explains how GPTZero reaches its verdict, why non-native English is especially exposed, and what to do next.
How GPTZero scores text
GPTZero has described two signals at the heart of its approach (GPTZero):
- Perplexity — how surprising each word is to a language model. Low perplexity means the model could easily have predicted your next word.
- Burstiness — how much that predictability varies across sentences. Human writing tends to mix short and long, plain and unusual sentences; AI text is often evenly smooth.
GPTZero says its current model uses these as some of several indicators, alongside a deep learning classifier. The output is a probability that the text was AI-written — not a record of how it was written.
Why ESL writing looks “predictable”
When you write in a second language, you tend to choose words you are sure of and sentence patterns you have practised. That is good writing strategy — but it produces exactly the low-perplexity, low-burstiness profile detectors associate with AI. In the Stanford study of seven detectors, human-written TOEFL essays were misclassified as AI 61.22% of the time on average, while essays by US eighth-graders were classified almost perfectly (Liang et al., 2023).
BeforeTechnology is very important in our life. It helps people in many ways. It makes communication easy. It also makes work fast.
AfterTechnology shapes almost every part of daily life — from how we talk to family abroad to how quickly a small team can finish a project.
Four short sentences built on the same pattern read as highly predictable. Varying rhythm and adding specific detail is more natural English — and more you.
The researchers also found that when the same essays were rewritten with richer, more varied word choices, far fewer were misclassified. The problem is the style of simple English, not the honesty of the writer. We explain the mechanism in depth in why AI detectors flag human writing.
How accurate is GPTZero?
No detector is accurate enough to serve as proof. An independent evaluation of 14 tools, GPTZero among them, found every one scored below 80% accuracy (Weber-Wulff et al., 2023). OpenAI withdrew its own classifier after it mislabeled human text as AI 9% of the time (OpenAI, 2023). See every figure in our AI detector statistics roundup.
What to do if you're flagged
- Stay calm and ask for the result. Ask which tool was used, what score it gave, and which passages were highlighted.
- Collect your process evidence. Google Docs or Word version history, earlier drafts, outlines, notes, browser history for your sources.
- Offer to explain your work. Walking your instructor through your argument, or writing a short passage under supervision, tests authorship directly.
- Cite the research. Point to the documented bias against non-native writers — calmly, as context, not as an attack.
- Put it in writing if needed. Our appeal letter template gives you a structure.
The full process is in falsely accused of using AI? What to do.
Lowering the risk next time
Write in a tool that saves version history, keep your notes, and revise the sentences that sound translated. Varying your sentence rhythm and adding concrete details from your own experience is better English and less “predictable” to a model. Our seven habits that prevent false flags go further.
Make your own writing read naturally
Paste a paragraph you wrote. ESL Humanizer fixes unnatural phrasing, keeps your meaning, highlights every change and shows a detection estimate — free for one paragraph.
Quick answers
Is GPTZero accurate?
It is more accurate on long, clearly AI-generated text than on short or mixed text. Independent testing (Weber-Wulff et al., 2023) found no detector it tested, GPTZero included, reached 80% accuracy. No detector score should be treated as proof on its own.
Why does GPTZero say my writing is AI when I wrote it?
Low perplexity and low burstiness — predictable word choices and similar sentence lengths — look machine-like to the model. Second-language writers often write this way on purpose to avoid mistakes.
What should I do if my teacher used GPTZero?
Ask to see the result, stay calm, and bring evidence of your writing process: version history, drafts, notes and sources. Offer to explain your work in person. Our step-by-step guide covers the rest.
Sources
- GPTZero (2023). What is perplexity & burstiness for AI detection?
- Liang, Yuksekgonul, Mao, Wu & Zou — Patterns (Cell Press) (2023). GPT detectors are biased against non-native English writers
- Weber-Wulff et al. — International Journal for Educational Integrity (2023). Testing of detection tools for AI-generated text
- OpenAI (2023). New AI classifier for indicating AI-written text