Are AI detectors biased against non-native English writers?
Yes, there is good evidence that they can be. The best-known study found that popular detectors flagged more than half of essays written by non-native English speakers as AI-generated, while scoring essays by native-speaking students almost perfectly. If your students or colleagues write in a second language, treat a high score with extra caution.
Cite
Global 100 Forum, "Are AI detectors biased against non-native English writers?", https://forum.global100.org/q/are-ai-detectors-biased-against-non-native-english-writers/, accessed 2026-10-11.- Maya LindqvistStaffEdits the AI text detection and academic integrity sections ·
The study is Liang, Yuksekgonul, Mao, Wu and Zou, "GPT detectors are biased against non-native English writers", published in Patterns in 2023. The Stanford team ran seven detectors on two sets of human writing: TOEFL essays by non-native speakers and essays by US eighth-grade students. The detectors misclassified over half of the TOEFL essays as AI-written. The US student essays were classified almost entirely correctly.
Key figures from Liang et al., 2023
Writing sample Author group Share wrongly flagged as AI TOEFL essays Non-native English speakers More than half, across seven detectors US eighth-grade essays Native English speakers Near zero TOEFL essays after vocabulary enrichment Non-native English speakers Fell substantially, showing the detectors react to word choice Why it happens
Many detectors lean on perplexity, a measure of how predictable each next word is to a language model. Writers working in a second language often use a smaller vocabulary and more common sentence patterns. That makes their text more predictable, which is exactly what the detectors associate with machine output.
The researchers tested this directly. When they used a model to enrich the vocabulary of the non-native essays, the false positive rate fell. When they simplified the vocabulary of native-speaker essays, the false positive rate rose. The tools were reacting to linguistic sophistication, not authorship.
What to do with that
- Do not use a detector score as the basis for an accusation, and be especially careful where the writer is working in a second language.
- Ask for process evidence such as drafts, notes or version history, which do not depend on how polished the prose is.
- Check your tool's documentation for any published testing on non-native writing. Many vendors do not publish it.
- Read the wider accuracy picture in how accurate are AI text detectors, really before setting any policy on scores.
1 more reply
Most helpful first- Maya LindqvistStaffEdits the AI text detection and academic integrity sections ·
Some institutions reached the same conclusion from the policy side. When Vanderbilt University disabled Turnitin's AI detector in August 2023, one of the reasons it gave was that detectors are more likely to label text by non-native English speakers as AI-written. That is a useful document to share with a committee that is deciding how much weight to give a score.
Write something first.
Give people something to work with: at least 30 words on what happened and what you tried.
That is too long. Keep it under 6,000 characters.
Write the question as the title, 15 to 140 characters, no links.
Pick a category.
Add a name (2 to 40 characters, no links).
That email address does not look right.
That was quick. Read the thread, then try again.
The form expired. Reload the page and post again.
Something went wrong with the form. Reload and try again.
Please complete the check and post again.
Limit reached for now. Try again later.
This thread is closed to new replies.
This thread no longer accepts replies.
Something went wrong with the form. Reload and try again.
Post a reply or question first (name and email), then this browser can vote, edit and accept answers.
You cannot vote on your own post.
Only the author (within 30 days) or the forum team can do that.
That email belongs to a forum team account. Use your sign-in link instead.
New accounts are paused for the moment. Try again later.
That was already posted.