Two simple prompts swung AI detector accuracy by up to 87 points, study finds
A study published on arXiv tested seven AI-content detectors against human-written essays and found that surface-level vocabulary changes dramatically skewed results. When researchers prompted ChatGPT to simplify word choices in eighth-grade essays, the average false-positive rate jumped from 5.19% to 56.65%, with no change to the actual content or authorship. In the reverse test, TOEFL essays written by non-native English speakers — initially flagged as AI-generated at a 61.22% rate — dropped to 11.77% after a prompt polished their vocabulary to sound more native. A third experiment showed that AI-generated text dressed up with literary language evaded detection in 87% of cases. The findings, based on data from early 2023, suggest these detectors respond primarily to lexical patterns rather than true authorship, raising concerns about their use in academic misconduct proceedings.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in