AI Writing Tools Built on Word Bans Are Already Losing the Detection Arms Race
A December 2024 study by Florida State University linguists identified RLHF fine-tuning — not training data or model architecture — as the reason ChatGPT overuses words like 'delve.' Separately, University of Tübingen researchers analysed over 15 million PubMed abstracts and estimated that at least 13.5% of 2024 biomedical papers show signs of LLM involvement, with words like 'meticulously' spiking 137% year over year. A follow-up FSU study complicates detection further, finding that AI-associated vocabulary is now spreading into everyday human speech via podcasts and YouTube. Despite this shifting landscape, most popular 'humanize my writing' tools on GitHub still rely on static banned-word lists, a method that becomes outdated whenever models or detectors evolve. The core problem is structural: no fixed word list can keep pace with the feedback loop between AI language models and the humans who increasingly consume and mirror their output.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in