A recent independent test of 14 AI text detection services, including Turnitin and PlagiarismCheck, found that none achieved over 80% accuracy, with many failing to reliably distinguish between human-written and AI-generated content. OpenAI's own classifier was shut down due to low accuracy, and Turnitin acknowledged practical error rates higher than initially stated. Researchers advise extreme caution, as these tools often misidentify human writing, particularly from non-native speakers, and struggle with paraphrased or translated texts, potentially leading to unfair rejections of work. AI
IMPACT AI text detectors are unreliable, potentially leading to unfair accusations and rejections of work, impacting freelancers and students.
RANK_REASON The cluster discusses the findings of an independent test and research on the accuracy of AI text detection tools. [lever_c_demoted from research: ic=1 ai=1.0]
- arXiv
- Deborah Weber-Wulff
- Inside Higher Ed
- James Y Zou
- OpenAI
- PlagiarismCheck
- Science Daily
- TechCrunch
- TOEFL
- Turnitin
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →