Can Turnitin Detect ChatGPT?
As artificial intelligence (AI) language models like ChatGPT gain popularity for generating human-like text, educators have raised concerns whether these tools enable students to easily bypass plagiarism checks.
So how well can Turnitin, one of the most widely used plagiarism detectors, identify essays and assignments written by AI? Let‘s take a comprehensive look.
The Rise of ChatGPT in Academia
ChatGPT burst onto the scene in late 2022 as a remarkably capable AI chatbot from Anthropic built on OpenAI‘s GPT-3 family of large language models.
GPT-3 and its successors like GPT-3.5 are trained on massive text datasets using deep learning to generate shockingly human-like responses on any topic.
Among ChatGPT‘s many capabilities, its prowess at crafting high-quality text on demand has led to exponential growth in use cases like content writing, coding, and essay writing.
In fact, usage has skyrocketed from 1 million users in the first week after launch to over 100 million monthly active users as of February 2023.
With students able to simply prompt ChatGPT to generate entire essays fitting requirements, academic integrity has emerged as a major concern.
Surveys suggest over 30% of students have used or plan to use ChatGPT for assignments, presenting a challenging new form of "assisted writing" for educators to address.
ChatGPT‘s Pros and Cons for Writing
Let‘s examine some of ChatGPT‘s key benefits and drawbacks when students improperly rely on it for academic work:
Pros
- Saves time by instantly generating paragraphs and entire essays on any prompt.
- Produces excellent grammar with few spelling or syntax errors.
- Can intelligently expand on points and generate "fake" citations and bibliographies.
- Obscures copying vs original work as the underlying GPT-3 model synthesizes responses from its training data.
Cons
- Lacks deeper understanding of prompted topics and often generates vague, inaccurate, or logically flawed responses.
- Output lacks original analysis and fails to demonstrate critical thinking.
- The constrained length and formulaic nature of essays is apparent.
- Citations and sources are fabricated rather than properly researched.
Here is a sample excerpt from an essay prompt on WWII history generated by ChatGPT:
"World War 2 began on September 1st, 1939 when Nazi Germany invaded Poland. This marked the start of the deadliest conflict in human history, with over 60 million lives lost across major theaters in Europe and the Pacific. Some key events include the Battle of Britain in 1940, Pearl Harbor in 1941, D-Day in 1944, and the atomic bombings of Hiroshima and Nagasaki in 1945. The war finally ended on September 2nd, 1945 with the formal surrender of Japan onboard the USS Missouri."
While well-written, this purely factual recollection lacks any original analysis or nuance expected in a historical essay. Yet on the surface it may able to fool both students and educators into thinking it is a human-written piece.
Challenges in Manual AI Detection
Before recent AI detector updates, instructors had difficulty reliably identifying student essays written by ChatGPT. Some key challenges include:
-
Human-like writing: With its training on diverse writing datasets, ChatGPT mimics much of the vocabulary, style, and coherency of an average human writer.
-
Knowledge synthesis: ChatGPT excels at recombining facts and data from its training corpus to appear knowledgeable on prompted topics, often beyond a typical student‘s capability.
-
New compositions: Rather than copy-pasting, ChatGPT generates unique compositions on the fly, making plagiarism checks ineffective.
-
Length flexibility: Essays can be produced at whatever length specified in the prompt, avoiding a common red flag.
-
Citation fabrication: ChatGPT can synthesize fake bibliographies and citations that appear legitimate on superficial inspection.
Instructors were left resorting to extensive manual verification such as oral examinations, multiple drafts, and scrutinizing writing tone – a losing battle as AI capabilities rapidly advance.
Plagiarism detectors lagged in developing the natural language processing capabilities to distinguish human vs AI writing patterns. But Turnitin has stepped up with robust AI detection to address this need.
How Does Turnitin Detect ChatGPT?
Turnitin relies on an advanced natural language processing model trained to identify telltale indicators of machine-generated text rather than natural human writing styles and reasoning.
Specifically, Turnitin‘s AI classifier examines submitted text across factors like:
-
Word choice: Humans exhibit originality and randomness, while AIs favor repeated words and phrases from their training data.
-
Sentence structure: AI text follows templated grammatical patterns, lacking syntactic variety.
-
Topic focus: AIs stick closely to the prompt with few logical leaps or semantic connections.
-
Coherence: AI essays often have fragmented flow and lack overall continuity.
-
Fluency: Unlike humans, AIs avoid typos, ambiguity, colloquial usage, and other imperfections.
By analyzing these linguistic factors across paragraphs and the full document, Turnitin assigns an "AI-generated text" percentage score.
For example, here is a snippet of a Turnitin report flagging an essay as 79% likely AI-written:

Anything above 25% indicates considerable AI contribution and should be verified manually.
Notably, Turnitin‘s AI detection works entirely separately from its plagiarism and similarity checks. Those identify text copied from existing sources, while AI detection focuses strictly on writing characteristics.
Compare to Other Plagiarism Detectors
Turnitin is relatively ahead of competitors in deploying capable AI writing analysis according to market evaluations:
| Plagiarism Checker | AI Detection Release | Accuracy |
|---|---|---|
| Turnitin | April 2022 | 98% |
| Grammarly | June 2022 | 63% |
| SafeAssign | July 2022 | 81% |
| Unicheck | November 2022 | 93% |
Turnitin‘s AI classifier model examines writing patterns in greater linguistic depth, enabling higher precision in distinguishing human vs AI authors.
The classifier is further strengthened by Turnitin‘s vast repository of real student paper submissions, unlike competitors relying on more synthetic training data.
Continuous model retraining will be critical as new AI writing styles emerge to maintain Turnitin‘s edge in detection accuracy.
Limitations and Evasion Attempts
Despite Turnitin‘s advanced AI detection capabilities, there are still limitations and potential ways students may try improperly evading detection:
-
Heavily editing AI output to introduce more variety, errors, and randomness. This is extremely time intensive.
-
Using multiple AI tools in conjunction to create mixed writing that may obscure individual patterns. But results in incoherent essays.
-
Introducing multi-step processes like generating paragraphs separately then compiling. But integration remains formulaic.
-
Purposefully generating off-topic or nonsensical text to mislead detectors. But produces low-quality work.
Overall, strategies focused solely on evasion typically lower essay quality and fail to demonstrate true understanding, undermining any dubious benefits.
Maintaining writing integrity as AI capabilities advance will require continuous detector innovation plus clear policies and education around ethics and responsible usage.
The Future AI Detection Arms Race
Today‘s plagiarism detectors are only calibrated to reliably identify writing from existing AI tools like GPT-3.5-based ChatGPT which have known limitations.
As OpenAI and competitors race to create ever-more capable generative AI models like GPT-4, detection systems will need to up their game.
Adversarial attacks to purposefully mislead detectors are also rising as an arms race akin to spam filters. This may require more holistic semantic analysis versus pattern matching.
Equally challenging will be identifying "human-in-the-loop" writing where students heavily augment AI writing. Distinguishing human enhancement from raw machine output will only get harder.
Is the cat and mouse game between cheating and detection futile? Or will the back and forth drive progress in AI accountability? Only time will tell.
For now, Turnitin provides capable protection against current AI misuse. But continued innovation and responsible policies will be essential as new models like ChatGPT itself inevitably emerge.
Recommendations for Instructors
Educators have a crucial role to play in responsibly leveraging Turnitin and other detectors while promoting academic integrity and original thinking.
Some tips include:
-
Familiarize yourself thoroughly with Turnitin‘s AI detection capabilities, limitations, and reports.
-
Frame assignments emphasizing analysis and critical views over basic fact collection which AIs excel at parroting.
-
Use multiple draft reviews, citations checks, and oral exams to verify student work beyond just initial submission.
-
Update academic integrity policies with clear repercussions for AI misuse while supporting fair and beneficial applications.
-
Discuss AI ethics with students and reinforce that their learning process, not outcomes, is what matters.
Overall, a holistic approach is needed to uphold standards in an increasingly AI-integrated world.
Conclusion
In conclusion, Turnitin‘s newly released AI text detector offers powerful capabilities to identify writing from ChatGPT and other generative language models with a high degree of accuracy.
By flagging AI-generated passages and their likelihood scores, Turnitin empowers educators to catch improper AI use in student work and maintain academic integrity as AI proliferation gathers pace.
However, the ever-escalating race between AI capabilities and detection will require ongoing vigilance and innovation from all players to balance productivity and ethics in an AI-powered future.
What are your thoughts on Turnitin and other detector‘s capabilities today, and how should educators respond responsibly? Let‘s discuss!