Key takeaways

  • This is The Stepback, a weekly newsletter breaking down one essential story from the tech world.
  • These tools work by comparing a written work against a database filled with content from across the web, scholarly articles, and more to…
  • Between potential false positives and uncertainty about whether work was duplicated intentionally, some educators have backed away from…

What happened

This is The Stepback, a weekly newsletter breaking down one essential story from the tech world. For more news about how AI is changing our daily lives, follow Emma Roth. The Stepback arrives in our subscribers’ inboxes at 8AM ET. Opt in for The Stepback here. Long before ChatGPT became a thing, educators and editors frequently used anti-plagiarism tools to see if writers were being honest about their work.

In February, a student at Adelphi University won a lawsuit against the school after his professor similarly claimed he used AI to write an essay. Though the lawsuit doesn’t say which AI tool the professor used to examine the student’s essay, Adelphi University has a licensing agreement with Turnitin.

As pointed out by the University of California, Los Angeles, AI detection tools are trained to pick up on patterns that could indicate AI use, such as repetitive terms and phrases, text that sounds too formal or informal, and nonsensical phrasing.

Some, like QuillBot, also measure the “unpredictability” of text, as “AI tends to make the most ‘obvious’ or most common language choices as compared with human-produced writing,” according to UCLA. They may also look for sentence structure that remains the same throughout as another sign of AI. But these measurements aren’t indicative of AI on their own, as some people may just have a writing style with these qualities.

Even though Turnitin touts low false positive rates, it maintains that its tool “may not always be accurate” and shouldn’t be used to take actions against a student. ” OpenAI even shut down its own AI writing detector in 2023 due to low accuracy. But AI writing accusations are still being flung across the web.

5 million followers across his social channels, Ozzy Osbourne’s son, Jack, accused journalist and Verge contributor Kat Tenbarge of using AI to write an article for Rolling Stone, while flaunting the results from an AI detector, Getsolved, as “proof” of his claim.

Why it matters

These tools work by comparing a written work against a database filled with content from across the web, scholarly articles, and more to check for matching sentences and phrases. Some, like Turnitin, offer a percentage that claims to illustrate how much of the student’s writing overlaps with other works.

Between potential false positives and uncertainty about whether work was duplicated intentionally, some educators have backed away from using that particular tool. But now, the hunt for copied content is evolving into a war on AI-generated work. As quickly as students have picked up ChatGPT, Google Gemini, and Microsoft Copilot, teachers have adopted so-called AI detectors just as fast.

A survey from the Center for Democracy and Technology found that 43 percent of sixth to 12th grade teachers in the US regularly used AI detectors between 2024 and 2025. Some universities already using Turnitin in their learning management systems found that the service automatically enabled AI detection when the tool launched in 2023.

Instead of comparing pieces of written content, AI detectors like GPTZero, Pangram, and the one created by Turnitin rely on their own AI models to guess whether something might not be human-written — a process that’s arguably even murkier than matching text on the web.

As noted by GPTZero, AI detectors use an algorithm to analyze a text’s wording, rhythm, and structure, as well as to pick up on patterns in length and tone that may be more common in AI-written text. This relatively subjective evaluation isn’t as solid as something you could back up by matching text online, and can get tripped up by writers who speak English as a second language.

Despite this, Turnitin has said that its AI detector falsely flags less than 1 percent of human-written content as AI, while Pangram claims its false positive rate is just 1 in 10,000. GPTZero claims it has a similarly low rate of mistaking human content for AI.

People online are already accusing each other of “sounding like AI,” but the ready availability of AI detection tools is only adding fuel to the LLM witch hunt. In some high-profile cases, AI writing accusations have directly impacted people’s livelihoods and reputations. Last month, the publisher Minotaur dropped a $2 million book deal over concerns that its author, Jerry Falade, used AI — something he vehemently denies.

There’s Thierry Rignol, a French national who sued Yale last year after a professor accused him of writing portions of his final exam with AI, resulting in a failing grade and a one-year suspension. The professor used GPTZero to scan Rignol’s writing for signs of AI, but the lawsuit argues that “AI surveillance and detection tools are known to unfairly target non-native English speakers” like Rignol.

What to watch

Tenbarge has refuted the claim in a video and a post on her website, but Osbourne hasn’t retracted his accusation or deleted the video, leaving Tenbarge to deal with the trolls. Instead of relying on tools to weed out AI, many schools are encouraging educators to rethink their lessons.

The University of Chicago, for example, suggests telling students to slow down their reading, breaking up longer writing assignments, and requiring students to reflect on their work.