evidoria

← Back to browse

Good practice Imported

GPTZero — AI-Text Detection and Its Academic-Integrity Risks

United States of America · New York · See the United States of America profile · See the New York profile

Evidence: Descriptive / self-reported Top 95% 20/100 · Ask Evidence Copilot about this practice

Launched in 2023, GPTZero scaled to 4 million users by 2024 for flagging AI-written student text. An independent 14-tool benchmark found such detectors 'neither accurate nor reliable,' and a 2025 lawsuit alleges a Yale student was wrongly suspended after a false positive.

none exceeded 80% accuracy; only five exceeded 70% accuracy
Detection accuracy across 14 tools tested (benchmark study) (2023)
30,000 uses
Users (first week) (first week, 2023)
over 1 million users
Users (mid-2023) (mid-2023)
roughly 4 million users
Users (July 2024) (July 2024)
3.5 USD million
Seed funding raised (2023)
10 USD million
Series A funding raised (2024)

Details

Promoter
GPTZero, Inc.
Period
2023–ongoing
Keywords
AI content detection, academic integrity, plagiarism detection, false positives, EdTech

Context

GPTZero was created by Princeton undergraduate Edward Tian in January 2023 as one of the first widely used tools for detecting AI-generated text in student submissions, using "perplexity" and "burstiness" text statistics.

Activities

The tool grew explosively: 30,000 uses in its first week, over 1 million users by mid-2023, and roughly 4 million users by July 2024, with a partnership announced with the American Federation of Teachers in October 2023. The company raised $3.5 million in seed funding (2023) and $10 million in a Series A (2024), and was later acquired by Superhuman.

Results

A peer-reviewed benchmark of 14 detection tools (Weber-Wulff et al., International Journal for Educational Integrity, 2023) found that none exceeded 80% accuracy and only five exceeded 70%, concluding that "the available detection tools are neither accurate nor reliable," with performance degrading further once text is lightly paraphrased. Press investigations (Futurism, Washington Post, Ars Technica) reported teachers relying on GPTZero to falsely accuse students — in one case nearly 20% of a class — including flagging the U.S. Constitution itself as AI-written. In February 2025, a Yale School of Management EMBA student sued the university alleging he was wrongly suspended for a year based on a GPTZero flag; a federal judge denied his injunction request in May 2025.

Conclusions

GPTZero is included as a documented cautionary case: it shows both how quickly an AI-detection tool can scale into near-universal classroom use, and how thin the underlying evidence base for that use can be, with real, litigated consequences for students. Institutions using such tools without independent verification of accuracy risk exactly this outcome.

Implementation

Indicative cost
High (€500k–€5M) — Commercial product; the company raised $3.5 million in seed funding (2023) and $10 million in a Series A (2024) before being acquired by Superhuman.
Time to results
Medium (1–3 years) — Launched January 2023; grew to roughly 4 million users by July 2024; still generating litigation as of May 2025.
Staffing & skills
GPTZero, Inc. (founded by Edward Tian)

Conditions for success

  • Partnership with the American Federation of Teachers (announced October 2023) to expand classroom adoption

Common failure modes

  • Peer-reviewed benchmark found none of 14 detection tools tested exceeded 80% accuracy, with performance degrading further after light paraphrasing
  • Press investigations documented false accusations of students, including flagging the U.S. Constitution as AI-written
  • A Yale EMBA student sued the university alleging wrongful year-long suspension based on a GPTZero flag (injunction denied May 2025, case ongoing)
  • Explicitly framed by the source as a cautionary case of a technology scaling faster than its evidence base

Commonly funded by

Own resources / municipal budget

Indicative funding routes for practices of this type — always check each programme's current calls and eligibility rules.

Do you run this practice? Claim it — verified implementers get a public contact pathway and can propose corrections.

Data sources

Where this practice's information was retrieved from, and when.

Similar practices you may find useful