Supported AI detectors skill

Five primary detectors plus optional extras.

by sergebulaev·MIT license·★ 3,059 Stars on the repo·GitHub ↗

Use now

Files of Supported AI detectors

sergebulaev/main1 file
detector-list.md
Show the full text130 lines

Supported AI Detectors

Last updated: 2026-04-25

Five primary detectors plus optional extras. Each entry covers: API endpoint, auth, known accuracy issues, and the citation that documents the issue.

Contents

    1. GPTZero
    1. Originality.ai
    1. ZeroGPT
    1. Sapling
    1. Copyleaks
  • Optional / extended detectors
  • Why the spread matters
  • Quick stats to drop in a reply

1. GPTZero

Known issues:

  • Stanford study (Liang et al. 2023) included GPTZero in the cohort that flagged 61.3% of TOEFL essays from non-native English writers as AI. ESL bias is documented and reproducible.
  • Inflates scores on technical / dense prose regardless of authorship.
  • Will not run on text under 250 characters; gives unstable scores under 100 words.

Citation: Liang, W., Yuksekgonul, M., Mao, Y., Wu, E., & Zou, J. (2023). "GPT detectors are biased against non-native English writers." Patterns, 4(7). https://doi.org/10.1016/j.patter.2023.100779


2. Originality.ai

Known issues:

  • Marketed as "99% accurate" but multiple independent tests put real-world accuracy in the 60-80% range.
  • Aggressively flags any text that has been edited by Grammarly or similar tools, since editing patterns mimic LLM patterns.
  • Sergey's team meeting test (2026): scored a hand-written article 100% AI while GPTZero scored the same article 82% and ZeroGPT scored 50%. 50-point spread on identical text.

Citation: Internal CCC team test, March 2026 meeting transcript (projects/coactor/transcripts/); also referenced in Sergey Bulaev's April 2026 LinkedIn post on detector unreliability.


3. ZeroGPT

Known issues:

  • Famously unstable — the same input pasted twice 30 seconds apart can return scores 20+ points apart.
  • Flags US Constitution, Bible verses, and Declaration of Independence at 90%+ AI when pasted as plain text.
  • Susceptible to trivial paraphrasing — adding two typos drops a 95% score to 30%.

Citation: Multiple replicated demos on Twitter/X 2023-2024; Vanderbilt University communication on disabling Turnitin (Aug 2023) cited similar instability across the detector category. https://www.vanderbilt.edu/brightspace/2023/08/16/guidance-on-ai-detection-and-why-were-disabling-turnitins-ai-detector/


4. Sapling

Known issues:

  • Tends to score lower than GPTZero/Originality on the same text — useful as a contrarian signal in the parallel test.
  • Worse on creative writing than on technical prose.
  • Does not handle markdown — strip formatting before sending.

Citation: Sapling's own published benchmarks (https://sapling.ai/ai-content-detector/benchmark) acknowledge ~3-5% false positive rate even in their best-case dataset.


5. Copyleaks

Known issues:

  • Adelphi University used Copyleaks-style detector output as the sole evidence in the case that became Newby v. Adelphi University (Oct 2025). Federal court ordered the violation expunged.
  • Heavily penalizes formal academic writing regardless of authorship.
  • Unstable across re-submissions of the same text.

Citation: Newby v. Adelphi University, U.S. District Court (E.D.N.Y.), October 2025. Coverage: Inside Higher Ed, "Court Orders University to Drop AI-Cheating Charge" (Oct 2025).


Optional / extended detectors

None of these have a free API, so the script cannot call them. Run them by hand and enter the scores with --manual; there is no --extra flag.

  • Turnitin AI Writing — disabled by Vanderbilt, Cambridge, others. No public API; institutional only.
  • Winston AI — https://gowinston.ai. Paid only.
  • Crossplag AI — https://crossplag.com. Paid only.
  • Writer.com AI Content Detector — free web UI, no API. Use --manual mode.
  • Scribbr AI Detector — free web UI, no API. Use --manual mode.

Why the spread matters

OpenAI shut down its own AI Text Classifier in July 2023 with this public statement: "low rate of accuracy" — internally measured at 26%. If the company that ships the model cannot reliably detect its own output, no third-party detector built on weaker signals can be trusted as ground truth.

Reference: OpenAI blog, "New AI classifier for indicating AI-written text" (Jan 31, 2023), updated July 2023 with discontinuation notice.


Quick stats to drop in a reply

  • 61.3% — TOEFL essays by ESL writers misclassified as AI by 7 detectors (Stanford 2023)
  • 5.1% — same detectors' false positive rate on US 8th-grade essays (Stanford 2023)
  • 26% — OpenAI's own classifier accuracy before shutdown (July 2023)
  • 50 points — spread observed on a single article in CCC team testing (2026)
  • 0 — number of US courts that have upheld a "detector said so" finding without corroborating evidence as of April 2026
1# Supported AI Detectors
2 
3Last updated: 2026-04-25
4 
5Five primary detectors plus optional extras. Each entry covers: API endpoint, auth, known accuracy issues, and the citation that documents the issue.
6 
7## Contents
8 
9- 1. GPTZero
10- 2. Originality.ai
11- 3. ZeroGPT
12- 4. Sapling
13- 5. Copyleaks
14- Optional / extended detectors
15- Why the spread matters
16- Quick stats to drop in a reply
17 
18---
19 
20## 1. GPTZero
21 
22- **Web**: https://gptzero.me
23- **API docs**: https://api.gptzero.me/v2/predict/text
24- **Auth**: `x-api-key` header. Free tier: 10k words/month. Paid from $9.99/mo.
25- **Returns**: `documents[0].class_probabilities.ai` (0.0-1.0) plus per-sentence breakdown.
26 
27**Known issues:**
28- Stanford study (Liang et al. 2023) included GPTZero in the cohort that flagged **61.3% of TOEFL essays** from non-native English writers as AI. ESL bias is documented and reproducible.
29- Inflates scores on technical / dense prose regardless of authorship.
30- Will not run on text under 250 characters; gives unstable scores under 100 words.
31 
32**Citation**: Liang, W., Yuksekgonul, M., Mao, Y., Wu, E., & Zou, J. (2023). "GPT detectors are biased against non-native English writers." *Patterns*, 4(7). https://doi.org/10.1016/j.patter.2023.100779
33 
34---
35 
36## 2. Originality.ai
37 
38- **Web**: https://originality.ai
39- **API docs**: https://docs.originality.ai/
40- **Auth**: `X-OAI-API-KEY` header. No free tier — $0.01 per 100 words minimum.
41- **Returns**: `score.ai` (0.0-1.0), `score.original` (0.0-1.0).
42 
43**Known issues:**
44- Marketed as "99% accurate" but multiple independent tests put real-world accuracy in the 60-80% range.
45- Aggressively flags any text that has been edited by Grammarly or similar tools, since editing patterns mimic LLM patterns.
46- Sergey's team meeting test (2026): scored a hand-written article **100% AI** while GPTZero scored the same article 82% and ZeroGPT scored 50%. 50-point spread on identical text.
47 
48**Citation**: Internal CCC team test, March 2026 meeting transcript (`projects/coactor/transcripts/`); also referenced in Sergey Bulaev's April 2026 LinkedIn post on detector unreliability.
49 
50---
51 
52## 3. ZeroGPT
53 
54- **Web**: https://www.zerogpt.com
55- **API docs**: https://api.zerogpt.com/api/detect/detectText
56- **Auth**: `ApiKey` header. Free tier: 5 requests/min. Paid plans available.
57- **Returns**: `data.fakePercentage` (0-100 integer), `data.isHuman` boolean.
58 
59**Known issues:**
60- Famously unstable — the same input pasted twice 30 seconds apart can return scores 20+ points apart.
61- Flags US Constitution, Bible verses, and Declaration of Independence at 90%+ AI when pasted as plain text.
62- Susceptible to trivial paraphrasing — adding two typos drops a 95% score to 30%.
63 
64**Citation**: Multiple replicated demos on Twitter/X 2023-2024; Vanderbilt University communication on disabling Turnitin (Aug 2023) cited similar instability across the detector category. https://www.vanderbilt.edu/brightspace/2023/08/16/guidance-on-ai-detection-and-why-were-disabling-turnitins-ai-detector/
65 
66---
67 
68## 4. Sapling
69 
70- **Web**: https://sapling.ai/ai-content-detector
71- **API docs**: https://sapling.ai/docs/api/aidetect
72- **Auth**: `key` field in JSON body. Free tier: 50 requests/day.
73- **Returns**: `score` (0.0-1.0), per-sentence `sentence_scores`.
74 
75**Known issues:**
76- Tends to score lower than GPTZero/Originality on the same text — useful as a contrarian signal in the parallel test.
77- Worse on creative writing than on technical prose.
78- Does not handle markdown — strip formatting before sending.
79 
80**Citation**: Sapling's own published benchmarks (https://sapling.ai/ai-content-detector/benchmark) acknowledge ~3-5% false positive rate even in their best-case dataset.
81 
82---
83 
84## 5. Copyleaks
85 
86- **Web**: https://copyleaks.com/ai-content-detector
87- **API docs**: https://api.copyleaks.com/documentation/v3/writer-detector/submit
88- **Auth**: 2-step. POST to `/v3/account/login` with email + key, get bearer token, then POST to `/v2/writer-detector/{scanId}/check`.
89- **Returns**: `summary.ai`, which Copyleaks has documented both as a 0-1 probability and as a
90 0-100 percentage. `test_detectors.py` normalises either shape and rejects anything that still
91 falls outside 0-100. Per-paragraph breakdown alongside it.
92 
93**Known issues:**
94- Adelphi University used Copyleaks-style detector output as the sole evidence in the case that became *Newby v. Adelphi University* (Oct 2025). Federal court ordered the violation expunged.
95- Heavily penalizes formal academic writing regardless of authorship.
96- Unstable across re-submissions of the same text.
97 
98**Citation**: *Newby v. Adelphi University*, U.S. District Court (E.D.N.Y.), October 2025. Coverage: Inside Higher Ed, "Court Orders University to Drop AI-Cheating Charge" (Oct 2025).
99 
100---
101 
102## Optional / extended detectors
103 
104None of these have a free API, so the script cannot call them. Run them by hand and enter the
105scores with `--manual`; there is no `--extra` flag.
106 
107- **Turnitin AI Writing** — disabled by Vanderbilt, Cambridge, others. No public API; institutional only.
108- **Winston AI** — https://gowinston.ai. Paid only.
109- **Crossplag AI** — https://crossplag.com. Paid only.
110- **Writer.com AI Content Detector** — free web UI, no API. Use `--manual` mode.
111- **Scribbr AI Detector** — free web UI, no API. Use `--manual` mode.
112 
113---
114 
115## Why the spread matters
116 
117OpenAI shut down its own AI Text Classifier in July 2023 with this public statement: "low rate of accuracy" — internally measured at 26%. If the company that ships the model cannot reliably detect its own output, no third-party detector built on weaker signals can be trusted as ground truth.
118 
119Reference: OpenAI blog, "New AI classifier for indicating AI-written text" (Jan 31, 2023), updated July 2023 with discontinuation notice.
120 
121---
122 
123## Quick stats to drop in a reply
124 
125- **61.3%** — TOEFL essays by ESL writers misclassified as AI by 7 detectors (Stanford 2023)
126- **5.1%** — same detectors' false positive rate on US 8th-grade essays (Stanford 2023)
127- **26%** — OpenAI's own classifier accuracy before shutdown (July 2023)
128- **50 points** — spread observed on a single article in CCC team testing (2026)
129- **0** — number of US courts that have upheld a "detector said so" finding without corroborating evidence as of April 2026
130 

Discussion