Best AI Detectors, Scored: Every One Flags Human Writing Sometimes
The best AI detector depends on who is asking: GPTZero for teachers, Originality.ai for publishers auditing freelancers, Pangram for free daily checks, Turnitin if your school already pays for it. All six flag human writing as machine-written sometimes, so no number they produce should ever be the only evidence in an accusation.
Updated August 2026. Every paid price below was read off the vendor’s own pricing page on July 30, 2026, and read again on August 5 and August 6. Three reads in seven days, not one figure moved.
The six best AI detector tools at a glance
| Tool | Best for | Standout | Starting price (verified Aug 6, 2026) | Our score |
|---|---|---|---|---|
| GPTZero | Teachers and schools | Most consistent tool in the one independent test we found | Premium $12.99/mo annually; free 10,000 words/mo (last seen Aug 5) | 8.5/10 |
| Pangram | Free daily checks, API work | 2,000 words a day free, no card; per-word API pricing | Free; Individual $20/mo | 8.3/10 |
| Turnitin | Institutions already paying | Publishes its own error rate and says the score isn’t proof | No price published; institutions only | 8.0/10 |
| Copyleaks | AI plus plagiarism in one report | One scan, both checks, seven LMS platforms | Personal $16.99/mo, or $13.99 annually | 7.8/10 |
| Originality.ai | Publishers auditing freelancers | Site-wide scans, editorial workflow, top-tier API | Pro $14.95/mo, or $12.95 annually | 7.5/10 |
| Winston AI | Cheapest annual bill | $10/month annually, images and deepfakes included | Essential $18/mo, or $10 annually | 7.5/10 |
Scores are judgment on documented evidence, not a lab benchmark.
How we scored this, and the test we haven’t run
We did not run these detectors against a controlled corpus. Anyone claiming a hands-on accuracy ranking without publishing their test documents is asking you to take their word for it, and here that word is usually paid for.
Here is what we did do. All six vendors’ pricing pages, opened and copied down as written, Turnitin’s 404 included, on July 30, again on August 5, again on August 6. Turnitin’s AI-detection FAQ, read end to end. The University of Chicago’s comparison, the only independent institutional test we could find. The Stanford paper on detector bias that not one of those vendor pages cites. Then two Reddit threads, one of them from teachers who got it wrong.
The tie-breaker is the thing we can’t buy: how each tool behaves on writing we can prove a human produced. That needs an artifact we don’t have yet.
Ratings are ours. No vendor saw a draft and nobody paid to be here.
Which is the most accurate AI detector?
Nobody credible knows, and the honest answer is worth more than a fake one.
The best independent evidence we found comes from the University of Chicago’s Academic Technology Solutions team, which tested detectors against unedited output from ChatGPT, Microsoft Copilot, Claude and the university’s own PhoenixAI, plus human-composed text. Its verdict: “no tool is infallible,” and “each had one or more vulnerabilities.” UChicago’s IT Services offers no supported AI detector at all.
Two findings are load-bearing. GPTZero was “the most consistent of the AI detectors examined” — 100% accurate on every AI sample except Copilot, where it managed 63% confidence, and 99% accurate on human writing. Originality.ai caught the most, flagging all four models’ output with 100% certainty, and it also flagged human-composed text as AI with 97% certainty. Best detection and worst false positive in the same tool: that trade-off is the whole category. Read it with the date in mind — the spreadsheet is stamped April 2025 and covers neither Pangram, Copyleaks nor Turnitin.
Then do the multiplication nobody in a sales deck does. Turnitin is the most conservative tool here, and its own documentation targets a false-positive rate under 1% on documents it scores above 20% AI. Take that target at its word and run 400 submissions through it in a term: it still allows about four wrongly flagged students, every term, in one course. Treat the percentage as a prompt to go looking for evidence rather than as the evidence: keep version history and draft timestamps, and run two tools, because where they disagree tells you more than either score alone.
The Stanford finding no vendor page we read cites
In July 2023 the Cell Press journal Patterns published Liang, Yuksekgonul, Mao, Wu and Zou of Stanford under a title that hasn’t aged a day: “GPT detectors are biased against non-native English writers.”
They ran seven widely used detectors over 91 TOEFL essays by non-native speakers and 88 US eighth-grade essays. The eighth-graders came back almost perfectly classified. The TOEFL essays drew an average false-positive rate of 61.3%. Every detector agreed on 19.8% of them, and at least one detector flagged 97.8%.
The mechanism is unglamorous. Most detectors lean on text perplexity — how surprised a language model is by your next word. A limited vocabulary produces low perplexity. So does an LLM. The classifier cannot tell those apart, and the authors proved it ran both directions: asking ChatGPT to enrich the TOEFL essays’ word choice dropped the false-positive rate from 61.3% to 11.6%, while simplifying the American essays pushed their misclassification up.
Then the result that should end the argument. Prompting ChatGPT to “elevate the provided text by employing literary language” collapsed detection rates on its own AI-generated college essays to near zero.
Read those two findings together, because they are the same finding. The tool that misses the student who typed one extra sentence into the prompt box is the tool flagging the international student who wrote every word herself. The authors’ recommendation was blunt: “we strongly caution against the use of GPT detectors in evaluative or educational settings.”
The caveats, stated rather than buried. Patterns ran the piece as an opinion article, not as a research paper. It is from 2023 and predates every model these tools now target, and it never names which seven detectors it tested, so none of it attaches to any product in the table above. It is evidence about the method, and the method hasn’t changed. Turnitin says it trained against this bias specifically, sampling second-language learners and non-English-speaking countries. That’s the right answer. It’s also Turnitin’s account of Turnitin’s dataset.
Best for teachers and schools: GPTZero
GPTZero wins classrooms for an unglamorous reason: in the only independent institutional test we found, it was the tool least likely to accuse a human. It leans cautious by design, and its sentence-level highlighting lets a teacher look at what got flagged instead of arguing with a percentage.
Premium is $12.99 per month billed annually for 300,000 words. Professional is $24.99 and adds 250-file batches and LMS integration (both verified Aug 6, 2026). The free tier is real too: 10,000 words a month, no card. One caveat on that last one. It rendered on our July 30 and August 5 reads and did not render at all on August 6, so treat the free plan as last seen Aug 5 until someone confirms it from a US connection.
Skip it if you need one number you can act on. The caution is the feature, and more AI text slips through than with Originality.ai. If your problem is freelancers filing machine copy, buy the aggressive one.
Best free AI detector: Pangram
Pangram has the most generous genuinely free tier here: 2,000 words a day, three image scans, 20-plus languages, browser extension and Google Docs, no payment method required. For a teacher spot-checking two essays a week, that is the whole job for nothing.
Paid is $20 a month for 300,000 words, with $60 off yearly and a seven-day trial; Professional is $65 for 1.5 million. The API is the category’s most transparently priced: $0.05 buys 100 words on Pangram 4 (verified Aug 6, 2026). One thing to hold in view: the “30 tools tested” comparison at organic #3 for this query is Pangram’s own blog post — content marketing by a competitor in the ranking.
Skip it if you need an institutional audit trail today. No independent test we found evaluates Pangram, so you’d be buying on the vendor’s own numbers.
The one that actually decides your grade: Turnitin
If you are a student, the only detector whose output matters is the one wired into your school’s LMS, and you can’t buy it or appeal to it directly. Turnitin publishes no price — turnitin.com/pricing returned a 404 on Jul 30 and again on Aug 6, 2026 — and it sells AI writing detection only with a Turnitin Originality licence or an iThenticate 2.0 add-on, through an account manager (re-confirmed Aug 5, 2026). Only instructors see the indicator.
It earns its score for candour. Its documentation, re-stamped August 4, 2026, targets a false-positive rate “under 1% for documents with over 20% of AI writing,” validated before every model release against 700,000 pre-ChatGPT papers, and buys that with deliberate under-reporting: a 50% score “could contain as much as 65% AI writing.” It scores nothing between 1% and 19%, rendering those as an asterisk so a low-confidence hit can’t become an accusation. It needs 300 words of prose, won’t touch code or bullet lists, and says the percentage “should not be used as the sole basis for action.” A July 2026 update collapsed its model ensemble into a single model at the same claimed error rate. Every vendor here should publish that paragraph. One does.
Skip it if you are an individual or an agency: it isn’t sold to you, and no consumer detector reproduces its score.
Best for AI plus plagiarism in one pass: Copyleaks
One Copyleaks scan returns two things at once, in the same report: an AI score and a plagiarism match. AI detection covers 30-plus languages, plagiarism matching over 100. It reads images, OCRs handwritten work, and plugs into seven LMS platforms. A department already paying for two separate checks is buying one subscription here instead of two.
Personal is $16.99 a month, or $13.99 billed annually at $167.88. Pro is $99.99, or $74.99 annually, with 25 seats. Read the credits twice: one covers 250 words or one image, monthly Personal carries 100 a month, annual Personal carries 1,200 for the year — the same rate, presented so annual looks twelve times larger (verified Aug 6, 2026).
Skip it if you might change your mind. Copyleaks refunds only within ten days and only if you’ve used no credits, and switching plans forfeits whatever credits were left.
Best for publishers auditing freelancers: Originality.ai
Originality.ai is built for the person paying for words, not the person grading them. It scans whole sites by URL, keeps 365 days of history on the top tier, and bundles plagiarism, readability, grammar and fact-checking into one credit pool. Its Chrome extension replays the writing process inside a Google Doc.
Pro is $14.95 a month, or $12.95 billed annually, for 2,000 credits at 100 words each. Enterprise is $179, or $136.58 annually, for 15,000 credits and API access. Credits expire monthly, and neither plan has a free tier (verified Aug 6, 2026).
Skip it if you will use the score to confront a person. This is the tool UChicago found flagged human writing as AI with 97% certainty. Aggressive detection suits auditing a supply chain; it’s wrong for anything with consequences attached.
Cheapest annual bill: Winston AI
Winston AI is the value play, but only if you commit for a year. Annual billing runs $10 a month for 100,000 credits, $16 for 200,000, $26 for 500,000, and every tier carries image and deepfake detection, OCR and plagiarism checks. Month to month those same three tiers cost $18, $29 and $49. The free option is a 14-day trial, not a standing tier (verified Aug 6, 2026).
The score stops where it does for two reasons. Winston’s site claims 99.98% detection, a figure no independent test we found reproduces. And Google’s AI Overview tells searchers Winston’s “pricing starts around $12/month,” while Winston’s own page said $18 monthly and $10 annually on all three days we read it, most recently Aug 6, 2026. That is the machine you are trusting to shortlist tools for you.
Skip it if you want a long look before committing. Fourteen days is the thinnest free trial here.
What happens when a teacher trusts the number
The thread worth reading before you buy anything here is False positives from ai detection in education destroyed my relationship with three students on r/Teachers, posted 7 November 2025, at 516 points and 259 comments (read Aug 6, 2026). A teacher ran a batch of essays, got three hits above 95%, and opened the academic integrity process on all three. All three had drafts and peer-review comments. One cried in the teacher’s office. Parents called the principal.
The top reply, past 1,200 points, is a computer science teacher: “AI checkers are garbage, do not use them please.” The second, just under 600, is procedural instead of ideological: don’t go scorched earth before you talk to the student, ask to see the drafts, and accept that the drafts, the peer edits and the rest were things you should have known about already.
The most useful comment isn’t about software. One teacher asks the student to summarise their own argument aloud and define any word that looked like reaching. If they can do both, it’s theirs. If not, the assignment gets redone. No subscription required.
Read it with your guard up. Several commenters accuse the post itself of being AI-written, and one of them ran it through a detector and reported back: “your text has 92% chances of being written by AI. That should tell you a lot.” That is either the best joke in the thread or the entire problem restated, and the argument never resolves. We quote the named commenters, not the poster’s account of events, for exactly that reason.
Why we don’t cover humanizers or bypass tools
We rank detectors. We don’t rank, link to or recommend the tools built to defeat them, and the SERP is the argument for that line. The second organic result for “best ai detector” is a r/humanizeAIwriting thread (posted 22 Oct 2025, 148 comments, read Aug 5, 2026) whose number-one “detector” pick is a humanizer, praised for helping you “keep structure intact but still pass checks,” seconded in the top comment by an account repeating the same three selling points, with a discount code handed out further down. Google’s most visible detector recommendation is an evasion vendor in a review’s clothes.
Its comments carry the real cost, though. One user: “I pasted a poem I wrote in 2011 to the Walter white one and it said 76% this is probably AI written.” Another describes making “intentional typos to avoid that” — people roughing up their own prose so a classifier will read them as human. A third dismisses ZeroGPT as “a horrible detection tool, because it often flags human work as ai if it is written formally or too clear/consise.” Same reason it isn’t in our table.
There’s a conflict on the legitimate side too. The organic #1 result and first AI Overview citation is a Substack post titled “I Tested 30+ AI Detectors,” ranking Winston AI second. The same byline appears on three posts on Winston AI’s own blog, still listed on Winston’s pricing page when we looked again on Aug 6, 2026, now dated Jul 15, Jul 21 and Jul 30. The ranking position is from our Jul 30 SERP pull. We aren’t alleging the ranking was bought. We’re saying a reader can’t tell.
FAQ
What is an AI detector?
An AI detector scores text on how likely a large language model produced it. The same product gets marketed as an AI content detector, an AI text detector or an AI writing detector. All work statistically, on patterns like word predictability and sentence uniformity, never on meaning, which is why clean formal human writing gets flagged most often.
Which AI detector is the most accurate in 2026?
No independent 2026 benchmark covers the whole field. The most recent independent institutional test we found, from the University of Chicago, called GPTZero the most consistent and found Originality.ai caught the most AI while also flagging human text as AI with 97% certainty. Its own conclusion: no tool is infallible.
Is there a genuinely free AI detector?
Yes, two. Pangram gives 2,000 words a day with no payment method (verified Aug 6, 2026), and GPTZero gives 10,000 words a month, last confirmed Aug 5. Winston AI’s free option is a 14-day trial, not a standing tier, and Originality.ai has no free subscription plan at all.
How much do AI detectors cost?
Entry pricing runs $10 to $17 a month. Winston AI is cheapest annually at $10/month, GPTZero is $12.99/month billed annually, Originality.ai is $12.95 annually or $14.95 monthly, and Copyleaks is $13.99 annually or $16.99 monthly (all verified Aug 6, 2026). Turnitin publishes no price and sells to institutions only.
Can AI detectors be wrong?
Routinely, and non-native English writers bear most of it. A 2023 Stanford study in Patterns found seven detectors misclassified 61.3% of non-native English essays as AI-generated, against near-perfect accuracy on US student essays. The University of Chicago separately recorded a human paragraph scored 100% AI by one tool and 97% by another.
Do teachers actually use these tools?
The one that counts is usually Turnitin, because it’s already inside the LMS and only instructors see its indicator. GPTZero and Copyleaks are what individual teachers reach for outside institutional systems; Copyleaks integrates with seven LMS platforms including Canvas, Moodle and Blackboard.
Changelog
- 2026-08-06 (later) — Second QA pass. Four things this page was asserting more confidently than its own notes allowed have been pulled back to what was actually read. GPTZero’s free tier no longer sits under an “Aug 6 verified” stamp in the comparison table or the GPTZero section: it rendered on July 30 and August 5, did not render on August 6, and the page now says exactly that instead of quietly restamping it. The opening freshness line says “every paid price” because Turnitin publishes none and GPTZero’s free plan is the open question. The Turnitin arithmetic now prints the condition attached to the vendor’s own target — under 1% on documents it scores above 20% AI — so a reader can check the multiplication rather than take it. And the Stanford section no longer claims the finding is quoted by “no vendor page”, a claim about pages nobody audited; it now says no vendor page we read, meaning six pricing pages and Turnitin’s FAQ. That scoping also removes a contradiction with this page’s own Turnitin paragraph. Two stale reverification notes were repaired: the Winston conflict-of-interest item was still listing blog dates a later note had superseded and was implying both halves were read the same day when only the Winston half has been re-read, and one test artifact was still asking for a Turnitin stamp that no longer exists. Copyleaks and Winston spec paragraphs rewritten out of comma-chain form. No price, score or verdict changed.
- 2026-08-06 — Third verification pass, and it closed the oldest open item on this page. turnitin.com/pricing was finally re-tested and returns 404 again, a week after the first 404, so “publishes no price” now rests on two reads rather than one. All five published price sets were re-read on the vendors’ own pages for the third time in seven days and not one figure moved. Two things did move and are recorded rather than smoothed over. GPTZero’s pricing page rendered no free-plan card at all on this read, where earlier reads showed one, so the free-tier claim is now dated to Aug 5 in the FAQ instead of being restamped; confirm it from a US IP before publish. And the r/Teachers reply this page cites has now read 602, 600 and 598 points across three visits, so the body stops printing a hard number for it and says “just under 600” — a figure that drifts four points in a week has no business being stamped as precise. Winston’s blog module still carries the same byline the Substack ranking it second does, now dated Jul 15, 21 and 30. No price, score or verdict changed.
- 2026-08-05 (later) — Editorial and fact-check pass. Both Reddit threads re-opened live and every quote on this page checked against them. Two counts were wrong and are fixed: the r/humanizeAIwriting thread is at 148 comments, not 145, and the second-ranked r/Teachers reply sits at 600 points, not 602. One sentence claiming the thread’s ZeroGPT verdict “matches UChicago’s exactly” was cut, because this page never prints UChicago’s ZeroGPT result and the reader had no way to check it; a direct quote replaces it. The Liang et al. section no longer calls the study “peer-reviewed” and now says what it is, an opinion article in Patterns. Turnitin’s error-rate arithmetic was corrected from four wrongly flagged students a year to four a term, which is what a sub-1% rate on 400 submissions actually produces. Added: the commenter who ran the r/Teachers post itself through a detector and got 92% AI. No price, score or verdict changed.
- 2026-08-05 — Verification pass. All five published price sets re-read on the vendors’ own pricing pages; every figure unchanged from July 30, so all stamps moved to Aug 5. Turnitin’s FAQ re-read and now stamped Aug 4, 2026: same sub-1% false-positive claim, plus a newly documented July 2026 change consolidating its model ensemble into a single model, both added above. New section added on Liang et al., Patterns (Cell Press, July 2023) — the 61.3% non-native-speaker false-positive finding and the “literary language” bypass result — which this page previously cited no equivalent of; both caveats (2023 vintage, detectors unnamed) stated on-page. Reddit coverage rebalanced: a 259-comment r/Teachers thread on three wrongly accused students now leads, and the r/humanizeAIwriting thread was compressed into the no-bypass section it evidences. UChicago comparison re-checked — still the April 2025 spreadsheet, no refresh. Keyword cluster re-pulled now the report cap has reset, closing the July 30 open item: keyword_overview confirms “ai detector tools” at 9,900/mo and SD 22, but returns SD 62 for the singular, so the plural stays an H2 ramp rather than the target. No price, score or verdict changed.
- 2026-07-30 — Page created; all prices read live on the six vendors’ own pages, plus Turnitin’s confirmed-absent price (its /pricing returns 404). Turnitin’s AI-detection FAQ and the University of Chicago comparison read in full, the latter’s April 2025 vintage disclosed above. SERP recorded: the AI Overview fires with ten references, six vendor-owned, and misstates Winston AI’s entry price; the organic top eight holds no independent reviewer.
Independent, no paid placements. We verify every price on the vendor’s own site and stamp the date; we publish who should skip each tool; we do not cover detection-bypass products; we keep this changelog public. See how we score AI tools and our independence charter. More categories on the AI tools index, every verified price on the pricing index. Related: AI study tools, AI courses for teachers, the Jasper AI review and AI tools for HR teams.