Every figure here carries the date it was read and the source it came from. How scoring works →

Winston AI

ai detectors

Visit Winston AI

· affiliate links never affect the score

Vouch Score v11 · data collected 2 September 2026
43.8/100
low
  • Capability11.1
  • Usability & control76.5
  • Value65.2
  • Commercial terms66.5
Sources: https://arxiv.org/abs/2306.15666 · https://www.trustpilot.com/review/gowinston.ai · https://gowinston.ai/pricing · https://gowinston.ai/terms · https://gowinston.ai/. Recomputable from the committed inputs; see methodology.

By Minel Gunesoglu, founder. I submitted nothing to Winston AI and ran no detection tests. What I did was read the academic paper that did run them, in the source PDF rather than in anybody's summary of it, and take its per-tool table rows off the page. I also read the vendor's pricing in a browser with both billing toggles open, read the Terms end to end, and read the consumer profile the vendor itself invites customers to. Every figure below is dated and traces to a saved copy of the document it came from.

The most useful number about this tool is one it never advertises, and the vendor's own marketing sits next to it awkwardly.

In 2023 a team of eight academics across seven European and Latin American universities put fourteen AI-text detectors through the same 54 documents. Winston AI, listed there under its domain name GoWinston, produced no false accusations at all: across the eighteen documents that were written by humans or translated by machine, it flagged none of them as AI. Six of the fourteen tools generated false positives. This one was in the half that did not.

The same table shows what that caution costs. Of the nine documents that were generated by ChatGPT and then rewritten with a paraphrasing tool, Winston AI correctly classified one.

Those two findings are not in tension. They are one property seen from two sides: a detector reluctant to call text AI will rarely accuse an innocent writer and will rarely catch a disguised one. Which of those matters more depends entirely on what you are buying it for, and this page reports both rather than choosing.

How Winston AI's 43.8 Score Was Built, Dimension by Dimension

DimensionReadingStrengthWhere it comes from
Capability11.1thinWeber-Wulff et al. 2023, machine-paraphrased document class, n=9
Usability & control76.5strongTrustpilot 3.9 on n=29 pooled with G2 4.4 on n=13, both solicited, read 2 September 2026
Value65.2strong$18 a month, entry paid tier, month-to-month
Commercial terms66.5strongTwo of four clause positions present in the Terms

Composite 43.8 of 100, which is 2.2 on a five-point scale, computed 2 September 2026 under score model v11. The composite is the geometric mean of the four readings, so a single low cell pulls it down harder than an average would, which is the intended behaviour and is why the capability row deserves the explanation below rather than a footnote.

Confidence is low, and the card names the single reason: capability rests on a sample of nine. The other three dimensions all grade strong. What the card also has, and what no other card in this category has, is a reading on every one of the four dimensions from evidence measuring that dimension. Nothing here is a substitute standing in for a cell we could not fill.

Is Winston AI Accurate? The Paraphrase Class That Sets the Capability Row

The scored figure comes from "Testing of Detection Tools for AI-Generated Text", published as arXiv 2306.15666 in June 2023 and peer-reviewed in the International Journal for Educational Integrity that December. Eight authors, at HTW Berlin, Riga Technical University, Uppsala, Masaryk, Universidad de Monterrey, Queen Mary University of London and Leeds. No vendor relationship was found. The paper names Winston AI among four companies that "claim to be the best on the market", which tells you the posture it was written in.

Nine researchers each prepared a set of documents. One class, labelled 06-Para, was generated by ChatGPT and then, in the paper's words, "rewritten automatically with the AI-based tool Quillbot (Quillbot, 2023), using the default values of the tool for modes (Standard) and synonym level". Nine documents, because "there were 9 documents in each class". Winston AI classified one of them correctly.

One of nine is 11.1%, and that percentage is our arithmetic rather than the paper's: the table prints counts per column and a percentage only for the row total. The count it prints for this tool in that column is 1.

This particular arm was chosen because it is the arm every other card in this category was scored on. Those three were measured by a different 2023 study, which put ChatGPT output through QuillBot and counted what survived; their readings on that axis are in the comparison table further down. Winston AI is not in that study. Rather than reach for a different question, this card reaches for the same question in a second paper: same paraphrasing tool, same generator era, same thing being asked.

That pairing is checkable rather than merely convenient. GPTZero is the one tool both research teams put through this test, and it reads 20.0% in one and 33.3% in the other. Two independent groups, the same neighbourhood. The limit is worth stating anyway, because it is real: Copyleaks, Originality.ai and GPTZero were all scored from a single table and are comparable to each other row for row. Winston AI is not in that table. Its figure comes from a different team measuring the same property, and anyone reading the capability column across this category should know which cell that applies to.

Winston AI's False Accusation Rate: Zero Across 18 Documents

The same paper carries a table the marketing pages do not quote. It counts, for each tool, how many classifications would have led to a false accusation, looking only at the eighteen documents that no AI wrote: nine human-written and nine machine-translated from a human original.

Winston AI: 0. Zero false positives, zero partial false positives.

For context on where that sits, six of the fourteen tools produced false positives, and the paper notes the risk "increasing dramatically for machine-translated texts". GPTZero, which has a card on this site, recorded 9 of 18, a ratio of 50.0%. Crossplag recorded 16.7%. Seven other tools, Winston AI among them, recorded none.

Two things keep this from being more than it is. The sample is eighteen documents in 2023, so zero out of eighteen is not the same claim as never. And the writing was academic prose prepared by researchers, not the full range of what a real classroom produces. What it does establish is that on this corpus, in this year, the tool did not accuse anybody.

Winston AI Detector Ranked 5th of 14 on Overall Accuracy

Across all six document classes, the paper's binary reading puts Winston AI at 67%, fifth of the fourteen tools tested. Its two other scoring approaches, which award partial credit for hedged verdicts, read 70% and 75%, both fourth. The paper names the tool directly in its findings: "Crossplag and GoWinston were the only other tools to achieve at least 70% accuracy", the other three above that line being Turnitin, Compilatio and the GPT-2 Output Detector.

The GPT-2 Output Detector placing third is worth a moment, and the authors flag it themselves: it was never trained on the model that wrote these documents.

Why Winston AI's Capability Row Reads 11.1 and Not 67

Both figures are real, both come from the same table, and the card publishes the lower one. The reason is comparability, not pessimism.

Publishing 67 would put a general-accuracy figure beside three siblings scored 50, 40 and 20 on paraphrase resistance specifically. The column would then rank four tools on a number that means one thing for one of them and something else for the other three, and this tool would appear to lead the category on the strength of a measurement none of its siblings was given. That is a ranking produced by the choice of instrument rather than by anything about the products.

So the scored cell holds the arm the category shares, and the overall figure ships here in the prose with its own rank and its own year. The paper's own conclusion about the whole field is the frame both belong in: detectors "are neither accurate nor reliable and have a main bias towards classifying the output as human-written rather than detecting AI-generated text", with roughly 20% of AI-generated text likely misattributed to humans, rising to about 50% once any obfuscation is applied. That is a statement about fourteen tools, this one included, and not about this one in particular.

One age caveat governs everything above. The data was collected between February and May 2023, against GPT-3.5 output. Every model Winston AI now advertises detecting was released afterwards. The measurement is real, dated, and cannot speak to a model it never saw.

Winston AI Reviews: 3.9 on Trustpilot, 4.4 on G2, and a Distribution With No Middle

The Trustpilot profile for gowinston.ai, read in a rendered browser on 2 September 2026, shows 3.9 out of 5 on 29 reviews. The distribution is the part worth looking at:

RatingShare
5 star73%
4 star0%
3 star0%
2 star3%
1 star24%

There is no middle. Not a thin middle, an empty one. A 3.9 average computed from that shape is not a lukewarm verdict from twenty-nine people, it is two incompatible verdicts averaged into a number that nobody actually gave.

This is the first record in this category that Trustpilot labels as solicited. The platform's company-level block is headed "Asks customers to review", and the line under it explains that the vendor invites its customers to leave reviews whether they are positive or negative. Under our scoring rules a solicited sample cannot be pooled with an uninvited one, so it is not pooled with any sibling here, and it should not be read as directly comparable to the self-selected records on the other cards.

One tension is recorded rather than resolved. The platform states at company level that the vendor invites reviews, and every one of the twenty visible review markers on the same page reads "Unprompted review". A company can run an invitation programme while the reviews that surface arrive on their own. This desk does not pick one signal over the other.

Winston AI on G2, Capterra, Gartner and Product Hunt: What the Other Platforms Hold

Four more platforms were checked on 2 September 2026, and only one of them holds enough to count.

PlatformRatingReviewsUsed?
Trustpilot3.929yes
G24.413yes
Capterra4.01no, below the floor
Product Hunt5.01no, below the floor
Gartner Peer Insightsnone0no, the profile carries no reviews

The G2 record clears the ten-review floor, and its declared collection regime matches Trustpilot's, so the two are pooled by sample size rather than one being preferred over the other. That gives 42 reviews at a raw 81.1, shrunk to the published 76.5 because forty-two is still a thin sample and the scoring model pulls a thin one toward a neutral anchor.

Pooling them is only defensible because they were collected the same way, and the G2 markers make that checkable: across the reviews rendered on that profile, eight carry "Incentivized" and eight carry "Source: G2 invite", with the platform's own tooltip explaining that the reviewer "was offered a nominal gift card as thank you for completing this review". A 4.4 built from invited and incentivised reviewers is not the same object as a 4.4 from people who arrived on their own.

The Gartner result is worth stating plainly because that page ranks fifth on Google for this tool's own brand term. It is a listing rather than a review record: a product description, a "Updated 29th May 2026" stamp, and a navigation item reading "Reviews and Ratings (0)".

None of these five pages can be read by an ordinary fetch. G2, Capterra, Trustpilot, Product Hunt and Gartner all refuse this desk's automated requests, and every figure above was read in a real logged-in browser.

What Winston AI Users Report About False Positives

The complaint that recurs is the mirror image of the paper's zero.

A February 2026 Trustpilot reviewer describes testing the tool's accuracy and reports that well-written, coherent, concise text is flagged as AI-generated. In an r/Professors thread dated 10 August 2026, asking for detector recommendations, the highest-scoring comment at +19 reports feeding in a thesis written more than fifteen years ago and having most of it come back flagged.

Set that against the controlled result honestly. Zero false accusations across eighteen documents in 2023, and users reporting false accusations in 2026, are not a contradiction to be resolved by discarding one. They are separated by three years and by everything about how the text was produced. Both are dated, both are sourced, and neither was verified by this desk.

Winston AI for Teachers: What an r/Professors Thread Says About the Category

The same thread carries something broader than any product, and it is reported here as category-wide because that is what it is.

The +19 comment states the writer is redesigning course structures to stop relying on take-home papers, returning to in-class blue books. The second comment, at +10, argues that the instructor is the detector, and describes documenting evidence by hand at a cost of an hour or two per paper.

Neither is a verdict on this tool. Both describe practitioners in the buying group for this category concluding that the category is not what they need. The academic paper reaches the same destination by a different road, concluding that the tools it tested "should not be used in academic settings" as evidence of misconduct.

A note on sourcing in this category: most of what surfaces for this brand on Reddit is promotional, and the thread used here is usable partly because its own readers policed it.

Winston AI Pricing: the Page Opens on $10, and the Monthly Price Is $18

The pricing page opens on the annual toggle. What it shows first is $10 a month.

Paying month to month costs $18. The full published ladder, read with both toggles open on 2 September 2026:

PlanAnnual viewMonth to month
Free$0$0
Essential$10/mo, $120/yr$18/mo
Tier 2$16/mo, $192/yr$29/mo
Tier 3$26/mo, $312/yr$49/mo

How that table was established is worth one sentence, because the monthly column is not text the page displays. The toggle would not switch: clicking it hung the renderer twice, so the monthly view never rendered. Both sets of numbers sit in the page's markup at all times, which means reading the text returns all of them with nothing to say which is which. What separated them was an inspection of every element holding a bare number, splitting them by whether they actually render, and then the vendor's own arithmetic: 10 ÷ 18 is 44.4%, 16 ÷ 29 is 44.8%, 26 ÷ 49 is 46.9%, against a toggle that advertises "Save 45%".

Every price on this site is read on a monthly basis, so $18 is the scored figure. A buyer who wants to pay monthly pays 80% more than the number the page displays on arrival.

Is Winston AI Free? A 14-Day Trial, Not a Standing Free Tier

The plan is called Free and prices at $0. What it grants is "2,000 credits / 14 day trial".

That is a trial with an expiry rather than a free tier that persists, so this card does not award the free-tier credit that a standing free plan earns. The Essential plan publishes 100,000 credits a month. What one credit buys was not stated on the pricing page, and is not guessed at here.

Winston AI Refund Policy: Non-Refundable, With Quebec Consumer Law Named

The Terms are direct: "All purchases are non-refundable, subject to any mandatory rights you may have under applicable consumer protection laws in your jurisdiction, including the Consumer Protection Act (Quebec) for Quebec residents."

That clause refuses refunds and then preserves statutory rights by name, which is a materially better position for a buyer than a flat refusal, and it scores as such. Cancellation is self-serve through the account.

Governing law is the Province of Quebec, with exclusive jurisdiction in Montreal. A buyer outside Canada is agreeing to litigate there, which is worth knowing before it matters rather than after.

Who Owns What You Submit to Winston AI

Section 24 carries the ordinary one-way indemnity: "You agree to defend, indemnify, and hold us harmless, including our subsidiaries, affiliates". A second indemnity appears later covering use of AI output. Nothing in the document runs an indemnity the other way. That is the market default, matching 19 of the 25 agreements read across this site.

The ownership clause is the part that stands out, and it is the cleanest of its kind read here in either of the two document-heavy categories on this site: "We do not assert any ownership over your submitted content. You retain full ownership of all content you submit."

What makes that notable is what does not follow it. Elsewhere the same sentence is routinely followed by a perpetual, irrevocable licence back covering commercial use and sublicensing, which returns most of what the first sentence gave away. Here there is no licence-back over submitted content at all.

Nothing in the Terms distinguishes what a free-tier user may do commercially from a paid user. That is recorded as unstated rather than assumed either way, and it is one of the two clause positions that leaves the commercial-terms cell resting on two of four rather than all four.

Winston AI Alternatives: The Other Cards in This Category

Four tools in this category carry a card. Ordered by composite, and the ordering is computed from the scores at build time rather than written by hand:

ToolCompositeCapability cellConsumer record
Copyleaks69.650.0not solicited
Originality.ai52.640.0not solicited
GPTZero45.320.0self-selected
Winston AI43.811.1solicited

Read that column with the caveat this page has already made: the first three capability figures come from one table, and Winston AI's comes from another team's table measuring the same property on the same paraphrasing tool. They are comparable in what they ask and not identical in how they were produced.

The false-accusation record cuts the other way and is worth putting beside the composite, because it comes from the one paper that measured both tools on the same documents: Winston AI recorded 0%, GPTZero 50%.

No single published study has measured every tool on this list against every other one. That is a fact about what exists in this field rather than an omission on this page, and it is part of why no head-to-head page exists here.

Winston AI Review Verdict: What Its Caution Buys, and What It Costs

43.8 is a low band, and the number is driven almost entirely by one cell resting on nine documents from 2023. The other three dimensions all grade strong.

The honest summary is narrower than the composite. If what you need is a detector that does not accuse people who wrote their own work, the strongest published evidence about this tool says it did not, on the corpus where that was measured, in the company of six other tools that also did not. If what you need is a detector that catches text someone has deliberately run through a paraphraser, the same paper caught one document in nine, and the field average on that class was 26%.

The paper those numbers come from concludes that the tools it tested are not suitable as evidence of academic misconduct. That conclusion covers all fourteen, and pointing it at this one alone would misuse it.

Three things about the purchase are unambiguous and do not depend on any accuracy figure. The month-to-month price is $18 rather than the $10 the page opens on. Refunds are refused except where consumer law requires them, and the governing law is Quebec. And what you submit stays yours, with no licence back to the vendor, which is not the market norm.

Card computed 2 September 2026, score model v11, on four measured dimensions of four. Confidence low: capability rests on n=9; the other three dimensions grade strong. Our methodology states what we measure and what we do not. We run no detection tests of our own and make no claim to have tested this or any other detector.

Published 2 September 2026

Scores and evidence on this page are re-checked monthly. Read about the person behind the scores, or find me on LinkedIn.

Licence terms, ownership and litigation status are reported here with the date they were read and a link to the source. They change, and a summary is not a clearance: nothing on this site is legal advice, and whether a particular use is safe for your work is a question for a lawyer in your own country. Verify a vendor’s current terms before you commit a deliverable to them.