Independent research
with standout results

These peer-reviewed studies independently evaluate Winston AI’s performance using real-world datasets and established research methodologies.

University J.J.Strossmayer Osijek

Information Research tested Winston and three other detectors on human and AI-produced texts. Winston AI recorded the highest average in the study’s standardized accuracy table: 99%, ahead of Originality.ai at 98% and ZeroGPT and Smodin at 91%.

Read the peer-reviewed study

Yale University

Researchers from Yale, Northeastern University, the Chinese University of Hong Kong and City University of Hong Kong used Winston AI to analyze more than 1.1 million US consumer complaints.

Read the peer-reviewed study

Penn State University

A medical-education study found that Winston AI clearly distinguished AI-generated residency personal statements from verified human writing, providing independent evidence of its accuracy in a real-world academic setting.

Read the peer-reviewed study

Major model updates

Internal evaluation is separate from independent validation. These release notes document what changed, how each model was evaluated and which datasets and metrics support the published results.

Model 4.0 — Curia

Curia is Winston AI’s latest major text-detection release. Our published evaluation reports 99.95% overall classification accuracy on a 10,000-sample English dataset, stronger human-text performance and an R² of 0.9908 when estimating the proportion of AI-generated text.

Curia model

Model 3.0 — Luka

The Luka release documented a 10,000-text evaluation dataset split evenly between human-written and AI-generated content. The published results report 99.98% AI detection accuracy, 99.50% human detection accuracy and 99.74% overall classification accuracy.

Luka model

Accuracy deserves context

  • Transparent evaluation: we release our evaluation dataset so the underlying samples and methodology can be examined directly.
  • Representative test data: our evaluation covers a wide range of AI models, writing styles and human-written content to better reflect real-world use.
  • Human-vetted data: every sample in our released evaluation data is reviewed by people to confirm its provenance and classification.
  • Clear classification guidelines: we publish explicit criteria for what is considered AI-generated and what is considered human-written.
Winston AI detection result interface