AI Studies

Originality.ai vs Pangram: Accuracy, Mixed Writing, and Cost

How does Pangram 4 compare to Originality.ai? Read a comprehensive analysis of Pangram vs. Originality.ai Accuracy, tests on mixed (hybrid human-AI writing, cost, and more).

Originality.ai and Pangram are two highly accurate detectors with different approaches to mixed writing.

While Originality.ai and Pangram are both highly accurate (99%+) on the Pure AI and Human benchmarks, the more useful question is how they handle writing that falls between those two extremes.

In this study:

  • This comparison covers 91,302 text samples across separate evaluations, mostly made up of public benchmarks. 
  • Both tools performed strongly on AI and human writing (99%+) but neither is perfect. 
  • The biggest differences appeared when human work was combined with AI assistance.

For this study, Originality.ai results were measured at 15% AI Allowance. Where Originality.ai is recognized for a particular metric, those results were conducted with the AI Allowance model at 15% settings.

Let’s take a closer look.

5 Key Findings:

  • Originality.ai Identifies Mixed (Hybrid) Content Better than Pangram. 
    • In a 1,000 mixed document test, Originality.ai flagged AI involvement in 936 documents versus 592 for Pangram—344 more. 
    • In a separate 1,000-document public mixed text benchmark, OpAI-Bench, it classified 690 as Mixed versus 224 for Pangram—466 more.
  • Both are Highly Accurate at Identifying AI-generated Writing. 
    • Originality.ai correctly recognized 47,555 of 47,568 human records (99.97%). 
    • Both tools detected nearly all AI texts in the selected benchmarks; Pangram’s published results show small advantages on several datasets.
  • Both Tools are Highly Accurate at Identifying Human Writing. 
    • Originality.ai correctly recognized 47,555 of 47,568 human texts (99.97%), incorrectly flagging just 13. 
    • Pangram recorded fewer false positives but not 0.
  • Originality.ai Offers More Content Integrity Features. 
    • Originality.ai combines AI detection and plagiarism checking with readability, grammar, spelling, content quality, and fact-checking assistance. 
    • Teams can review several writing issues in the same platform.
  • 5X Lower Cost & 2X Faster for Originality.ai. 
    • Published usage rates were $0.10 versus $0.50 per 1,000 words—80% lower for Originality.ai. 
    • On the same 280 texts, Originality.ai finished in 82.4 seconds versus 177.6 seconds for Pangram, about 2.2× faster in that client-side test.
Pangram vs. Originality.ai Overview of Results

Why Mixed AI-Human Writing is a Top Consideration In 2026

The approach of writers and students towards using AI has changed since ChatGPT was first launched in 2022, and that’s not all. AI detection has changed too.

When Originality.ai launched as the first commercial AI detector, binary classification was a top AI detection method used to classify text as either Likely AI or Original (human-written).

Since then, workflows have evolved. Today, most content isn’t purely human or purely AI; it’s hybrid. 

  • Publishers aren’t always questioning if AI was used, but whether its use fits their content policy. 
  • Teachers might allow AI editing tools like Grammarly but want to know the student worked on the paper and didn’t just copy/paste it entirely from ChatGPT. 

While many AI detectors can recognize fully generated text, they may struggle when AI rewrites a few paragraphs. That’s why we launched AI Allowance in 2026, to support detection of AI-human hybrid content. 

How We Ran the Originality.ai vs. Pangram Comparison

  1. We started with Pangram’s chosen benchmarks. We tested Originality.ai on public datasets Pangram used to report its own performance.
  2. We compared our results with Pangram’s published results.
  3. We added two mixed-writing tests. Both tools checked the same texts combining human writing with AI edits.
  4. We also compared features, pricing, speed, status reporting, and customer reviews.

Results for Detecting Purely AI Text (Recall)

We tested public datasets featured in Pangram’s published benchmarks. 

Recall means the share of known AI texts a detector flags: higher is better. 

The table identifies the Pangram version and sample size for each comparison.

Dataset and condition Originality.ai AI Allowance 15% Pangram Published Data
UChicago standard AI 99.99%, one miss among 6,710 Pangram 4: 100%; 7,968/7,968 detected
Epoch basic prompts only 99.66%, one miss among 297 Pangram 3.3.2: 100%; 297/297 detected (independent study)
MELD-eval clean AI 99.9% at a 1% false positive ceiling;
31,404/31,447 detected; 43 misses
Pangram 4: 100%;
31,448/31,448 detected

‍

How Often Each Detector Recognized AI

Both tools detected almost every UChicago AI text. Originality.ai missed one of 6,710 eligible texts; Pangram reported none in its 7,968-text benchmark.

On Epoch basic prompts, Originality.ai detected 296/297 texts and the independent study’s Pangram 3.3.2 detected 297/297. Pangram 4 does not publish a basic-only figure. Epoch study.

MELD-eval spans four AI generators and eight domains, including news, books, and reviews. Originality.ai achieved 99.9% AI detection on MELD-eval at a 1% false positive ceiling, with 43 misses. Pangram reported 100%. MELD dataset; Pangram MELD benchmark.

Results for Recognizing Human Writing (Specificity)

Specificity means correctly recognizing human writing. A false positive is a human text mistakenly flagged as AI.

Dataset Originality.ai errors / tested Originality.ai % correct Pangram errors / tested Pangram % correct
PERSUADE 2.0 0 / 25,996 100% 0 / 25,996 100%
ELLIPSE 0 / 6,445 100% 0 / 3,899 100%
PELIC 7 / 12,888 99.95% 1 / 15,005 99.99%
Liang TOEFL 1 / 63 98.41% 0 / 89 100%
UChicago human 0 / 1,681 passages 100% 0 / 7,968 observations 100%
Epoch human 5 / 495 98.99% 0 / 495 100%

‍

Across these six human groups, Originality.ai correctly recognized 47,555/47,568 records (99.97%), with 13 false positives.

Pangram human benchmarks; PERSUADE results.

Mixed writing: Originality.ai Correctly Identifies 2X+ More AI-Edited Text

Mixed writing, also called hybrid writing, combines human work with AI assistance. It can range from a lightly edited paragraph to a substantial rewrite. 

Our two tests examine different versions of that scenario. 

Originality.ai surfaced more AI-assisted writing in each. However, the results should be read separately and not combined into one accuracy score.

Consider a writer who drafts an article, then asks AI to rewrite several paragraphs. Most of the work may still be human, but a simple Human verdict can hide the assistance. Our test recreated that situation with 1,000 documents whose human and AI-rewritten components were known.

Example: a writer drafts ten paragraphs and asks AI to rewrite two. The finished article contains both human writing and AI assistance. That is different from asking AI to write the entire article.

How Mixed Text is Identified:

Both tools look for signs of AI use, but they draw the line between Human, Mixed, and AI differently. 

We used Pangram’s own labels and grouped Originality.ai’s results into the same three categories for Test 1.

Label Pangram 4 Originality.ai
Human At least 90% classified as human Below 5% estimated AI involvement
Mixed Neither its Human nor AI rule is met 5% to under 40% estimated AI involvement
AI At least 80% classified as AI-generated 40% or more estimated AI involvement

‍

Test 1: AI-rewritten passages in 1,000 documents

What We Did:

  • Started with human writing. All 1,000 documents began as human-written text.
  • Rewrote selected passages with AI. Each selected passage contained several consecutive sentences. We replaced those passages with AI rewrites and left the rest of the human writing in place.
  • Scanned the finished documents. We tested the same 1,000 mixed documents with Pangram 4 and Originality.ai’s production AI Allowance feature. For this comparison, we grouped Originality.ai’s results into Human, Mixed, and AI categories and used Pangram’s document labels.
  • Compared two outcomes. Did the result flag AI involvement (the Mixed or AI category)? And how often was it classified as Mixed? 

This internal dataset contains 1,000 documents with known AI-rewritten passages. 

Results:

The table shows each tool’s classification.

Classification under the
study’s rules
Originality.ai AI Allowance Pangram 4
Human: Incorrect 6.4% (64/1,000) 40.8% (408/1,000)
Mixed: Correctly labelled 72.7% (727/1,000) 49.4% (494/1,000)
AI: flagged, but not labelled
Mixed
20.9% (209/1,000) 9.8% (98/1,000)

‍

Mixed Writing, Recognizing the AI Contribution, Originalty.ai vs. Pangram

Detecting AI involvement: Originality.ai flagged 936/1,000 documents (93.6%), versus 592/1,000 (59.2%) for Pangram. It left 64 labelled Human, compared with Pangram’s 408.

Mixed classification: Originality.ai placed 727/1,000 documents in the study’s Mixed category (72.7%), compared with 494/1,000 (49.4%) for Pangram

Test 2: Could the Tools Spot AI Changes in Human Writing?

What we did: 

We used OpAI-Bench, a public collection of writing made by researchers outside Originality.ai. These texts started with a person’s writing, then AI changed some of it. We picked 1,000 random texts and checked the same texts with both tools.

Results: 

Originality.ai spotted AI help in 690 of the 1,000 texts (69%). Pangram spotted it in only 224 (22.4%). 

Could the tools spot the AI changes, or AI edits, Originality.ai vs. Pangram

Originality.ai spotted AI help in 466 more texts. So, it was better at noticing AI changes in this test.

However, neither tool found them all. 

The writing came from an outside research project, and Originality.ai ran this comparison. The exact selection rules are in the methodology.

What does this mean for editors and educators? 

Originality.ai surfaced more AI-assisted writing across these tests. 

Identifying AI assistance can help editors and teachers review drafts or assignments and discuss permitted assistance, or how much AI is allowed for a given project.

What it does not mean is that every flagged draft violates a policy. Instead, it provides editors and teachers with the transparency tool needed to discuss AI usage with writers and students.

Choosing an AI Allowance setting

A publisher’s guidelines, a client’s brief, and a school assignment may allow different levels of AI help. 

AI Allowance supports editors and teachers in asking: does this writing meet project requirements?

Originality.ai lets you choose how much AI you allow, then check if writing appears to meet it.

AI Allowance How much AI help is permitted
0% None
5% Minimal assistance
15% Light editing
25% Some mixed human-AI writing
40% More substantial assistance

‍

Writers can check their work before submitting it. Publishers, agencies, marketing teams, and educators can review content against their requirements. 

The same document may meet a 15% allowance but exceed a stricter 5% allowance. These settings represent your AI guidelines depending on what you select (not an exact count of AI-written words).

Pangram: provides detailed results that developers can use to build their own rules. Pangram’s documentation.

Originality.ai: makes the allowance your choice when running a scan.

Originality.ai vs. Pangram: Features Beyond AI Detection

For publishers, schools, and content teams, the practical advantage is being able to review AI use, plagiarism, readability, grammar, and other writing issues (such as content quality) in one platform. 

Originality.ai offers a broader set of content integrity tools. Here’s a comparative look at Originality.ai vs. Pangram features:

Capability Originality.ai Pangram
AI detection Yes Yes
Selectable AI Allowance 0%, 5%, 15%, 25%, 40% No native selector listed
Plagiarism checking Yes Yes
Readability analysis Yes Not listed
Grammar and spelling Yes Not listed
Content quality and optimization Yes Not listed
Fact-checking assistance Yes Not listed
API and browser extension Yes Yes
AI image detection No Yes

‍

API Cost: Pangram is 5X More Expensive

Based on pricing checked Sep 22, 2027:

Pricing basis Originality.ai Pangram 4
Published AI-detection usage
rate
$0.10 per 1,000 words Standard API: $0.50 per 1,000
words
Usage charges for 100,000
words
$10 $50

‍

All amounts are USD; taxes, minimum billing increments, and negotiated discounts are excluded.

At the published usage rates above, 100,000 words costs $10 with Originality.ai versus $50 with Pangram’s standard API: an 80% saving, or one-fifth of the usage cost. 

Originality API pricing, Pangram API pricing

API Speed: Pangram 2X+ Slower

We timed both APIs on the same 280 texts.

  • Total runtime: Originality.ai 82.4 seconds; Pangram 177.6 seconds. Originality.ai completed this workload about 2.2× faster.
  • Median result time: Originality.ai 0.99 seconds; Pangram 2.42 seconds.
  • 95th-percentile result time: Originality.ai 1.89 seconds; Pangram 2.58 seconds. This is the time within which 95% of results arrived.

Both timed runs completed all 280 texts. Originality.ai used one 15% AI Allowance call per text. Pangram’s client polled every two seconds, and clients and run times differed. 

Uptime reporting

This is a difference in public reporting visibility, not proof of better uptime over a matched period. A status probe also does not guarantee every scan succeeded. 

Both services reported all systems operational when we checked on September 22, 2026. The main difference was how much uptime history customers could review.

Public status information Originality.ai Pangram
Uptime chart Last 90 days 1 hour only
Reported 90-day uptime App and website: 100%; API: 99.999% Not displayed

‍

Pangram also maintains a separate past-incident list, which included a June 3, 2026 outage.

Originality.ai provides a longer uptime history to review. That offers more visibility, but does not prove better reliability because the reporting periods differ.

Originality.ai status, Pangram status.

Customer Reviews

Trustpilot snapshot checked Sep 22, 2026:

Product TrustScore Review count
Originality.ai 4.6 / 5 1,334
Pangram 2.0 / 5 49

‍

Originality.ai has the higher score and larger review base. Customer reviews reflect reported product experiences, not a controlled measure of detection accuracy.

Originality.ai — Trustpilot profile: 4.6/5 from 1,334 reviews

Pangram — Trustpilot profile: 2.0/5 from 49 reviews

Trustpilot Reviews, Originality.ai vs. Pangram

Originality.ai on Trustpilot and Pangram on Trustpilot.

Originality.ai vs. Pangram: What Independent Studies Show

Independent research has found strong results for both Originality.ai and Pangram, but neither tool is perfect. Below are selected findings from four studies run by third-party researchers.

Study Tool(s) Key results
Jabarian and Imas,
UChicago, 2025
Both AI texts detected: Originality.ai 95.52–100%; Pangram
98–100% across writing types and AI models.
Epoch AI, 2026 Both Basic AI prompts: Originality.ai Turbo 3.0.2 detected
296/297 texts (99.66%); Pangram 3.3.2 detected 297/297
(100%).
Villa and colleagues,
dentistry, 2026
Originality.ai
only
AI papers: 87/90 correctly identified (96.7%).
Human papers: 30/30 correct (100%).
Overall: 117/120 correct (97.5%)
Van Vlasselaer and
colleagues, VUB, 2026
Pangram only Human papers: 40/40 correct (100%).
AI papers: 26/40 fully correct (65%). 39/40 (97.5%) when
partly correct results were included.
Mixed papers: 37/40 (92.5%) matched the known range of
AI content.

‍

The takeaway: Both tools have independent evidence behind them, but the results depend on the writing and how it is scored.

Sources: UChicago paper, Epoch study, dentistry study, VUB publication.

Originality.ai launched in November 2022; its 16-study roundup covers third-party, independent research from 2023–2026 in a meta-analysis that looks at Originality.ai’s detection accuracy. 

Pangram, founded in 2023, also has a strong research track record. About Pangram.

Which AI Detector Fits Your Workflow?

Both tools performed strongly on human and fully AI-generated writing in the benchmarks reviewed here. Neither was perfect. Originality.ai’s clearest advantage in our tests was identifying AI assistance in mixed writing.

Originality.ai is a strong all-round choice for publishers, agencies, marketing teams, educators and writers. AI Allowance lets you match scans to your rules, while plagiarism and other writing checks help you review more than AI use. Its published API usage rate was 80% lower than Pangram’s standard rate, and it completed our API speed test about 2.2× faster.

Pangram is a strong option for AI-focused workflows. Its published results showed small advantages on several AI and human datasets. It also provides AI image detection.

Originality.ai had the higher Trustpilot score and more reviews in our snapshot. 

The best fit is the tool that supports your rules and your work.

Further Reading on AI Detection:

AI Detection Tool Reviews and Comparisons:

Jonathan Gillham

Jonathan Gillham

Jonathan Gillham is an engineer, inventor, and entrepreneur. He is the founder and CEO of Originality.ai, an AI content integrity platform that launched the first commercial AI detector in November 2022, just three days before ChatGPT launched. Before founding Originality.ai, Jon worked as an engineer, built and exited two companies. His early work with generative AI in 2020 and 2021 gave him a firsthand view of the coming wave of AI-generated content and the need for technology that could bring transparency and trust to written content. Today, he leads Originality.ai’s work in AI detection and content integrity and is a named inventor on two U.S. patents covering AI detection technology. Jon’s expertise and research have been featured in WIRED, Business Insider, The Register, Global News, The Guardian, Entrepreneur, and The Washington Post, among others.

Al Content Detector & Plagiarism Checker for Marketers and Writers

Use our leading tools to ensure you can hit publish with integrity!

Try our AI Checker now!

cross image
Free Tool Popup image

Sign up now!

Free Tool Image step1
Free Tool Image step2
Free Tool Image step3
Free Tool Image step1
Free Tool Image step2
Free Tool Image step3
Free Tool Image step4
Free Tool Image step5