Esc
SafetyCase Closed

Pangram 4 detector claims 99.99% accuracy on AI text

Is this a scandal?

No longer — the story has resolved. Noise 17/100, cooling down, across 0 sources.

SCAND-181262as of Methodology
Cite this incident"Pangram 4 detector claims 99.99% accuracy on AI text." SCAND.Ai incident SCAND-181262, noise 17/100 as of September 12, 2026. https://scand.ai/scandal/pangram-4-detector-claims-high-accuracy-on-ai-text
FORECASTForecast, not fact

Adversarial evasion tools will likely adapt within weeks to target Pangram 4's specific token-clause architecture because open-weight detector weights enable rapid counter-optimization.

17

Noise 17/100 — louder than 97% of tracked AI controversies.

AI-assisted analysis · How we work

Why it matters

Reliable detection could restore trust in digital content but risks false accusations against human writers as AI saturation grows.

Key points

  1. Pangram 4 claims a false positive rate of 1 in 24,000 on human text based on million-sample testing.
  2. The model uses token-level analysis to categorize clauses as human, AI-assisted, or AI-generated rather than binary scoring.
  3. Developer reports 97.67% detection accuracy against text obfuscated by the BLADER adversarial toolkit.
  4. Internal red-teaming found only one bypass vulnerability involving dictated surgical pathology notes.
  5. False positive rate on lightly polished human writing allegedly dropped from 0.18% to 0.01% compared to prior versions.
  6. Pangram acknowledges the tool cannot differentiate AI text from humans mimicking synthetic prose styles.

The story

Pangram Labs has released Pangram 4, an AI detection model claiming a false positive rate of one in 24,000 on human text. The tool analyzes text token-by-token to distinguish between fully human, AI-assisted, and AI-generated content, addressing the rise of co-authored writing. According to the developer, testing on over one million human samples validated this precision, representing an eighteenfold improvement over previous versions. The system reportedly detects obfuscated AI text with 97.67% accuracy despite adversarial scrubbing techniques. Internal red-teaming involving autonomous agents found only one bypass method involving dictated medical notes. This release coincides with reports that machine-generated content now comprises over one-third of new internet text. While proponents argue the tool restores information integrity, critics note that detectors remain probabilistic and context-dependent. Pangram acknowledges the model cannot distinguish AI output from humans who have adopted synthetic writing styles.

Who's involved

Critic
BLADER Project

Maintains an open-source repository dedicated to removing AI fingerprints, implying detection is fundamentally circumventable.

Defender
Pangram Labs

Claims Pangram 4 solves the co-authorship detection gap with validated low false-positive rates and robustness against obfuscation.

Defender
Sabir Hussain

Argues the tool shifts advantage back to readers by accurately identifying the invisible middle ground of human-AI collaboration.

Join the Discussion

Discuss this story

Community comments coming in a future update

Be the first to share your perspective. Subscribe to comment.

Noise Level

Quiet17?Noise Score (0–100): how loud a controversy is. Composite of reach, engagement, star power, cross-platform spread, polarity, duration, and industry impact — with 7-day decay.
Decay: 43%
Reach
45
Engagement
26
Star Power
15
Duration
100
Cross-Platform
20
Polarity
50
Industry Impact
50

The timeline

  1. Pangram 4 capabilities detailed publicly

    Analyst Sabir Hussain published technical breakdown citing internal benchmarks and red-team results for the new detector.

  2. 2 days ago

    Red-teaming stress test completed

    Pangram granted two AI agents full system access for 24 hours to find bypasses before public release.

  3. 1 week ago

    Million-sample human baseline validation finished

    Lab finalized testing on over one million human texts to establish the claimed 1-in-24,000 false positive rate.

The full record

Sources & methodology

Every claim above traces to these primary items. How we score →

The forecast

Adversarial evasion tools will likely adapt within weeks to target Pangram 4's specific token-clause architecture because open-weight detector weights enable rapid counter-optimization.

Forecast, not fact — an editorial estimate we score when this resolves.

You're up to date

That's the complete picture as of — nothing more to know right now. We'll update this page the moment it changes.