
AI Detector Robustness: Data and Methods from Our RAID Reanalysis
Full aggregate tables from Violet's RAID reanalysis: detector accuracy under 11 attacks, per-document flip rates, per-domain false positives, methods.
VioletTag
5 postsfiled under “AI research.”

Full aggregate tables from Violet's RAID reanalysis: detector accuracy under 11 attacks, per-document flip rates, per-domain false positives, methods.

Measured on 480,000 generations: paraphrasing makes one detector 5.4pp more accurate and another 17.0pp worse. Evasion is a detector lottery, not a strategy.

One pooled 5% setting, eight writing domains of truth: measured per-domain detector false-positive rates range 0.96% to 20.4%.

Measured on 480,000 generations: one symbol substitution flips 93.7% of a detector's per-document verdicts. Aggregate accuracy hides the collapse.

Why synonym shuffling cannot guarantee clean provenance—plus RAID measurements of 11 attacks collapsing AI detectors, and Violet's integrity protocol.
Turn social and market signals into an ecosystem-growth, go-to-market, or technical engagement built around the work your team needs.
Explore engagements