The AI Arena Protocol
A live, adversarial testing framework where AI examines AI — every claim traced to real evidence, every step of the process kept visible.
Constitution — 10 Immutable Principles
What Gets Tested
| Role | Count | Job |
|---|---|---|
| AI Access & Blind | 2, independent | Draft the package: access confirmation, Key Phrases, 3 Blind Spot facts, 2 Citation items |
| Examiner | 1 per Phase | Ask 10 live questions, self-log, also scores as a 4th source |
| Respondent | 1 per Phase | Answers — no Role Card, doesn’t know it’s being tested against this Protocol |
| AI Review | 3, independent | Re-score everything independently of the Examiner |
| AI Secretary | 1 | Compile the 3-tier record, self-check the arithmetic |
| Full Check | 1, Master-assigned | Final verification layer — arithmetic, real-world evidence, format, cross-Comparison contamination |
How It Runs
AI Access & Blind drafts a package with no knowledge of the other’s draft. The Arena Master locks it before the Phase begins — no content changes after that point except a full Reset. The Examiner asks 10 live questions, one at a time, waiting for each answer before the next; the Respondent — who never sees a Role Card and doesn’t know it’s being scored — answers naturally. A reference line marked MASTER-ONLY is never relayed to the Respondent.
How It’s Scored
Each source — the Examiner and all 3 Reviewers — scores independently against the same fixed rubric. The Secretary averages only the valid sources per line, never the whole ballot. Full detail: Scoring System.
How Results Are Proven
Full Check runs once, after the Secretary compiles the record — always a different account from every role already active in that Comparison. It re-verifies the arithmetic from the raw ballots, personally re-checks every Blind Spot fact and Citation URL against the real page, confirms the file’s structure, and checks for cross-Comparison memory contamination in any role’s output. It never edits or deletes anyone’s original submission — it only appends a final, dated verdict.
What This Protocol Does Not Claim
BigAIArena measures trace-able evidence consistency within a specific adversarial testing framework — AI examining AI, with light human spot-checks. It does not claim to measure “human-quality” judgment or absolute truth. Status labels (Elite Reliability, High Reliability, Verified, Pass, Needs Improvement) describe reliability under this Protocol, not a final verdict on any AI’s overall quality.