OPU · Review Tool

Cross-model evaluation.A stress test, not a verdict.

Multiple AI systems can help reveal contradictions and blind spots. Their agreement does not substitute for provenance reconstruction, executed benchmarks or qualified independent review.

AI-assisted review

Independent models can widen the review surface. They do not become the jury.

The historical MZN evaluation protocol used multiple frontier models to stress-test definitions, identify blind spots and compare interpretations. The useful output is disagreement, question quality and reasoning trace—not synthetic validation.

01

Framework & scope

Give models the same phase boundary, claim definition and review question before comparing conclusions.

02

Evidence-aware questioning

Ask models to separate public evidence, restricted material, missing evidence and independent-validation requirements.

03

Compare disagreement

Model variance is useful. It can reveal prompt sensitivity, category ambiguity and assumptions that need human review.

Canonical guardrail

Cross-model agreement is not independent validation. It is an AI-assisted review signal that can improve diligence design.

Where it fits

Use it after the claim is defined, not before.

1 · BoundaryDefine the eligible claim.
2 · Model reviewPressure-test interpretations.
3 · Human diligenceInspect artifacts and evidence.
4 · Specialist validationTechnical, IP, legal and commercial conclusions.