A certification protocol that grants an AI research claim only after matched agents fail to recover the result without its history
The Discovery Certification Protocol turns claims about AI research agent results into executable recovery and feedback tests. Gate 1 validates useful improvement on a sealed evaluation, Gate 2 gives matched agents the registered starting information and observed web content while withholding the target research history, and any valid method that reaches the numerical target supplies a recovery witness and vetoes the claim. Two controlled audits in SQLite optimization and virtual catalyst control produced zero recoveries in 96 episodes with an upper bound of 0.0468, and paired studies gave 30 truthful recoveries against zero neutral ones. A deterministic LLM-free verifier reproduces every decision from frozen evidence, which is the part worth stealing for any internal claim-review process.
↳ Follow the thread