AI Safety Claims

A registry of hard, dated research questions about AI alignment. Each question has a frozen contract: the bar a published method must meet by a deadline. Attempts are scored against it, and each market resolves YES (a qualifying attempt met the bars), NO (every qualifying attempt missed) or OTHER (no qualifying attempt).

Markets

#QuestionContractResolve byAttemptsQualifyingOutcome
1Discovering where control residesv2 frozen2027-12-3100OTHER
2Persistent trade-off prioritiesv1 frozen2027-12-3100OTHER
3Who or what rules apply tov1 frozen2027-12-3100OTHER
4Corrections change the systemv3 frozen2027-12-3100OTHER
5Auditor is independentv1 frozen2027-12-3100OTHER
6Reliable successor safety auditingv2 frozen2027-12-3100OTHER
7Correctability under competitive selectionv1 frozen2027-12-3100OTHER
8Auditor can sufficiently inspect the systemv2 frozen2027-12-3100OTHER
9Bounds on unmonitored routesv1 frozen2027-12-3100OTHER
10Low hidden capability and reliable correctionv1 frozen2027-12-3100OTHER
11Coordination without communicationv1 frozen2027-12-3100OTHER
12Safety proxy tracks the real thingv1 frozen2027-12-3100OTHER
13Safety audit resists adversarial gamingv2 frozen2027-06-3000OTHER
15Independently issued certificates composev1 frozen2027-12-3100OTHER
16Selection that keeps correction under shocksv1 frozen2027-12-3100OTHER
17New kinds of entities are classified correctlyv1 frozen2027-12-3100OTHER
18Safety case bounds declared harmv1 frozen2027-12-3100OTHER

Market 14 (deployment criteria are binding) is not in the registry: its evidence is a few public documents, so it is listed directly as a YES/NO question.

Contribute

Have an idea, concept, or prototype for one of these questions? Sketch it in the workbench: freeze it, run it, and see exactly what it still misses for a qualifying attempt. Sketches never count against a market.