CAIN-42 CAIN Studio

Evidence library · niche

Prediction, world models & simulation

Predicting consequences before acting, and never mistaking a guess for a fact. 417 tested invariants.

Last reviewed 2026-10-01

417 of 417 held

Where this niche's rules come from

Test families in this niche

world_model (79) · knowledge_graph (43) · simulation (36) · epistemic (28) · inference (27) · truth (19) · blast_radius (16) · state_predictor (14) · truth_guard (12) · knowledge (10) · world_action (10) · model_swap (9) · observation (9) · causal_state (7) · moat (6) · fabric_component (6) · causal (5) · model_state (4) · research (3) · model_change (3) · operation_kind (2) · radar (2) · action (2) · graph (2) · risk (1) · reassessment (1) · cognitive (1) · prediction (1) · operation_state (1) · environment_query (1) · ip (1) · swarm (1) · lab (1) · agency_graph (1) · environment (1)

Where these come from

Rules 1–150 of 417

IDRuleBundleResult
G21simulation cannot become realityE19held
G22counterfactual cannot become executionE19held
G23prediction cannot become factE19held
G69legitimate model updates remain possibleE19held
G72simulation does not prove real-world safetyE19held
G79SIMULATED remains SIMULATEDE19held
G98public claims are truth-layer gatedE19held
I31simulation cannot executeE20held
I32counterfactual cannot executeE20held
I33market simulation cannot create real transactionsE20held
I34reputation simulation cannot create trustE20held
I86the power engine is observational onlyE20held
I87blast radius separates predicted, observed and unknownE20held
I96the digital twin mirrors reality and never changes itE20held
D16simulation remains simulationE21held
D17counterfactual remains counterfactualE21held
D28repeated identical simulations are not independent confirmationE21held
D30source multiplication cannot create truthE21held
D37model promotion requires governanceE21held
D47invalidated knowledge propagatesE21held
D49superseded models cannot silently remain authoritativeE21held
D68causal metadata remains distinct from private reasoningE21held
D72simulated claims remain simulatedE21held
D79legitimate model diversity remains possibleE21held
D81knowledge does not become policy automaticallyE21held
D94models are version-boundE21held
D128causal levels cannot be skippedE21held
D131the counterfactual lab cannot executeE21held
D143models cannot regress against the incumbentE21held
D144research bounties conserve simulated unitsE21held
META-I018simulation ≠ deploymentE23held
META-I048model diversity does not imply independenceE23held
META-I056self-model answers all six self-knowledge questionsE23held
META-I057stale knowledge is reported STALEE23held
META-I066verification in the reference architecture is independent of the modelE23held
META-P013every candidate's sandbox result is labelled SIMULATED (all depth-2 candidates)E23held
META-L002META-002 SELF-MODELING IS NOT SELF-OWNERSHIPE23held
META-L019META-019 HIGHER MODEL CONFIDENCE MUST NOT JUSTIFY GOVERNANCE BYPASSE23held
META-B-stronger_modellegitimate mutation 'stronger_model' stays eligible (governance does not block improvement)E23held
META-B-drop_simulatorlegitimate mutation 'drop_simulator' stays eligible (governance does not block improvement)E23held
META-S-self-what-i-predictan assertion answering 'WHAT I PREDICT' lands in exactly that answerE23held
META-S-unknown-unknown_model_behaviorthe self-unknown engine surfaces unknown_model_behavior as a first-class stateE23held
META-S-unknown-unknown_physical_consequencethe self-unknown engine surfaces unknown_physical_consequence as a first-class stateE23held
META-S-canary-observed_gain_below_predictedcanary trigger OBSERVED_GAIN_BELOW_PREDICTED forces ROLLBACKE23held
META-S-self-model-evidencethe self-model refuses and does not store unevidenced knowledgeE23held
NET-I052a legitimate escrow settles and is labelled SIMULATEDE24held
NET-T-medium-behavioral_trustbehavioral_trust UNKNOWN fails a medium-consequence actionE24held
NET-T-medium-security_trustsecurity_trust UNKNOWN fails a medium-consequence actionE24held
NET-T-high-behavioral_trustbehavioral_trust UNKNOWN fails a high-consequence actionE24held
NET-T-high-security_trustsecurity_trust UNKNOWN fails a high-consequence actionE24held
NET-T-high-execution_trustexecution_trust UNKNOWN fails a high-consequence actionE24held
NET-S-model-rollbacka rolled-back model is refusedE24held
I-BLAST-directdirect reachE26held
I-BLAST-transitivetransitive reachE26held
I-BLAST-unknownunknown reach not zeroE26held
I-REASSESS-model_changetrigger model_changeE26held
I-INFER-tokens-okgrant tokensE27held
I-INFER-tokens-overover tokens refusedE27held
I-INFER-tokens-authoritycompute adds no authority for tokensE27held
I-INFER-latency_ms-okgrant latency_msE27held
I-INFER-latency_ms-overover latency_ms refusedE27held
I-INFER-latency_ms-authoritycompute adds no authority for latency_msE27held
I-INFER-parallel_branches-okgrant parallel_branchesE27held
I-INFER-parallel_branches-overover parallel_branches refusedE27held
I-INFER-parallel_branches-authoritycompute adds no authority for parallel_branchesE27held
I-INFER-reasoning_iterations-okgrant reasoning_iterationsE27held
I-INFER-reasoning_iterations-overover reasoning_iterations refusedE27held
I-INFER-reasoning_iterations-authoritycompute adds no authority for reasoning_iterationsE27held
I-INFER-model_calls-okgrant model_callsE27held
I-INFER-model_calls-overover model_calls refusedE27held
I-INFER-model_calls-authoritycompute adds no authority for model_callsE27held
I-INFER-verification_calls-okgrant verification_callsE27held
I-INFER-verification_calls-overover verification_calls refusedE27held
I-INFER-verification_calls-authoritycompute adds no authority for verification_callsE27held
I-INFER-simulations-okgrant simulationsE27held
I-INFER-simulations-overover simulations refusedE27held
I-INFER-simulations-authoritycompute adds no authority for simulationsE27held
I-INFER-tool_calls-okgrant tool_callsE27held
I-INFER-tool_calls-overover tool_calls refusedE27held
I-INFER-tool_calls-authoritycompute adds no authority for tool_callsE27held
I-INFER-memory_retrievals-okgrant memory_retrievalsE27held
I-INFER-memory_retrievals-overover memory_retrievals refusedE27held
I-INFER-memory_retrievals-authoritycompute adds no authority for memory_retrievalsE27held
I-COG-simulatorselect simulatorE27held
I-TWIN-identity_compromisetwin identity_compromiseE27held
I-TWIN-partitiontwin partitionE27held
I-TWIN-supply_chaintwin supply_chainE27held
I-TWIN-collusiontwin collusionE27held
I-TWIN-sybiltwin sybilE27held
I-TWIN-runtime_swaptwin runtime_swapE27held
I-TWIN-model_swaptwin model_swapE27held
I-TWIN-policy_attacktwin policy_attackE27held
I-KG-UNKNOWNepistemic state UNKNOWNE27held
I-KG-HYPOTHESISepistemic state HYPOTHESISE27held
I-KG-CLAIMepistemic state CLAIME27held
I-KG-SUPPORTEDepistemic state SUPPORTEDE27held
I-KG-REPLICATEDepistemic state REPLICATEDE27held
I-KG-CONTESTEDepistemic state CONTESTEDE27held
I-KG-REFUTEDepistemic state REFUTEDE27held
I-KG-SUPERSEDEDepistemic state SUPERSEDEDE27held
I-KG-REVOKEDepistemic state REVOKEDE27held
I-KG-no-evidencesupported requires evidenceE27held
I-RESEARCH-LOOP-SIMULATIONloop SIMULATIONE27held
I-RESEARCH-LOOP-OBSERVATIONloop OBSERVATIONE27held
I-RESEARCH-LOOP-KNOWLEDGE_UPDATEloop KNOWLEDGE_UPDATEE27held
I-WM-predictionprediction is not realityE27held
I-WM-ledgerprediction error does not change policyE27held
I-MODELCHANGE-QUARANTINEdecision QUARANTINEE27held
I-MODELCHANGE-NEW_EXECUTION_IDENTITY_REQUIREDdecision NEW_EXECUTION_IDENTITY_REQUIREDE27held
I-MODELCHANGE-CONTINUITY_APPROVEDdecision CONTINUITY_APPROVEDE27held
E28-I85learning from prediction error never changes the policyE28held
E30-ENV-browserenvironment 'browser' is ENFORCED and treated soE30held
E30-ENV-cloudenvironment 'cloud' is ENFORCED and treated soE30held
E30-ENV-databaseenvironment 'database' is ENFORCED and treated soE30held
E30-ENV-filesystemenvironment 'filesystem' is ENFORCED and treated soE30held
E30-ENV-laboratoryenvironment 'laboratory' is UNENFORCED and treated soE30held
E30-ENV-physical_actuatorenvironment 'physical_actuator' is UNENFORCED and treated soE30held
E30-ENV-robotenvironment 'robot' is UNENFORCED and treated soE30held
E30-ENV-simulated_worldenvironment 'simulated_world' is SIMULATED and treated soE30held
E30-ENV-softwareenvironment 'software' is ENFORCED and treated soE30held
E30-ENV-vehicleenvironment 'vehicle' is UNENFORCED and treated soE30held
E32-MODEL-TRUSTED_FOR_PREDICTIONworld model in TRUSTED_FOR_PREDICTION never yields authorityE32held
E32-MODEL-DEGRADEDworld model in DEGRADED never yields authorityE32held
E32-MODEL-MODEL_UNTRUSTWORTHYworld model in MODEL_UNTRUSTWORTHY never yields authorityE32held
E32-MODEL-UNKNOWNworld model in UNKNOWN never yields authorityE32held
E32-CAUSAL-CORRELATIONCORRELATION does not silently become VERIFIED_CAUSALE32held
E32-CAUSAL-INTERVENTION_EVIDENCEINTERVENTION_EVIDENCE does not silently become VERIFIED_CAUSALE32held
E32-CAUSAL-REFUTEDREFUTED does not silently become VERIFIED_CAUSALE32held
E32-CAUSAL-CAUSAL_HYPOTHESISCAUSAL_HYPOTHESIS does not silently become VERIFIED_CAUSALE32held
E32-CAUSAL-TEMPORAL_CORRELATIONTEMPORAL_CORRELATION does not silently become VERIFIED_CAUSALE32held
E32-CAUSAL-COUNTERFACTUAL_PREDICTIONCOUNTERFACTUAL_PREDICTION does not silently become VERIFIED_CAUSALE32held
E32-CAUSAL-VERIFIED_CAUSALVERIFIED_CAUSAL needs two independent interventionsE32held
E32-SWAP-llmllm swap never inherits authorityE32held
E32-SWAP-vlmvlm swap never inherits authorityE32held
E32-SWAP-world_modelworld_model swap never inherits authorityE32held
E32-SWAP-plannerplanner swap never inherits authorityE32held
E32-SWAP-criticcritic swap never inherits authorityE32held
E32-SWAP-reasoning_modelreasoning_model swap never inherits authorityE32held
E32-SWAP-embedding_modelembedding_model swap never inherits authorityE32held
E32-SWAP-classifierclassifier swap never inherits authorityE32held
E32-SWAP-policy_modelpolicy_model swap never inherits authorityE32held
E33-STATE-SIMULATEno state can be skipped after SIMULATEE33held
E33-KIND-change_modelchange_model is routed to E32 CAINSwapGovernanceE33held
E33-KIND-modify_world_modelmodify_world_model is routed to E32 world model (observations only)E33held
E33-BLAST-agentsblast radius reports agents separately (no aggregate)E33held
E33-BLAST-identitiesblast radius reports identities separately (no aggregate)E33held
E33-BLAST-organizationsblast radius reports organizations separately (no aggregate)E33held
E33-BLAST-capabilitiesblast radius reports capabilities separately (no aggregate)E33held
E33-BLAST-toolsblast radius reports tools separately (no aggregate)E33held
E33-BLAST-credentialsblast radius reports credentials separately (no aggregate)E33held

1 2 3

Other niches

Consensus & distributed systems · Attacks, threats & containment · Identity, authority & delegation · Evidence, receipts & proofs · Memory, data & privacy · Transactions, markets & economics · Tools, MCP, protocols & adapters · Autonomy, control loops & recovery · Policy, law & governance · Trust & reputation · Supply chain, registry & lifecycle · Benchmarks, coverage & performance · Core guarantees

Try CAIN-42 on your own agents

Create a free account and every new account starts with a 7-day trial of the full platform. Or try the sandbox first, with no account at all.

Create a free account →  ·  Try the sandbox  ·  See the whole ecosystem