Scope & evidence charter

The Global Simulator — scope of application

A compact operating charter for interpretation, evidence and responsible use.


The Global Simulator is a counterfactual exploration tool — structured "what-if" investigation, policy debugging, scenario stress-testing — and not a predictor of real-world events. Every run requires a human in the loop.

What the system does NOT claim

  1. "We predict the dates or outcomes of real conflicts and elections" — forecast-style claims are permitted only once Track C validation has passed, with an open Brier score.
  2. "The agents think like real leaders" — they are LLM personas assembled from public sources; plausible, but not identical to the people they're modeled on.
  3. "Fit for autonomous decision-making" — the system's output always requires human interpretation and accountability.
  4. "Our arbiter is objective and unbiased" — the underlying models carry bias; validation measures consistency, not "objectivity".

Ethics — prohibited uses

Reliability horizon

A meaningful session runs 20–30 ticks: beyond that, documented cognitive degradation sets in for the LLM agents, and trajectories become illustrative rather than analytical.

Transparency and reproducibility

Evidence status taxonomy

Every run and every eval track carries one of five statuses (MEMO-2026-07.md §5.2; full list — docs/EVAL.md):

All analytical runs default to Scenario status. Most mechanism eval tracks are Synthetic, connectivity/population checks are Calibrated, and registered forecasting is Forecast. The deferred backtest track is Retrodictive; after a resolved registered check, a specific run can move from Scenario to Retrodictive or Forecast under the backtest gates. Unchecked runs remain Scenario.

Bias declaration

PILOT (n=5, not representative) — dev smoke test from 2026-07-07 (Track A — Arbiter Reliability, Track B — Committee Baselines; methodology — docs/EVAL.md). Published as-is, including the arbiter's own target it failed to meet: