Educational companion dossier · Fact, interpretation, lived experience, clinical education, fiction, and mechanics are labeled separately. Scope & safety
REAL-WORLD INTERPRETIVE

SIMULATION LITERACY

Calibration, Validation, and Sensitivity

How to tune a model, test it against evidence, and reveal which assumptions drive its conclusions.

LEVEL 1

ORIENTATION

Why this matters

REAL-WORLD INTERPRETIVE

One-sentence brief

A model that reproduces one historical curve may still fail elsewhere or for the wrong reasons. Validation and sensitivity prevent a fit from becoming false authority.

REAL-WORLD INTERPRETIVE

Three key points

  1. Calibration selects parameters; validation tests performance on evidence not used for tuning.
  2. Sensitivity asks how conclusions change when assumptions change.
  3. Passing a bounded test is evidence within scope, not proof of reality or production readiness.
LEVEL 2

WORKING BRIEF

Evidence, context, and limits

REAL-WORLD INTERPRETIVE

Calibration

Calibration estimates or selects parameters so model outputs align with chosen data or constraints. It should document targets, loss functions, search ranges, prior assumptions, and overfitting risk.

  • Separate calibrated from fixed parameters.
  • Use data quality and uncertainty in the objective.
  • Do not hide manual adjustments.
REAL-WORLD INTERPRETIVE

Validation

Validation compares the model with observations, holdout periods, alternative datasets, or known accounting identities. A model may be useful for one purpose and invalid for another.

  • Define success before testing.
  • Use out-of-sample or forward evaluation where possible.
  • Report failures and subgroup gaps.
REAL-WORLD INTERPRETIVE

Sensitivity analysis

Sensitivity analysis varies inputs, structures, and rules to identify what controls the result. Local sensitivity tests small changes; global sensitivity explores combinations and interactions.

  • Prioritize assumptions with high impact and low evidence.
  • Show direction and magnitude of change.
  • Do not present one parameter set as inevitable.
REAL-WORLD INTERPRETIVE

Identifiability and equifinality

Different parameter combinations or mechanisms can produce similar outputs. A good fit does not automatically identify the true causal process.

  • Compare multiple plausible explanations.
  • Use diagnostic observations, not only aggregate fit.
  • Mark parameters that cannot be estimated from available data.
REAL-WORLD INTERPRETIVE

External validity

A model calibrated in one period or jurisdiction may fail under different institutions, mobility, climate, access, reporting, or behavior. Transfer requires new evidence and review.

  • Recalibrate or bound transfer claims.
  • Do not use one country as the universal baseline.
  • Publish where the model has not been tested.
REAL-WORLD INTERPRETIVE

Release decision

Automated validation can support a release gate, but it does not replace legal, ethical, accessibility, scientific, regional, or lived-experience approval.

  • Keep human approval state explicit.
  • Rerun after material changes.
  • Preserve failed baselines as regression evidence.
LEVEL 3

COMPLETE DOSSIER

Limitations, game links, and review context

DISPUTED / MULTIPLE ACCOUNTS

Known limitations and gaps

  • This is a non-operational educational transformation: it does not build or implement a game, executable simulator, forecasting service, emergency tool, or decision system.
  • This page is an educational transformation of supplied research leads; it does not authenticate every citation, equation, product claim, or institutional attribution in those files.
  • A scenario, model run, or map is not an observation, forecast, legal finding, public-health instruction, emergency warning, or proof of future behavior.
  • No source code, nuclear-effects formula, casualty calculation, target-selection method, cyber-intrusion procedure, exploit chain, or instruction for bypassing safeguards is published.
  • Geography, nationality, ethnicity, religion, language, migration, disability, health, poverty, or political identity are not inherent danger, compliance, intelligence, competence, or worth variables.
  • Specialist scientific, public-health, accessibility, legal, regional, ethics, security, and lived-experience review remains pending; the page stays open to correction.
LEVEL 4

RESEARCH EDITION

Sources, methods, and stable links

REAL-WORLD INTERPRETIVE

Linked reports

REAL-WORLD VERIFIED

Method and corrections

This page follows the public method for provenance, confidence, source independence, alternative accounts, limitations, review state, and visible correction.

NEXT

CONTINUE

Related learning

Page complete Calibration, Validation, and Sensitivity Page label: REAL-WORLD INTERPRETIVE