MERIDIAA TheoremPath project
The recorded Meridia country in relief, with mountains, water and illustrative settlements.

Simulated worlds. Scientific questions.

Meridia

A world to investigate.
Evidence to question.

Test how AI systems work with incomplete evidence, against a world whose history you can inspect.

Enter the world
01 / Recorded world

Seed 4711 · Initial landscape
Display relief exaggerated · Building glyphs illustrative

Explore

01 / The observatory

Examine the world.

A recorded country, its residents and the methods used to study it. Explore the evidence at each scale. No live agent runs on this page.

Loading the recorded world…

02 / A measured experiment

Same worlds.
Different missing evidence.

Income can influence who answers a survey and whose income is left blank. Compare four recorded settings across twelve worlds, with the same underlying truth in every paired comparison.

Separate study · Worlds 74001–74012 · Bayesian county adult-income estimates

Loading the recorded experiment…

What this comparison establishes

Each paired arm shares retained truth, source records, sampled households and keyed random draws. When an income dependence is switched off, its corresponding mean probability is matched to the reference; realized rates can still differ.

Signed error is 100 × (estimate − truth) / max(|truth|, 1), averaged equally over finite county targets and then equally over twelve worlds. Mean absolute percentage error takes the absolute value before averaging. The estimator and its settings stay fixed.

This diagnoses an observation mechanism in twelve development worlds, with one survey draw per arm. It is not a fitted correction, a confirmation on fresh worlds or an agent benchmark. The keyed instrument differs from the legacy version, so this view compares only keyed arms.

A claim you can inspect.
All four settings, exact world-level values and source hashes.

Download experiment values

03 / The research process

Separate what happened
from what was observed.

Meridia generates events and imperfect records of those events. A task specifies what a method can see. Retained truth lets researchers inspect its conclusions.

  1. 01

    Generate a world

    People, households, institutions and a dated event history.

  2. 02

    Define the evidence

    Declare the observation rules, missing information and available records.

  3. 03

    Run a method

    Use a specified packet and output format. Preserve its result and runtime.

  4. 04

    Inspect the result

    Compare estimates and supporting evidence with the recorded world.

Scope of this demonstration

A model, with limits.

This demonstration contains a simplified population and institutional system. Its history is replayed; actions do not change the future. Disease transmission, a coupled climate system and live agent execution are not part of this example.

Current research focuses on observation mechanisms, uncertainty and reproducible agent evaluation. Performance in a generated world alone does not establish performance in a real country.

04 / Research with Meridia

Bring a question
worth testing.

We’re developing Meridia for teams studying how agents work with evidence and uncertainty. Start with one workflow and a result your team can inspect.

Discuss a research pilot