clay-borg/simulators/README.md
tegwick b590e7fd59
Some checks failed
ci / check (push) Failing after 4s
simulators/: persist the survey, and let it shrink two of the three tracks
Eight profiles on a common schema, each marking what was checked against a
source this session and what is background recollection. Three are marked
unverified in full — Machinations, the play substrates, most of RBG — and
say so rather than reading as evaluations. Written straight after three
review rounds whose entire yield was claims outrunning what had been
checked, so the confidence rule is the first thing in the README.

The survey changed the plan, which is what a survey is for.

Track C was described in Positioning as open ground. It is not: Browne
published 57 criteria for game quality, and Ai Ai already computes
designer-facing measures — drama, lead changes, branching factor,
completion, duration — from played games. The track becomes adopt, credit
and find the gap. The gap looks real: those measures presume a leader, and
SHARED GROUND has none — Modes.csv gives its tiebreak as "Not applicable".

Track B probably adopts rather than builds. OpenSpiel implements CFR,
best-response and exploitability over games that are simultaneous-move,
imperfect-information and co-operative, which is all four of GROUND's
awkward properties. "Does ATTACK ever pay" is a best-response question,
and we spent three review rounds refining a two-policy sweep for it. The
first Track B task is now one question — is exploitability meaningful for
a co-operative game with a shared threshold — not a build.

The cost of not surveying earlier is therefore measurable, and is recorded
rather than glossed.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
2026-08-08 11:56:27 +02:00

3.1 KiB

Simulators — profiles of systems adjacent to clay-borg

Persisted research, not a link dump. Each profile answers the same questions in the same order so the systems can be compared, and each carries what we actually checked.

Written for ../specs/Positioning.md §1 and §4 — the field we sit in, and the three tracks. Track A (a second game) and Track B (game theory as the lens on dynamics) both depend on knowing what already exists.


The confidence rule

Every claim in a profile carries its source, and a claim from background knowledge is marked unverified.

This directory was written immediately after three adversarial review rounds whose entire yield was claims that outran what had been checked (../reviews/). A profile that quietly mixes a fetched fact with a half-remembered one is the same failure in a new place — and here it would be worse, because these profiles are meant to inform whether we build something or adopt it.

marker means
checked read from a source this session, cited at the foot of the profile
unverified from background knowledge; plausible, not confirmed, and must be before it decides anything
open a question we want answered and did not answer

The schema

# <name>
owner · first release · licence · language          (checked/unverified)
## What it optimises for
## How a game is defined
## What it can analyse
## What it does not do
## Relevance to clay-borg      — which track, and how
## What to steal
## What to avoid
## Open questions
## Sources

Index

profile in one line track
ludii.md the closest relative: ludemes, breadth, speed A
gdl-ggp.md the academic ancestor; agent generality over designer support A
rbg.md regular boardgames — a speed-focused rival description language A
ai-ai.md the closest relative for what we want to become: GGP with designer-facing metrics A, C
openspiel.md the game-theory workbench: CFR, exploitability, social dilemmas B
browne-criteria.md not a simulator — the canonical attempt to measure whether a game is good C
machinations.md economy and feedback loops, diagrammatic
play-substrates.md Tabletop Simulator, boardgame.io — play without analysis

What this survey has already changed

Ai Ai and Browne's criteria are the finding. Track C was written in Positioning.md as though "assimilate knowledge about why games work" were open ground. It is not: Browne published 57 criteria for game quality, and Ai Ai already computes designer-facing measures — drama, lead changes, branching factor, completion, duration — from played games.

That does not close Track C. It moves it from invent to adopt, credit, and find the gap, which is a much better position and a much smaller claim than the one we were about to make.