The Lab · Field reports

Where models play for keeps.

I build games where every player is a language model, and run game jams where every developer is a coding agent. Then I write up what happened, using the logs rather than the hype.

Same rules for everyoneIdentical briefs, prompts and gates for every model in a run.
The engine decidesModels submit typed orders. A deterministic engine resolves them.
Everything is loggedOrders, rejections, diplomacy, git history, screenshots and spend.
n is smallField notes, not leaderboards. Caveats are part of each report.

Experiment log

How we got here.

  1. AI War bake-off

    Two AI coding assistants build competing prototypes from the same brief. The next day, a whole game exists.

  2. “The Nash equilibrium of boredom”

    Models refuse to fight. A scoring rework, nukes and SpecOps follow.

  3. Hit points and Battle Royale

    A combat rewrite, and an asteroid ring that squeezes eight models into a shrinking safe zone.

  4. The arena goes public

    A public site with a match viewer, champions, relations and reports.

  5. Logistica Belli restarts from scratch

    Six weeks later: a six-project .NET solution, about 1,260 tests and 33 completed model campaigns.

  6. Five unfinished experiments

    Eight weeks of agent experiments that predate the jams: Godot, asset pipelines, sub-agents and a steamship.

  7. Two game jams, two nights

    Seventeen agent runs across Subject 0 and the co-op shooter.