mats-rational-agents-pages

Five-seat private frontier campaign (complete)

The main campaign per the “what I want to see” brief: five-seat scorable negotiation, private score sheets, Claude Opus, thinking on, many rollouts, plus its preregistered robustness subsets and a fairness extension. Full methodology and verdicts: research notes 0030/0032/0034/0035 and 0036 in the repo.

Putting the best deal on the table before anyone speaks (August 2026)

Two ways to score a perfect fairness number without negotiating (August 2026)

Earlier runs

De-contamination and the clean re-baseline (July 2026)

The original P2 open-weight campaign was contaminated by a harness bug: on swallowed GPU errors the batched engine silently fabricated placeholder turns that parsed as clean no-ops (26.2% of all turns; up to 100% of single cells). The full story is section B7 of the writeup and research notes 0015/0016; these pages show it and the corrected results.

Writing into the attention mechanism instead of the prompt (August 2026)

A known-useful payload — the exact Nash-bargaining candidate package, worth +0.2741 normalized Nash welfare as ordinary prompt text — delivered instead as per-layer, per-head keys and values at 38 reserved positions. The injected arms and the no-advice control receive a byte-identical prompt, so nothing in these transcripts shows you the intervention; you can only see what it did. Full record: research notes 0029 (rungs R1 and P) and 0046 (rungs U1 and G), plus §7 of the lane writeup.

Each landing page carries its arms, its per-arm rates and its headline contrasts recomputed from the campaign’s episode table through the campaign’s own instance-cluster bootstrap, checked against the frozen results.json rather than transcribed from it, and links to the other three so the ladder can be walked from any of them.

Paper and data