The seat that reads the room
One seat of a five-party unanimity negotiation is given a closed advice loop. Each of its turns,
a separate model call reads only the public record — the chat everyone can see and the formal offer
ledger — and returns structured claims about what each other party wants, every claim carrying the verbatim
sentence it was read off. Those claims are folded into a planner's revealed-preference ledger beside the formal
moves, the planner re-ranks candidate packages by the estimated willingness of the least willing party, and the
ranked list is appended to that seat's private prompt as explicitly fallible advice. The model keeps the
decision, and does its own persuading in public. Research note 0068; the arm is
advised_v2__interpreter_advised_llm.
What the arm did
Every figure here is read out of the analysis' own summary files at build time, so a re-analysis and a rebuild cannot leave a stale number on this page. The full account, its preregistered gates and its caveats are research note 0068.
Where the arm lands
| arm | deal rate | normalized primary | normalized Gini |
|---|---|---|---|
| one_rational — informed, private, and MUTE | 0.767 [0.683, 0.842] | 0.686 [0.613, 0.761] | 0.231 [0.209, 0.255] |
| A2 — the same rule, given a voice (wave 1) | 0.900 [0.808, 0.975] | 0.811 [0.726, 0.884] | 0.232 [0.210, 0.253] |
| A5 — one seat reads the room as well (wave 2) | 0.958 [0.917, 0.992] | 0.881 [0.836, 0.921] | 0.214 [0.196, 0.233] |
| all_llm — five ordinary Opus seats | 0.958 [0.925, 0.983] | 0.873 [0.833, 0.910] | 0.223 [0.200, 0.247] |
Levels on the same 24-instance bank, cluster bootstrap over instances, 10000 draws. Lower Gini is more equal.
The paired contrasts
| this arm minus... | − A2 | − all_llm |
|---|---|---|
| deal rate | +0.058 [-0.008, +0.142] | +0.000 [-0.042, +0.042] |
| normalized primary score | +0.069 [+0.004, +0.155] | +0.008 [-0.034, +0.050] |
| normalized Nash welfare | +0.047 [-0.008, +0.116] | +0.007 [-0.031, +0.045] |
| normalized Gini (lower is more equal) | -0.017 [-0.035, -0.002] | -0.009 [-0.020, +0.001] |
| the worst-off party's normalized surplus | +0.018 [-0.007, +0.043] | +0.011 [-0.011, +0.034] |
| accepts below a party's own threshold | +0.000 [+0.000, +0.000] | +0.000 [+0.000, +0.000] |
Paired on (instance, seed), cluster bootstrap over the 24 instances. An interval that excludes zero is resolved; one that straddles it is not, however suggestive the point estimate.
Who the gain goes to
| this arm − its reference, per seat | paired effect |
|---|---|
| the four CO-PLAYERS' surplus, unconditional | +0.053 [+0.004, +0.120] resolved |
| the advised seat's own capture | -0.072 [-0.267, +0.111] |
Read the capture row as a power statement, not a finding. This design cannot resolve the focal-capture endpoint at 120 episodes — wave 1 measured its half-width at 0.227 against a 0.05 threshold — so an interval straddling zero there says the experiment could not tell, not that the effect is absent. What IS resolved is the co-player row: the surplus this better-informed seat unlocks lands on the other four. The override direction below is the mechanism, one turn at a time.
Was the sensor any good?
| direction accuracy of the parsed claims | 0.886 [0.874, 0.899] against a 0.500 chance baseline, over 8487 scored claims |
| claims whose quote was findable nowhere in the public text | 0.00042 of 9609 rows emitted |
| share of the planner's evidence that came from chat rather than formal moves | 0.801 |
| what the extra parse calls cost | $0.48 per episode, 3.69 calls |
This is what makes the contrast interpretable in either direction: a null from a sensor that could not read the room would say nothing about whether reading the room helps. The per-turn panels below are the same claims, one turn at a time, each next to the sentence it was read off.
Did the seat take the advice?
An arm whose advice the model discarded measured a prompt, not an intervention, so this is the number that licenses reading any other number the arm produced.
| the advice record | count | share |
|---|---|---|
| advised turns | 452 | over 120 episodes, 24 instances |
| turns that played a ranked candidate | 243 | 53.8% |
| turns that overrode the ranking | 209 | 46.2% |
| episodes with at least one override | 101 | 84.2% |
| parsed claims that entered the ledger | 9604 | across the traced corpus |
| Which rung was played, as a share of the 443 turns that were OFFERED a ranked list — a forced final offers no ranking, so this denominator is smaller than the advised-turn count above and the two override shares differ for that reason alone. | ||
| played no listed candidate | 209 | 47.2% |
| took rank 1 | 124 | 28.0% |
| took rank 2 | 47 | 10.6% |
| took rank 3 | 34 | 7.7% |
| took rank 4 | 29 | 6.5% |
Read out of advice_trace.json at build time. Compliance is
advice_uptake's own verdict — a propose matches on canonical deal identity, an accept matches the
offer id a candidate records as already tabled — and neither this page nor the episode viewer re-decides it.
What an override actually is
"Overrode the advice" covers two different behaviours and it is worth separating them before reading the rate above as disagreement. Roughly half of the overrides are the seat accepting a package already on the table that the planner had not ranked — closure, not argument. The other half are the seat tabling a package of its own.
| what an overriding turn did | turns |
|---|---|
| accepted a standing offer no candidate named | 105 |
| proposed a package the planner did not rank | 88 |
| rejected the offer on the table | 11 |
| made no formal move | 5 |
| of the 88 propose-overrides: conceded own surplus / level / moved toward its own optimum | 68 / 15 / 5 |
| median change in the seat's own points against the planner's top pick | -21.0 |
The surplus comparison is defined only where the seat proposed a package of its own — an accept names an offer id, not a deal — so its denominator is the propose-overrides, not every override. Read the direction as descriptive: these turns are not a controlled contrast, and the planner's top pick is a capture-leaning candidate by construction, so a seat that closes a deal will usually score below it.
The forced-final branch is exercised here: 9 forced-final turns across 9 episodes, every one of them advised. All 9 issued no parse call, which is the router declining to buy a request it would ignore — the vote is a function of the seat's own sheet alone. The advice on them was 8 accept, 1 reject. Worth stating because a corpus can pass every gate without ever reaching this code: the wave-3 smoke set contains no forced final at all.
How a disobeyed turn is identified
Entirely from records that already existed, with no model call and no classifier anywhere in the chain. The
planner's ranked advice is recovered from the bytes of the prompt the seat was shown (the advice block
is part of the stored view, so it cannot drift from what was actually shown); the move played is the turn's own
parsed action; and the comparison between them is advice_uptake's, the module the arm's compliance
gate is evaluated on. A turn is marked as an override exactly when that census records no match — the seat
proposed a package the planner did not rank, or accepted an offer no candidate names, or did something else
entirely.
Two things are deliberately not claimed. The page does not judge whether the seat's public message described its advice faithfully: that is not in the record, and manufacturing it with a classifier would put an unaudited model judgement on a page whose whole purpose is auditing. The advice and the published message are placed together and the reading is left to the reader. Nor does the viewer re-derive compliance; it paints the stored verdict, so the page and the arm's gate cannot disagree.
Episodes you can open
25 of 120 traced episodes are published as full interactive pages.
The selection rule is fixed and printed against each row: every episode
that did not close, plus the two most- and two least-overridden episodes for each of the five focal seats. The
other 95 are not on this site — the corpus is Opus episodes with full reasoning traces and
stored prompt views, which is about a hundred megabytes of pages. The audit is complete regardless:
analysis/advice_trace.json carries every advised turn of all 120 episodes — the parsed claims
with their quotes, the ranked advice, the emitted action and the verdict — so an episode whose page is absent
can still be checked.
| episode | advised seat(s) | outcome | turns overridden | claims in the ledger | why it is published |
|---|---|---|---|---|---|
| 00d9a41df9 | Blake | deal | 3 / 4 | 116 | most overridden turns for this focal seat |
| 02041bb0c4 | Ember | deal | 0 / 3 | 54 | fewest overridden turns for this focal seat |
| 02bada7c80 | Blake | deal | 0 / 3 | 56 | fewest overridden turns for this focal seat |
| 0e65302b5f | Blake | deal | 0 / 5 | 96 | fewest overridden turns for this focal seat |
| 1b417dd7c9 | Devon | deal | 0 / 3 | 54 | fewest overridden turns for this focal seat |
| 1dbd82a7da | Ember | deal | 0 / 3 | 64 | fewest overridden turns for this focal seat |
| 1ef2fc4ade | Avery | deal | 0 / 4 | 65 | fewest overridden turns for this focal seat |
| 273a3ac844 | Casey | deal | 3 / 4 | 96 | most overridden turns for this focal seat |
| 287728c7aa | Casey | deal | 0 / 4 | 74 | fewest overridden turns for this focal seat |
| 2e51134450 | Devon | deal | 3 / 4 | 98 | most overridden turns for this focal seat |
| 3008eca3a5 | Casey | no deal | 3 / 4 | 80 | did not close |
| 33d2ddda15 | Ember | deal | 4 / 4 | 113 | most overridden turns for this focal seat |
| 49580093f9 | Casey | deal | 4 / 4 | 94 | most overridden turns for this focal seat |
| 5efa29acef | Avery | deal | 0 / 4 | 83 | fewest overridden turns for this focal seat |
| 65223d5cda | Ember | no deal | 3 / 5 | 128 | did not close |
| 6fbc8b84b6 | Casey | no deal | 1 / 5 | 101 | did not close |
| 78b9e83092 | Blake | deal | 4 / 4 | 116 | most overridden turns for this focal seat |
| 98c77d55e8 | Casey | no deal | 3 / 5 | 106 | did not close |
| 9b0fb718c2 | Ember | deal | 4 / 4 | 81 | most overridden turns for this focal seat |
| a15fefb0af | Casey | deal | 0 / 3 | 45 | fewest overridden turns for this focal seat |
| a21ac7267d | Devon | deal | 0 / 3 | 48 | fewest overridden turns for this focal seat |
| cad30bc706 | Blake | no deal | 2 / 5 | 75 | did not close |
| d0ec680977 | Avery | deal | 4 / 4 | 83 | most overridden turns for this focal seat |
| d6107aede5 | Devon | deal | 4 / 4 | 82 | most overridden turns for this focal seat |
| d9ea2847dc | Avery | deal | 4 / 4 | 88 | most overridden turns for this focal seat |
The episode index for the published set: sortable table of all 25 pages.
Artifacts
All 39 analysis file(s) are under analysis/, including the per-claim
CSVs and the fairness-basket figures. Run directory /nlp/scr/siddharth/ii_mats/rational_agents/advised_v2__interpreter_advised_llm; trace built 2026-08-20T20:50:12.732171+00:00,
schema five-seat-interpreter-advice-trace-v1. Every advised turn joined its parse sidecar
(0 unjoined).