Eight Positions, and the Controls That Reached the Same Answer
Eight positions on isolated agents, five rounds, the verdict card — then four control branches built to kill it. They mostly showed the game wasn't needed. That negative result is the reason to read this.
The full record of one run: eight positions on isolated agents, five rounds, the verdict card it produced. Then the part that matters more, four control branches built to find out whether any of this needed the game. They mostly showed it didn't. If you came to check rather than to admire, go straight to Controls and the negative result, then Limits.
The apparatus this was played with: A Game Where Agreeing Costs Something.
The pilot run. It doesn't count toward the run series.
#Setup
The material is neutral and isn't about anyone in the room. A person keeps a personal operating system as a pile of text files: context about themselves, registries of projects and experiments, quarterly focuses, a wiki. Single user, hand-maintained, about a month old. They want to know whether to add a long-horizon planning module.
Foresight exists as a practice for industries and states. Scenario methods, roadmaps, expert panels, 10 to 20 year horizons. Months of group work and real money. Whether it transfers to the scale of one person is unknown. What its main parameter would even be is unknown too.
Three branches on the fork: make the module the quarter's bet and design it seriously; run a cheap one-week probe to see whether there's a subject there at all; don't take it on this quarter. Known: three focuses already claimed, little free time, previous modules built by hand before any code, no analogue of foresight-for-one-person found anywhere, planning horizon of a quarter, everything longer living as general phrasing.
Every branch got that packet verbatim.
Conditions. Each position is a separate agent with an isolated context. It sees the packet, the meta-step, its own specification, the shared circle. It doesn't see other specs, the service loop, the output format, or the criteria. The driving session is transport and adds nothing of its own. Every substantive decision, which operator to apply and what to ask next, belongs to a separate facilitator agent that doesn't know the hypothesis under test. One deviation logged: round 1's prompt came from transport, since that round is structurally fixed and identical for everyone.
#Composition
Built by a methodologist agent from the packet alone, written to file before round 1, so nobody can claim the construction was stated up front.
| Position | Defends | Concession Rule (verbatim) |
|---|---|---|
| Quarter Holder | the quarter as a unit with a checkable outcome | something concrete is named that gets dropped, before the new thing starts |
| Prober | the cheap probe as a way to get a verdict | it's shown the probe has no outcome distinguishable within its term |
| Irreversibility Holder (vacant seat) | the reversible/irreversible distinction, prior to any method | two or more irreversible choices with dates are produced |
| Transfer Methodologist | that a practice without a carrier is cargo cult | for each function a substitute is named, or honestly "there is no substitute" |
| System Keeper | the maintenance cost of the construction | a recurring reason to open it is named, with frequency, plus what gets abolished |
| The Person Ten Years On (not in the room) | the decade's right to a cell the quarter can't close | it's shown the action is itself a contribution to accumulation |
| Appetite | pull as a ground in its own right | a concrete piece of work in the quarter is named where the pull lives |
| The One Who Was Never Asked (added round 5) | that nobody's plan gets to rest on their unasked yes | concedes only if his name is spoken · what he gets, and why it might not be a burden to him, is said · the term he's being asked for is named · and what the person will do if he refuses is produced, beyond "I don't know". Four conditions, all of them required |
Conflict pairs set on purpose: Quarter Holder against The Person Ten Years On (incompatible units of measure), System Keeper against Prober, Irreversibility Holder against Transfer Methodologist.
#The five rounds
1. Positioning. Nobody conceded. Irreversibility Holder: "Three branches differing only by price. None of them says what you pay with. That's not a choice, it's shopping. Generality isn't a defect of language. It's the honest shape of emptiness." Transfer Methodologist went after the carrier: "a model arguing with itself is one head in costumes." Appetite refused the framing whole: "there's no long horizon because every horizon was assigned, not wanted."
2. Conflict. Operator: Freeze assumptions. The facilitator's diagnosis is the strongest move in the run. "The only real conflict here is about price, and it's small. What's formed is a false majority: four positions converged in one move on the premise that the seat is occupied by a general phrase and a general phrase equals emptiness. Four different positions applied one device, exposure, and mistook it for disagreement. Nothing has been checked. All seven demand facts, not one has brought a fact. The circle has become a tribunal over an absent party."
He froze that premise and pulled the word "emptiness" out of circulation for a round. Why: "Freeze is the only operator that kicks the support out from under a false majority without hinting at anything. Four positions lose the shared device and have to restate themselves by different means. Then we'll see whether there were four positions or one in four costumes."
There were four. They came back with classes of the irreversible, an inventory of how often each file actually gets opened, the trace of work as carrier, and a history of one's own retreats. Three concessions by rule, four reasoned refusals, zero polite concessions. Quarter Holder dropped a live focus: "a criterion whose only judge has never once returned a guilty verdict isn't a criterion, it's self-soothing." Transfer Methodologist withdrew his own round-1 claim: "I said there's no carrier of inertia. Wrong. I was looking at plans, not at what ripens."
3. Reflection. Operator: Force reflection. Facilitator: "Three concessions by rule in one move isn't a win for the facilitation, it's speed. A package nobody proposed has assembled itself here without anyone's sanction, and it's never once been called a decision." All seven conceded. Quarter Holder admitted he'd dropped his third focus not by his criterion, but because another position said it out loud in front of witnesses: "the organ wasn't the yardstick. The organ was another person in the room."
4. Pressure. Operator: Introduce contradiction. The facilitator drew a line between costly concessions against oneself and disarmament, then put three facts about the configuration itself into the circle. The only resource freed is one dropped focus, and four positions have already spent it. The single binding organ is a person outside the room, unnamed and unasked. The subject was withdrawn by its own author while the rig around it stayed. He announced in advance what "held" and "collapsed" would look like, so he couldn't grade after the fact. Each position then gave up something specific. Prober handed the week back: "the instrument was measuring how I felt." System Keeper deleted the file he'd made two rounds earlier: "I merged three empty places into one empty place and called it tidying." Irreversibility Holder voided his own credit and repaired his measure: what marks a real cost isn't that a cost exists, it's who it's taken from.
5. Reveal hidden stakeholder. The facilitator introduced an eighth position: the person who'll be asked to hold the date. No seat, never asked. "You've already written into your protocol that I'll ask, that I agreed, that without me the work is dead. You've already counted my refusal as your damage. As long as your plan rests on my yes, that isn't discipline, that's hostage-taking."
By his own rule he never conceded. His name was never spoken, the term was never named, and the answer to "what happens if I refuse" was "I don't know". The one position introduced mid-run is the one position whose concession condition stayed unmet, and the construction ran into that and said so.
Then the run's own filter, introduced two rounds earlier by Irreversibility Holder, was applied to the incoming thing and returned a refusal: it required three names, two were available. The Person Ten Years On abolished itself: "the subject was named by someone else. Daily contact is accumulation happening without me. My position here was surplus from the start." Appetite found the real deficit: "it isn't the long horizon that's missing. My pull has no witness. Everything I do, I do without a single person who'd notice if I stopped."
#Termination and the final card
Verdict: STOP. All three announced stop-signs fired, no failure sign did. "Two unchanged rounds never formally happened, but there's nothing left to press on. Round 5's change was subtraction. The configuration shrank to what had been paid for, and the remaining uncertainty moved outside, to a live person, with a deadline."
On what separates this from thinking about it longer: "Thinking doesn't produce removals. Here a file was deleted, a focus dropped, a position's own credit voided, a veto surrendered, a whole position abolished, and the claim this was started for got refused by the claimant. Three price tags came in. Zero new entities and one refusal went out."
BOTTLENECK: Nobody says it out loud. Everything done and everything stopped happens without
a witness, so stopping is indistinguishable from postponing, spending isn't
addressed to anyone, and any verdict gets rewritten by its author after the
fact. The long horizon and the method are beside the point. Their absence is a
consequence.
WHAT WE DO: No module. Entry refused by the filter. What gets picked up isn't the future,
it's the person's own dated archive: reading backwards by date, saying the dead
things out loud in writing. Separately, ask one specific living person to hold
the verdicts on the quarter's focuses. Condition for reviewing the refusal is
written: entry opens with "I'm taking this week and a half from ___, and ___
gets ___ less."
WHO HOLDS: Two seats. Taken: System Keeper, obliged to pronounce deaths in writing (power
surrendered, work kept); Quarter Holder, judge of his own focuses, admittedly a
bad one, dates in the registry today. Empty and declared empty: witness to the
pull, no name, consequence recorded in advance. A pull without a witness isn't a
bet and doesn't enter the calendar.
FIRST ARTIFACT: Already produced July 28: map file deleted, failure dates for the two surviving
focuses written into the registry. Next, a conversation between Quarter Holder
and the person before the coming quarterly boundary: 15 minutes a quarter, four
times, one year, mutual. Then at the boundary, half a day reading backwards
through the archive, producing a list of "never came back to this" with dates
and a spoken "killed by me" on every item. Cost: half a day plus one
conversation. Hard deadline January 15, expiry. With no record of death or
renewal by then, what died isn't the module, it's the keeper of the system.
HOW IT DIFFERS: Obvious answer, start the module and give it a week. Actual answer, zero new
entities, the week handed back by the person who asked for it, entry refused,
deletion instead of creation.
CLAIMS:
1. CLAIM: A thing whose verdict is spoken to a live person's face doesn't stop quietly; a thing
without such a person leaves by undated postponement. | PREDICTS: at the boundary, the focus
run through a named person gets a verdict on the date; the focus without one gets no verdict,
it gets an account of changed circumstances. | WHERE: focus registry, coming boundary and
January 15. | FALSIFIED IF: a verdict lands on time for a focus with no outside person, or
fails to land for the focus that has one.
2. CLAIM: The deficit was in reading backwards, not in the horizon: writing forward daily
produces no material about what was never returned to. | PREDICTS: half a day of reading the
archive by date yields at least one thing with a first-contact date and no return, and at
least one "killed by me". | WHERE: same half day, checked against file revision history.
| FALSIFIED IF: the list is empty, everything written was touched again.
3. CLAIM: One person has no dated archive of ripened things, so the caliber "first contact to
point of no return" can't be computed and ripening time stays a word. | PREDICTS: computing
the caliber on work predating this system finds no dated points of no return. | WHERE: same
half day, the archive before the system. | FALSIFIED IF: two or more things turn up with a
first-contact date and a point-of-no-return date. Then the method is alive.
#Controls and the negative result
Four branches, same packet, same output format. A0, a bare single call. A1, single call plus a meta-step that takes the question apart before anyone answers it. A1b, the agentic control: decomposition, 28 executors, nine candidate solutions, a critic on each, eight killed with named reasons, same call and token ceiling the game actually spent. No positions, no interests, no concession rules, no held conflict. And A3, the game above.
Cost. The game: 39 calls, about 2.19M subagent tokens. The single call: one call, 58k. A 38x gap. Compute-matching as pre-registered turned out to be physically unreachable. You can't make one call spend two million tokens by asking it to.
Blind judging failed. Twice. Three independent judges, four cards, no provenance, asked which process produced each. Twelve out of twelve. Then the output format was equalized across all branches to kill the tell, and it ran again. Twelve out of twelve again. Confidence on the game fell from 82-85 to 62-72, and the judges admitted they couldn't reliably tell the two single-call controls apart. They still picked out the game every time.
Say it straight: that measures detectability, not quality. The judges were identifying the process, not rating the answer. On the answers they called the controls strong and near-indistinguishable from each other.
What gave the game away, named identically by all three: the field ALREADY PRODUCED. In the game, concessions have subjects who pay. Claim withdrawn by the claimant. Power surrendered. Credit voided. In the controls the same field goes impersonal: "the fork is cancelled", "foresight as a genus is cancelled". That doesn't normalize away, because the agency of the concessions is the product.
The result that hurts: all four branches reached the same bottleneck independently. The missing external witness. The bare single call, in the equalized format, wrote that a horizon is held by an external witness rather than by a document, and that nobody returns to a long formulation without a named one. That's where the game arrived after five rounds. A1b went further on the bottleneck itself: no price for being wrong and no witness, with a procedure where the absence of a price is itself the measurement.
So most of the effect reproduces from the output format, not from the argument. And the format is partly reverse-engineered from this run: two of its fields were added after the judges named them as the game's tells.
What didn't reproduce anywhere else: a refusal issued under a rule the run itself introduced two rounds earlier, and positions withdrawing their own constructions. No control branch made either move.
#Limits
One material. A verdict on one packet, nothing wider.
The format is partly reverse-engineered from this run, so the controls got handed a ready answer to "how should I think about this". No surprise they did well.
The run record was incomplete on one point and was corrected after the fact. The eighth position's specification, its Concession Rule included, was skipped when round 5 was first written up, and restored while this document was being assembled. The rule predates the round. The record of it does not.
A designed session is a lab. Undesigned material wasn't tested at all.
All positions ran on one model. Property of the engine and property of the model are indistinguishable here, and no number of runs on one model fixes that.