A 32-contender debate tournament, covered touch-and-go. The individual rounds are specimens; the bracket is the experiment — what a format selects for when an audience votes.
Links: Debates index, Constructed ≠ Arbitrary, The Weighting Problem
32 contenders, single elimination, opening round branded “Thunder 32.” Two finalists advance to a live WW1 event. Format per round: 5-min openings → sponsor → ~15-min open floor → (sometimes) a moderated question round → closings. Moderators rotate. Advancement is decided by audience write-in ballot on a 48-hour window at wordwardebate.com — a mechanism that matters (see Cross-round findings).
Method: reviews are built from a cleaned auto-caption transcript plus real-time commentary captured while listening. That mode produced the densest discussion in the bracket (sixteen sections off one hour of tape) because reactions are caught at the moment they occur rather than reconstructed after. Its ceiling is length — it scales with runtime, not with signal — so the WW1 main-card debates (1:56 and 2:07) will need a different approach: timestamped notes on marked moments, or a selection pass before commenting.
The sampling fact behind the information channel, stated plainly (2026-08-17). > Chris: “I seem to know about 1/2 the contestants personally.” Named so far: Glenn Lawrence (military service), Tareyak (feminism rerun), one of the two in the hate-speech round, Kewl Vic (party loyalty), plus Charsky and Bunn from the AI round. This is worth stating as a finding rather than a footnote, because it cuts two ways. It is evidence for finding 2 — a bracket in which one observer personally knows half the field is, by construction, seeded from a single community, which is exactly the condition under which affiliation rather than reach decides a 30–105-ballot round. And it is a standing caveat on the predictions ledger: several calls are made with information no transcript contains, so a correct prediction is not automatically evidence that the stated reasoning was right.
Coverage policy: deliberately not exhaustive. Rounds get picked up when something structurally interesting happens, not to complete the bracket. Chris knows and has met a number of the contenders, so the series is partly followed as an audience member — which also means there is an information channel here beyond the transcripts.
The reason the bracket is worth tracking as a unit rather than as a pile of reviews. These only appear across rounds:
Moderator interference — the third data point breaks the pattern, and the real variable is attribution. After two rounds the read was “format property.” A targeted check of the Boomers round (moderator: Pisco) does not support that. Three moderators, three behaviours:
| Moderator | Round | Behaviour | Verdict |
|---|---|---|---|
| Molyneux | Masculinity | Argues a side at length; supplies the aff’s best case himself after closings | Interference |
| Kyla Turner | Feminism | Restates a debater’s losing answer as the far stronger one he never gave, attributed to him | Interference — and the worse kind, because it launders |
| Pisco | Boomers | Forces the aff to specify culpability vs. mere causality; isolates where the two already agree; offers a counter-argument explicitly flagged (“here’s a response that I’m kind of channeling”) | Not interference — arguably the correct technique |
So the earlier “questions vs. content” line was too crude. A moderator may supply content; what matters is whether it is labelled as the moderator’s own construction and put to both sides, or laundered into a debater’s mouth as something they said. Pisco supplies a demographic-collapse counter and names it as his; Kyla supplies a weight-equivalence argument and calls it Miles’s. That is the whole difference. Chris’s series-level read (“interfering too much so far”) survives as a claim about particular moderators, not the format.
Confirmed on the full tape (Boomers review §5) — the earlier caveat is retired. Pisco’s technique holds across the whole round, and the round’s structural spine (culpability vs. mere causality) exists only because he installed it. He also asks the best moderator question in the bracket so far: “What would it take for you to be convinced?” — the falsifiability question, put neutrally. The review logs one boundary case: he blocks a debater’s etymological move mid-round, which is a moderator adjudicating substance. It clears the rule only because he states it as his own rather than laundering it. A stricter rule would fail it — which is a useful stress test of where the line actually sits.
The etymology pair — the same intervention from two moderators, on opposite sides of the rule (2026-08-17). The stress test above now has its matching half. In the hate-speech round, the Aff uses discriminate in its descriptive sense — telling things apart — which is a real sense of the word and the one his equality argument depends on. Kyla blocks it, requiring him to clarify “what you mean which is not just to tell the difference between things” — a clause that does not ask which sense he intends but rules one out, and the sense ruled out is the one his case needs. The Neg’s definition survives and the segment proceeds on it.
Set against Pisco: identical move, opposite side of the line. Pisco blocked an etymological move and owned it as his own construction, which is why it cleared the attribution rule as a boundary case. Kyla’s arrives dressed as a request for clarification, so the room hears a neutral moderator asking a question rather than a moderator settling a contested term. That is a cleaner statement of where the line sits than either round alone: the rule is not “don’t supply content,” it is “don’t launder a ruling as a question.” Chris’s read of her across this round is blunter than the tape looks — “Kyla is constantly steering the narrative… she immediately cuts in and champions the left’s definition” — and it overturns the reviewer’s initial verdict that she was clear of the round-2 charge.
The scoring mechanism confounds every inference from the bracket. Time-boxed write-in ballots also measure reach, so advancement is over-determined: legibility and popularity both predict it. “This format selects for legibility over correctness” stays unfalsifiable against “it selects for audience size” until a round appears where the two point opposite ways. Kyla explicitly instructs voters against popularity — worth noting the organizers are aware of it.
Two sub-mechanisms, from Chris’s own experience as a debater rather than from the tape. “Reach” was too coarse — it was being read as raw audience size, and at least two distinct things hide inside it:
Rounds fail at different layers, and the layer is diagnostic. Arguments over a contested term have a stack — definitional → mechanism → empirical. Round 1 never escaped layer 1 (58 minutes on what masculinity is). Round 2 cleared it in ninety seconds and then failed at layer 3, running two parallel monologues against undefended metrics. Two distinct failure modes, both fatal, neither obvious from a summary.
The stack needs a floor — layer 0. The Boomers round never reaches layer 1, because the two sides never share a standard for what counts as evidence at all (one side’s causal mechanism is planetary transits, asserted at “almost always 100%”). This is diagnostically distinct from a layer-1 failure: layer 1 is two people who could adjudicate a definition and don’t; layer 0 has no adjudication procedure to reach. Neither debater is being unreasonable given their own standard, and the round is still dead on arrival. The tell is a position that cannot name its own defeater — Pisco asks Mariah directly what would change her mind and gets no condition back.
Topic-talk vs. resolution-work. The craft lesson, and it generalizes past this series: every minute is affirming the proposition, negating it, or wasted — however interesting. The test for any line of argument is whether you can state the sentence connecting it to the resolution. Round 2’s dating-app segment was the central argument in microcosm and still read as a digression, purely because neither debater tied it back. Framing is the debater’s job, not the moderator’s.
Superseded as best-in-bracket (2026-08-11). Chris’s read after nine rounds: “this was the best debate so far, even above the AI debate” — Turner vs. Bourdeau. Different virtue, though, and worth keeping both: the AI round is two well-matched debaters doing the resolution-work correctly; the fragility round has a live hinge that both sides keep fighting over once the term is pinned. One is a model of technique, the other of a real disagreement. The AI round remains the control case for this finding; the fragility round is the control case for The Load-Bearing Word.
Control case found — the AI round. Charsky and Bunn waste nothing: both weigh explicitly, Charsky’s closing opens “reasons why you should favor me winning a debate,” and both state the crux out loud (the counterfactual baseline). Which raises the confound: is the finding about the format, or simply about these two being better debaters? A format that permits both outcomes isn’t obviously selecting for either. Open on the review as seed 1.
True position, lost round. — substantially weakened by the results. The original claim: twice the side the vault read as correct lost on execution, so the series’ real finding would be about formats rather than topics. The round-1 results do not support it.
| Round | Position the vault reads as true | Outcome |
|---|---|---|
| Masculinity | Garcia | Lost, 25.8% — confirms |
| Rich-without-luck | Neg / The Aftermath (Chris’s read) | Won, 71.88% — counterexample |
| AI | Charsky’s side, on the vault’s own three-layer rule | Won, 79.05% — counterexample |
Third counterexample, and it retires the finding (2026-08-17). The military-service round was the designed test: the vault said the argument went to Brunet and the ballot would go to Lawrence on rhetoric and appeal. Brunet won 78.6 – 21.4. And the conditions favoured the wrong answer as hard as this bracket can arrange — the Aff was a USAF veteran arguing a veterans-first proposition to an audience that advertises patriot credentials, so rhetoric and audience prior both pointed at Lawrence. He lost by 57.
So the running tally is one confirmation (Garcia) against three counterexamples, and the three counterexamples are +43.8, +58.1 and +57.2. Garcia is better explained by finding 6 than by any bias against correctness. This finding is now retired rather than merely demoted; what replaces it is finding 15.
One confirmation and two counterexamples, both of them decisive margins. The honest revision: the format does not systematically punish the correct side. What round 1 actually shows is closer to the opposite — the two rounds where the vault had a clear view of who was right resolved in favour of that side, and the one loss (Garcia) is better explained by finding 6 (delivered ≠ received) than by any bias against correctness. Retained as a watch item for round 2, not as a finding.
Two ways to lose an audience — and a technique/format mismatch. “Delivered ≠ received” splits into distinct failure modes. Garcia lost his audience at the level of prose (a private vocabulary nobody could parse); Charsky loses his at the level of delivery — clear prose, announced structure, dense citation, all of it delivered fast and monotone in the manner of competitive policy debate. Those habits are optimized for a trained judge flowing arguments on paper, where dropped points score as conceded. Under an audience write-in vote the same technique inverts: density becomes noise and citations evaporate at speed. Not a deficiency — a mismatch between a tuned technique and a different scoring function. Extension (2026-08-14) — the mismatch also operates on form, not just delivery. The finding above is about delivery — Charsky’s fast, dense, citation-heavy competitive-policy manner inverting under an audience ballot. The military-service round shows the same mismatch one level up, in argumentative structure.
The correct technical answer to an inflating affirmative there is a counterplan: advocate national service — non-topical by construction, competitive because “prerequisite” is a necessary condition that a sufficient alternative destroys, and guaranteed to solve the harms because the affirmative attributed them to national service himself. Under a trained judge that is simply a negative ballot. Under an audience vote it is a liability, because the affirmative and most of the room read the negative’s agreement as concession:
Chris: “Glenn would think ‘Yeah, neg agrees with ME! so I win!’ not understanding that he is refuting the resolution… I don’t think many people would understand this tactic and it might be losing because of it.” … “structured debate lovers would enjoy this, but it alienates 90% of the audience.”
So the generalisation: the gap between winning a round and being seen to win it widens with the technical sophistication of the move. A format scored on perceived clash cannot register a win that consists of manufactured agreement. This is a real constraint on what a competent debater should do in this bracket, not merely on how fast he should talk — and it is the strongest version of finding 2’s worry, because it says the format doesn’t just fail to reward the correct response, it can actively punish it.
Untested escape: strip the machinery and keep the ballot question concrete — “neither of us is defending tonight’s resolution; should a firefighter who never enlisted be allowed to vote? He says yes, I say yes, the resolution says no.” Same structure, no jargon. Nobody in the bracket has tried it.
Second specimen, one round later — and it confirms the finding from the other direction. The military-service case was hypothetical: a technically-decisive move that wasn’t run, judged in advance as a probable vote-loser. The feminism rerun supplies the live comparison, because there the audience-facing counter was run and the tactical one wasn’t. Facing a squirrel affirmative (one instance satisfies me; you must refute every interpretation), Tareyak simply declined the burden framing and kept re-asking the general question, while the available tactical kill — pin the case to the single example in its constructive, refuse everything later as new argument — went untaken. Chris grades his own tactical counter down: “the neg’s approach has broader appeal because mine is ‘too tactical’, and people wanted to hear the Neg’s framing of the debate.”
The structural reason, now stated as a rule and promoted to The Negative’s Easy Burden: the tactical counter’s win condition is my opponent failed to discharge a burden, which requires the room to accept a burden allocation nobody announced; the framing counter’s win condition is my account of the subject is better than his, which is what the audience already thinks it is voting on. So the finding generalises past delivery and past argumentative form to win conditions themselves: a lay ballot can only register a victory whose condition it already recognises. That is the cleanest statement this bracket has produced of what an audience vote can and cannot measure.
The format punishes the disposition that makes discussion valuable. The AI round was the best conversation of the bracket — both men truth-seeking rather than playing bloodsports — and the debater who qualified himself honestly reads as unstable. Concessions that are intellectual virtues function as competitive liabilities, and a performance-scored vote cannot distinguish an honest qualification from a wobble. Chris: “I like both of them because they want discussion over blood-sports, but in this competition they need to play the game.”
Not every round has a proposition in contention — the verbal dispute. Guptill vs. Tejada is well-formed at every layer of finding 3’s stack and empty at the top: definitions get settled (twice), evidence standards are shared, both men reason competently — and they agree. Strip the contested word and the Neg (“society won’t speak hard truths and punishes those who do”) and the Aff (“society coddles, affirms delusions, rewards the unearned”) assert the same claim, which is why the Neg answers “Absolutely” when handed the Aff’s central evidence. The tell is a debater saying “it sounds like you agree with me — don’t you technically concede?”: he is right about the agreement and wrong that it’s a concession, because you cannot concede to someone asserting your own view under a different label. Diagnostic: state both sides without the contested term; if the disagreement disappears, it was verbal. This is orthogonal to the layer model rather than another rung on it — the question is there a proposition in contention? is asked before the stack is entered.
The ordering isn’t random — it’s a stable convention that systematically inverts, which is worse. Corrected 2026-08-12 after Chris pushed back on the “haphazard” reading; he was right and the checked version is sharper. Across all ten covered rounds:
| Signal | Pattern |
|---|---|
| Introduction order | The Neg is introduced first in 10 of 10. “Debater one” is always the side denying the prompt — a completely stable convention |
| Speaking order | The Aff opens in 7 of 10. The three exceptions (feminism, AI, too-sensitive) are all Kyla-moderated — but Kyla also runs Aff-first rounds, so it is inconsistency within a moderator, not a per-moderator rule |
| Closing order | ⚠ Not a single convention — see the correction below. The rule stated in Kyla-moderated rounds is whoever opens first closes first (“since Forest got the first word, Jose, you’re going to get the last word”), but Mullally runs a different rule, and runs it inconsistently |
Correction — closing order varies by moderator, and one moderator varies within himself (2026-08-17). Prompted by Chris’s suspicion that “the closings are not always the same,” checked against seven archived transcripts. His worry that speaker labels would make this unrecoverable turns out not to bind: the moderator announces the closing order out loud in every round, so the convention is readable straight from the text.
| Round | Moderator | Announced | Effective rule |
|---|---|---|---|
| Feminism (R2) | Kyla | “whoever got the beginning word, the opposite gets the final word” | opener closes first |
| AI | Kyla | “since Charsky got the first word, you get the last word” | opener closes first |
| Too sensitive | Kyla | “since Forest got the first word, Jose… the last word” | opener closes first |
| Algorithm | Kyla | “Ryan, you got the first word. Silvio will get the last word” | opener closes first |
| Hate speech | Kyla | “since Octavius, you got the first word. Chase, you get the last word” | opener closes first |
| Western therapy | Mullally | “Spencer as the affirmative, you have the last word” | Aff closes last |
| Feminism rerun | Mullally | “David started the debate. You go first and then David… the last word” | Aff closes last |
| Party loyalty | Mullally | “Joe, because you have the affirmative, you’re going to get a closing first” | Aff closes FIRST |
So: Kyla is perfectly consistent across five rounds on a positional rule (opener → first closing). Mullally uses a side-based rule instead — and applies it in opposite directions, giving the affirmative the last word twice and the first closing once, each time announced as though it followed from being the affirmative. Chris’s read is confirmed and sharpened: “most often Aff closes first, but a few times Neg closed first” is right, and the reason is that two incompatible conventions are in play and neither is the format’s.
Why it matters rather than being trivia: the last word is a real advantage in a round scored by audience recall, and formal systems allocate it deliberately — the side carrying the burden gets the bookends (The Negative’s Easy Burden § the institutional admission). Here the advantage is assigned by which moderator drew the round, and in Mullally’s case by which rule he reached for that night. That is a per-round coin-flip on a structural edge, in a bracket where rounds turn on a dozen ballots.
Second correction — the opening-order count looks wrong and needs a verification pass. This finding records the Aff opening in 7 of 10 with three exceptions (feminism R2, AI, too-sensitive). The transcripts support feminism R2 (Ruelas, listed first, gets the beginning word) and too sensitive (Guptill opens) — but not AI: Charsky takes the first word, and Charsky is the affirmative. If that holds the count is 8 of 10, not 7. Flagged rather than silently amended, because the Aff/Neg assignment for that round is inferred from its review rather than from the intro. Also note this refutes the stronger version Chris proposed — “we are 100% that Aff opens” is not right; two genuine exceptions survive the check.
So the intro order inverts the speaking order in the normal case — the Neg is named first and the Aff talks first. Two signals, pointing opposite ways, neither announced. A contender who reasons “the Aff opens, and I’m not opening, so I must be the Neg” is right 70% of the time and catastrophically wrong the rest — which is precisely what happened to Jose Tejada, who drew one of the three Neg-first rounds and was introduced second, so both cues misled him simultaneously. He argued his opponent’s case for a fifth of the round.
Cheap fix for the organizers: open with the same side every time, or say out loud which side is which before the clock starts. The moderators already state the position in the bios; the failure is that the ordering contradicts it. Debate convention says the side carrying the burden opens; here the order is effectively arbitrary and is never announced as arbitrary. A contender who reads it as a signal of his own assignment will be wrong about a third of the time, which is one plausible contributor to Jose Tejada spending a fifth of his round arguing his opponent’s case. Caveat: the assignment was stated explicitly in the intro, so this is a contributing defect, not an excuse. Cheap fix available to the organizers: fix the order, or say out loud that it’s random.
The audience prior is the specific pool’s, not the general public’s — and this pool is agency-affirming. The sharpest lesson from the one wrong prediction. Chris forecast the Aff in the wealth round on the reasoning that “people who are not rich want to believe the topic” and “it seems easier to blame something else than to solve the problem.” That is a sound intuition about the general population and it inverted against this audience: the Aff took 28%.
Look at the two largest margins in the bracket and what they have in common:
| Round | Winner’s position | Margin |
|---|---|---|
| Too sensitive | Tejada — people are coddled; struggle and merit are real | +89.1 |
| Rich-without-luck | The Aftermath — you can build wealth by your own action | +43.8 |
Both are agency-affirming: outcomes follow from what you do. Both won enormous. A bracket whose contender bios advertise “the working man’s resume,” “straight white Christian father patriot,” and Infowars credentials is not drawing a general-population sample — it draws people for whom the personal-responsibility answer is the flattering one.
So finding 2’s motivated-reasoning sub-mechanism needs a correction rather than a deletion: it is real, and its direction must be read off the actual voting pool. Applying a general-population prior to a self-selected audience will get the sign wrong, which is exactly what happened here. Caveat: two specimens, and the agency reading is cleanest for those two — Owlish’s win is better explained by delivery, and Charsky’s by argument quality. Offered as the leading hypothesis for round 2 to test, not as an established finding.
The neutral-dilution hypothesis — why later rounds should measure performance better. Chris’s structural prediction, registered before round 2 airs:
Chris: “I figured the early rounds were mostly popularity contests, but as we move up, they will attract larger audiences so a larger streamer doesn’t win by default against better performance.”
The mechanism is dilution. A round-1 ballot is decided largely by voters who arrived as partisans of one contender, so affiliation (finding 2) dominates. As the bracket advances, two things change: both contenders have already cleared a round, so neither has a raw-reach edge by default; and the series accumulates an audience that follows the tournament rather than any one debater. The neutral fraction rises, and with it the weight of performance.
This is a genuine, falsifiable prediction with a clean test: margins should compress and correlate less with follower counts as rounds progress. If round 2 produces another 89-point blowout for the better-known name, the hypothesis is in trouble.
The reach analysis — attempted 2026-08-11, and the attempt answered the question a different way.
Chris: “we would have to run some sort of reach analysis for each contestant to see if this was just a popularity contest.”
The plan was to collect each contender’s subscriber/follower counts and regress vote share on relative reach. It cannot be run as designed, for two reasons that are themselves the finding:
⚠ The search failure has a cause, identified by Chris 2026-08-12: “part of what makes the reach research hard is they are mostly using real names.” Contenders are billed as Josh Smith, Nic Turner, Silvio Cruz — names that return nothing, while their actual channels run under handles. Chris supplies the worked example: Josh Smith’s channel is “JS Urban Adventures,” with a decent following, plus the streaming show The Dreaded Conservative (200+ episodes). Neither is reachable from the name in the bio. So the reach data mostly exists and is simply not addressable by the identifier the bracket publishes — which downgrades the conclusion below from “there is nothing to measure” to “it cannot be measured this way.” A workable method would go handle-first: pull each contender’s channel from the debate description, their own socials, or by asking, rather than searching the legal name.
Confirmed 2026-08-17, and the hate-speech round supplies the identity mapping in both directions on one tape — which is the strongest evidence yet that the data exists and only the identifier is wrong. Kyla bills Chase McPherson by legal name and then reads out the handle (@_vigilante_tv, YouTube and Twitch); she bills Octavius Thunder by handle and then gives the legal name (Eli Halpern). Chris supplied the McPherson mapping independently before the tape was pulled. So the bracket is not consistently using real names or handles — it uses whichever the contender prefers, and the intros carry the cross-reference. The practical upshot: the reach data is recoverable by watching the first ninety seconds of each video, which is a far cheaper method than any search, and the bar to running finding 12’s analysis is now labour rather than availability.
So the correct reformulation: round 1 is not a popularity contest in the streamer sense — it is a turnout contest among personal networks. That is a meaningfully different claim, and it explains the 52–3 result better than reach ever could.
One case where the profile gap is legible, and it cuts the right way. In the AI round, the higher-profile figure is clearly Bunn — a credentialed academic (PhD Wisconsin–Madison, Assistant Professor at Covenant College, published in Comment and Providence, an ISI Top-20-Under-30 award, an established X presence). Charsky has no findable public footprint at all. Bunn lost 83–22 on the largest pool in the bracket. Whatever decided that round, it was not the participants’ relative standing — which is the strongest available evidence that the AI result is a genuine quality signal.
⚠ The decisive fact about this bracket: rounds are decided by 30–105 ballots. Discovered 2026-08-11 while attempting finding 12’s reach analysis, and it reframes every inference above.
Two independent derivations converge. (a) Reverse-engineering the reported percentages. Several results are published to two decimals, which constrains the denominator tightly — 71.88/28.12 admits only n=32 (23–9) or its multiples; 94.55/5.45 only n=55 (52–3); 79.05/20.95 only n=105 (83–22). (b) View counts on the source videos, pulled from YouTube: 163–1,136 per debate, against a channel with 6,360 subscribers. Each inferred n falls at 15–25% of its video’s views — ordinary turnout. The two methods were derived independently and agree.
| Round | Views | Reported | Inferred ballots |
|---|---|---|---|
| Masculinity | 703 | 74.2 / 25.8 | 23 – 8 (n=31) |
| Boomers | 1,136 | 61.7 / 38.3 | 29 – 18 (n=47) |
| Feminism | 591 | 60.5 / 39.5 | 23 – 15 (n=38) |
| AI | 810 | 79.05 / 20.95 | 83 – 22 (n=105) |
| Rich-without-luck | 210 | 71.88 / 28.12 | 23 – 9 (n=32) |
| Too sensitive | 276 | 94.55 / 5.45 | 52 – 3 (n=55) |
Caveat: multiples of each n fit equally well (n=64, 96… for the wealth round), so these are lower bounds; the view-count cross-check is what makes the smallest solutions most likely. Percentages could also be rounded in a way this reconstruction doesn’t model.
What this does to the findings:
Audience size is small and shows no clear trend within round 1. Views ran 703 → 1,136 → 591 → 810 → 210 → 276 → 239 → 163 across Aug 3–10. The release interval is now pinned: the channel has been posting on a 48-hour cycle (see the production note below), so each successive video in that series had two fewer days of exposure than the one before it. It is tempting to read a collapse, but later videos have had far less time to accumulate (the Aug 10 debate had one day; the Aug 3 debate had eight), so the sequence cannot establish a decline. What it does establish is the scale: no round-1 debate has cleared 1,200 views, against a 6,360-subscriber channel.
Subscriber datapoint, 2026-08-17: 6,410 — up ~50 over roughly two weeks of the bracket running. Worth recording because finding 11’s neutral-dilution mechanism requires the tournament to accumulate an audience, and a ~0.8% channel growth across the whole of round 1 is not the order-of-magnitude change finding 13 says would be needed for performance to outweigh a mobilised bloc. Early evidence against the hypothesis, from the one number that tracks it directly.
The ledger’s own calibration is the finding: the argument reads hold up, the ballot reads don’t — and they fail in one direction. With ten blind calls now resolved, the pattern is legible and it is not flattering to the cynical prior this hub started with.
| Resolved ballot call | Outcome |
|---|---|
| Owlish advances on delivery | ✅ |
| Charsky wins round 4 | ✅ |
| Tejada advances | ✅ |
| Calvin wins round 5 | ❌ |
| Hamm advances | ❌ |
| Turner (Aff) wins | ❌ |
| Lawrence (Aff) wins on rhetoric | ❌ |
| Summerhays/Gilkison — abstained | ✅ (the abstention was right; the vote diverged 43 points from the substance read) |
⚠ REWRITTEN 2026-08-18 — this finding was refuted by the next two results, one day after it was written. The claim above was that the misses are one-directional: that every wrong ballot call but Hamm came from expecting rhetoric, sympathy or affiliation to beat argument, and therefore that the fix was to stop discounting argument quality. Then the last two cards posted:
| Round | Chris’s read | Result |
|---|---|---|
| Party loyalty | Joe — better arguments, presented well | Kewl Vic +62.0 |
| Hate speech | McPherson — “had a plan… experience dealing with these types of debates” | Thunder +25.0, closing on “satanic pedophilic overlords” |
Both misses run the opposite way. Here Chris predicted the better arguer and the worse one won — in one case by 62 points, in the other over a closing statement about global enslavement. So the asymmetry claimed above does not exist; with nine calls resolved the misses run in both directions, and “predict the ballot from the argument” would have gone 0-for-2 on its first outing.
The lesson is about the vault’s method, not about the audience. Finding 15 was fitted to seven data points and broke on the eighth and ninth. That is precisely the error finding 13 warns about — small-n noise read as signal — committed here on the vault’s own ledger rather than on the debaters’. A pattern over seven binary outcomes is roughly what chance produces anyway; it should never have been promoted to a “posture going into round 2.” Standing correction: no predictor gets promoted off fewer than ~10 resolved calls, and any predictor that would have been contradicted by the two most recent results is not yet a predictor.
What actually survives, stated conservatively: the analytical reads still hold up — argument-quality verdicts, concession-tracking, layer diagnoses, and load-bearing-word calls have all survived contact with results. What remains unmodelled is the mapping from those verdicts to ballots, and three of nine is not distinguishable from a coin. The honest posture into round 2 is therefore agnostic rather than corrected: run the autopsy over all fourteen at once, and register a model only if it beats the base rate across the whole set rather than across a flattering slice of it.
The bloodsports penalty — a contender’s own diagnosis, and the most promising predictor on the table (2026-08-18).
Chase McPherson, to Chris after his loss: the format penalizes contestants who engage in “bloodsports” tactics.
Sourcing first: this is a defeated contender explaining his own defeat, which is exactly the kind of claim that needs discounting — the self-serving version attributes a loss to a format flaw rather than to performance. It is logged anyway because it is independently supported by material already in the vault, some of it recorded before this round aired, and because McPherson is well-placed to know: he is a practised debater and Chris’s read of him was positive (“had a plan… experience dealing with these types of debates”), so the thesis is not being used to paper over a bad performance.
The convergent evidence, none of it collected with this hypothesis in mind:
But the crude version is false, and the counterexamples locate the real line. Technical moves per se are not punished: Cruz won on a severance and Bourdeau won on a severance, both textbook technical manoeuvres, and Charsky won by 58 with fast, dense, citation-heavy competitive-policy delivery that finding 6 predicted should hurt him. So “the audience punishes technique” is refuted.
Sharpened hypothesis — the audience punishes making the opponent the subject. What separates Joe and McPherson from Cruz, Bourdeau and Charsky is not technicality but target: the winners argued about the resolution, the losers argued about the man across from them. This slots directly into finding 4’s ledger as a third category — every minute is affirming, negating, or wasted, and opponent-talk is wasted and costly, because a room that came to hear a question adjudicated reads it as evading the question. It also explains the counterexample pattern that “bloodsports tactics” cannot.
Standing relative to finding 17: finding 16 is now the junior hypothesis. Chris’s subject-prior model explains 7 of 11 codable rounds directly and the twelfth by affiliation, without appealing to posture at all — and it explains both rounds finding 16 was built on: hate speech fits the prior (a free-speech pool backing the free-speech side), and party loyalty fits affiliation (Vic is the larger channel). So opponent-talk may be doing no independent work. It survives only for the residual — the four rounds where the prior lost — and should be tested there rather than everywhere.
Falsifiable, and cheap to test in the autopsy: code each round for opponent-directed minutes and see whether it predicts the loser across all fourteen. If a debater who spent visibly more time on his opponent than on the resolution has lost every time, that is the first predictor in this hub with a real mechanism behind it. Registered before the coding is done, so it can fail honestly.
Chris’s subject-prior model — it outperforms everything else this hub has tried, and it is now the front-runner (2026-08-18).
Chris: “my thoughts were just that it was a combination of popularity (Kewl Vic) and voting for the belief they hold anyway… good rhetoric is a factor as it changes the margins — McPherson was 4th closest when most would disagree with him — but I think people are voting on the subject, and only the debaters if they know them.”
A three-part model, and each part is separately checkable: (a) the ballot mostly tracks the resolution, not the performance; (b) affiliation intrudes only where the voter knows a contender; (c) rhetoric moves the margin, not the winner.
Coded against all twelve decided rounds. “Prior” = the position this specific pool is expected to hold, read off the bracket’s own audience signature (agency-affirming, free-speech, anti-establishment, patriot-coded — finding 10):
| Round | Winner (side) | Pool’s prior favours | Fit |
|---|---|---|---|
| Feminism helping | Ruelas (Neg — not helping) | Neg | ✅ |
| Rich without luck | Aftermath (Neg — luck isn’t required) | Neg | ✅ |
| Too sensitive | Tejada (Aff — we are) | Aff | ✅ |
| Algorithm fair | Cruz (Neg — not fair) | Neg | ✅ |
| Therapy for men | Summerhays (Aff — doesn’t work) | Aff | ✅ |
| Hate speech | Thunder (Aff — protect it) | Aff | ✅ |
| Boomers responsible | Rex Jones (Aff) | Aff (weak — judgment call) | ✅ |
| Masculinity in crisis | Owlish (Con — not in crisis) | Pro | ❌ |
| AI making us dumber | Charsky (Neg — it isn’t) | Aff | ❌ |
| Fragile culture | Bourdeau (Neg — culture isn’t the cause) | Aff | ❌ |
| Military service to vote | Brunet (Neg — no) | Aff | ❌ |
| Party loyalty illogical | Kewl Vic (Neg) | unclear — affiliation | ➖ |
Seven of eleven codable rounds, with the twelfth explained by clause (b). Against the vault’s own 3-of-9 ballot record, that is a decisive improvement — and it was reached without watching a single tape, which is itself the point.
And the four failures are not random: they are the four rounds where the vault recorded the largest execution gap. Owlish’s legibility against Garcia’s private vocabulary; Charsky, the best debater in the bracket; Bourdeau extracting three concessions; Brunet naming the inflation in thirty seconds. So the natural completion is a two-factor model:
Default to the audience’s prior on the resolution. The prior is overridden only by a large execution gap — and rhetoric otherwise moves the margin, not the winner.
That fits 11 or 12 of 12, and it subsumes rather than discards the existing findings: finding 10 supplies the prior’s direction, finding 2 supplies clause (b), and finding 6 supplies what counts as an execution gap.
Clause (c) has direct support, and Chris identified the specimen. McPherson argued a speech-restriction case, as a self-described center-left streamer, to a free-speech pool — the prior was against him about as hard as this bracket allows — and he held it to 25.0 points, the 4th-closest margin of twelve, against a median of 45.6. He lost the round and won the margin, which is exactly what clause (c) predicts.
A structural detail that fell out of the ranking and supports the two-factor split: the margins are bimodal. Four rounds land at 12.4–25.0, then there is a gap with nothing in it, and eight land at 42.8–89.1. A single noisy predictor would smear across the range; two regimes look like prior-aligned blowouts versus rounds where something contested the prior. Worth testing directly — the four close rounds should be the ones where prior and execution point in opposite directions.
⚠ The discipline this finding must obey, given finding 15 died yesterday of exactly this error. The prior codings above were assigned after the results were known, which is precisely how a model earns a fit it hasn’t earned. This is therefore a hypothesis, not a finding, and its test is registered in advance:
Chris falsified his own model within minutes of proposing it, and the objection is better than the model (2026-08-18).
Chris: “the Garcia debate already falsifies me :) It seems the audience leans right, and this would give Garcia the win, but here it seems rhetoric played the bigger part at least in this one. Of course this is a multivariate problem that the variables change each round (different voter block each round).”
The self-falsification is correct and was already coded ❌ above. But the second clause is the deeper point and it changes what kind of object the prior is. Finding 17 as written treats “the pool’s prior” as a fixed parameter of the bracket. It isn’t. At 30–105 ballots drawn largely from the contenders’ own networks, each round convenes a different electorate — so the operative quantity is that round’s pool’s prior, which moves with whoever mobilised. A resolution’s prior and a round’s prior are not the same number.
That is not a refutation so much as a specification of the error term, and it does real work:
Chris’s own framing of why it beat the vault’s model is worth keeping as the summary: the vault has been analysing debaters, and the audience has been voting on topics.
Chris’s standing complaint, and the reason this section exists: “a small pet peeve with the WW format — since it is not real time, they can’t keep the bracket updated, so it is hard to see how things are going.”
Exactly right, and it is structural rather than sloppy. A live event resolves on the night; this one resolves 48 hours after each release, on a staggered schedule, with results published as image cards on a page that carries no bracket. So at any moment the tournament is in a superposition — some rounds decided, some inside their window, some unaired — and nobody outside the organizers can see the standings. The rounds are all we get; the tournament is invisible. Since this hub exists to track the bracket as a unit, maintaining the state here is the fix.
Standing as of 2026-08-18 — 12 of 14 aired rounds decided. (Chris reported all results as posted; in fact two are still missing — see the note under the table.)
| # | Round | Result | Margin | Advances |
|---|---|---|---|---|
| 1 | Masculinity | Owlish 74.2 – 25.8 Garcia | +48.4 | Owlish |
| 2 | Feminism | Ruelas 60.5 – 39.5 Anton | +21.0 | Ruelas |
| 3 | Boomers | Rex Jones 61.7 – 38.3 Meyers | +23.4 | Rex Jones |
| 4 | AI | Charsky 79.05 – 20.95 Bunn | +58.1 | Charsky |
| 5 | Rich without luck | The Aftermath 71.88 – 28.12 Constable | +43.8 | The Aftermath |
| 6 | Too sensitive | Tejada 94.55 – 5.45 Guptill | +89.1 | Tejada |
| 7 | Algorithm | Cruz 56.2 – 43.8 Hamm | +12.4 | Cruz |
| 8 | Therapy for men | Summerhays 71.4 – 28.6 Gilkison | +42.8 | Summerhays |
| 9 | Fragile culture | Bourdeau 73.7 – 26.3 Turner | +47.4 | Bourdeau |
| 10 | Military service | Brunet 78.6 – 21.4 Lawrence | +57.2 | Brunet |
| 11 | Therapy culture | Smith vs. Ouedrago | — | ⏳ closed, unposted |
| 12 | Feminism rerun | David S. vs. Tareyak | — | ⏳ closed, unposted |
| 13 | Hate speech | Thunder 62.5 – 37.5 McPherson | +25.0 | Octavius Thunder |
| 14 | Party loyalty | Kewl Vic 81 – 19 Kung Fu Joe | +62.0 | Kewl Vic |
| — | (two matchups) | never listed | — | not aired |
Two rounds remain unposted and both are overdue: therapy culture (Smith vs. Ouedrago) and the feminism rerun (David S. vs. Tareyak), each closed well before the two that just published. So the pipeline is not merely lagging — it is publishing out of order, which is a fourth production irregularity and rules out “they post as windows close.”
What the margins say — this format produces blowouts. Across twelve decided rounds the mean margin is 44.2 points and the median 45.6. Eight of twelve were decided by 40+, and exactly one has been competitive (Cruz/Hamm, +12.4). Read against finding 13 — rounds turning on 30–105 ballots — a 45-point median means a typical round is decided by a dozen or so people breaking one way. This is the baseline finding 11 predicts should compress as the bracket advances: if round 2’s margins don’t shrink, neutral dilution is in trouble.
Three rounds are closed with no card posted, including hate speech, whose window has clearly expired. Given the two never-listed matchups and the round-2 filming already underway, the publication pipeline is running behind the tournament itself — a third production irregularity, consistent with reading the format as loosely administered.
Assembled 2026-08-18. The organizers publish no bracket (defect 4), so this is reconstructed from result cards, the channel, and the site.
12 of 16 round-2 slots are known:
| Advancing | Beat | On |
|---|---|---|
| Owlish | Garcia | delivery over substance |
| Martae Ruelas | Anton | +21.0 |
| Rex Jones | Meyers | +23.4 |
| Luke Charsky | Bunn | the bracket’s best-argued round |
| The Aftermath (Petro) | Constable | +43.8 |
| Jose Tejada | Guptill | +89.1, the largest margin |
| Silvio Cruz | Hamm | +12.4, the only close round |
| Spencer Summerhays | Gilkison | +42.8 |
| Rowan Bourdeau | Turner | +47.4 |
| Chris Brunet | Lawrence | +57.2 |
| Octavius Thunder | McPherson | +25.0 |
| Kewl Vic | Kung Fu Joe | +62.0 |
4 slots unresolved: two await result cards (Smith vs. Ouedrago; David S. vs. Tareyak) and two come from the matchups that were never listed or aired.
One name recovered for the missing four. The channel hosts contender audition videos, and cross-referencing them against the known field leaves exactly one unmatched: Lee Horseradish — an audition posted, no matchup, no result. That is a strong candidate for one of the two unaired rounds. (The other three remain unknown; the site has no roster page — /wwd-contender-series-1 returns Page Not Found — and /vote-now-1 is now empty, so no ballot is currently open anywhere.)
No round-2 content exists publicly yet: no matchup videos, no announced pairings, no dates. Contenders told Chris filming is underway, so the tape exists and the publication lag is the constraint.
Compact view of the ledger; the ledger rows carry the reasoning, this carries the record.
| Round | Chris’s call | Result | Verdict | Note |
|---|---|---|---|---|
| Masculinity | Owlish, on delivery not substance | Owlish +48.4 | ✅ | Called the mechanism, not just the winner |
| Masculinity | Garcia keeps losing on legibility | 25.8% | ✅ | Open for future appearances |
| AI | Charsky | Charsky +58.1 | ✅ | The load-bearing hit: argument quality beat legibility and audience prior at once |
| Too sensitive | Tejada | Tejada +89.1 | ✅ | Called before the tape’s second half was credited |
| Rich without luck | Calvin (Aff) | Aftermath +43.8 | ❌ | Prior was general-population, pool is agency-affirming → finding 10 |
| Algorithm | Hamm — “might be close” | Cruz +12.4 | ❌ | Wrong winner, exactly right on shape — the only competitive round in the bracket |
| Therapy for men | declined to call | Summerhays +42.8 | ✅ | Abstention vindicated: substance read and ballot diverged 43 points |
| Fragile culture | Turner (Aff) | Bourdeau +47.4 | ❌ | Pre-registered meaning holds: the concessions were invisible |
| Military service | Lawrence on rhetoric; argument to Brunet | Brunet +57.2 | ❌ ballot / ✅ argument | The designed test of rhetoric-beats-argument. It failed by 57 |
| Meta | Bracket selects legibility over correctness | — | ❌ refuted | The bracket’s most valuable single result |
| Hate speech | McPherson (Neg) | Thunder +25.0 | ❌ | Called against Chris’s own position; the better debater lost to a closing about “satanic pedophilic overlords” |
| Party loyalty | Kung Fu Joe (Aff) | Kewl Vic +62.0 | ❌ | Called against a personal acquaintance — and the acquaintance won by 62 |
| Therapy culture | Smith, on reach alone | ⏳ | — | The one pure affiliation call; still unposted |
| Feminism rerun | Tareyak (Neg) | ⏳ | — | Still unposted |
Record: 3 of 9 ballot forecasts correct, one correct abstention, one meta-claim refuted. The one-directional pattern claimed here yesterday did not survive the next two results — see finding 15, rewritten.
Open bookkeeping: an Aff/Neg tally would be worth having — does either side enjoy presumption with this audience? — but three of the ten side assignments need verifying against their tapes before the count means anything, so it is deliberately not asserted here.
⚠ Retrieval note (2026-08-17): results are published as images, so they cannot be fetched programmatically. A pass over /voting-results returns navigation, a copyright line, and lazy-loaded image placeholders; the only machine-readable result on the whole page is one image’s alt text (“Rex Jones wins Round 1 of Word War Debate Contender Series”). No percentages appear in the markup at all. So every tally in the table below was — and must be — read off a rendered page by a human or a browser session, and the finding 13 reconstruction depends on those two-decimal figures being transcribed accurately. Attempting the fetch again without a rendering browser is wasted effort.
Status check the same day: /vote-now-1 shows only Kung Fu Joe vs. Kewl Vic open among contender rounds (plus a non-bracket “Candace vs. Andrew” event), so the military-service, feminism-rerun and hate-speech ballots have closed and their results exist but are not yet transcribed here. It also independently corroborates the production note above: no ballot has opened for either unlisted matchup, so neither has aired.
Source: wordwardebate.com/voting-results. Six of the seven covered rounds have posted; Cruz vs. Hamm and Summerhays vs. Gilkison are still inside their 48-hour windows, so the blind protocol continues for those two.
| Round | Result | Margin |
|---|---|---|
| Masculinity — Garcia vs. Owlish | Owlish 74.2% – 25.8% | +48.4 |
| Feminism — Ruelas vs. Anton | Ruelas 60.5% – 39.5% | +21.0 |
| Boomers — Rex Jones vs. Meyers | Rex 61.7% – 38.3% | +23.4 |
| AI — Charsky vs. Bunn | Charsky 79.05% – 20.95% | +58.1 |
| Rich-without-luck — Constable vs. The Aftermath | Aftermath 71.88% – 28.12% | +43.8 |
| Too sensitive — Guptill vs. Tejada | Tejada 94.55% – 5.45% | +89.1 |
| Algorithm — Cruz vs. Hamm | Cruz 56.2% – 43.8% | +12.4 — the closest in the bracket |
| Therapy for men — Summerhays vs. Gilkison | Summerhays 71.4% – 28.6% | +42.8 |
| Fragile culture — Turner vs. Bourdeau | Bourdeau 73.7% – 26.3% | +47.4 |
| Military service — Lawrence vs. Brunet | Brunet 78.6% – 21.4% | +57.2 |
Retrieved 2026-08-17. Ten result cards are archived at raw/debates/wordwar-results-2026-08-17/ — the site publishes results only as images, so those files are the primary source. Not yet posted: therapy culture, the feminism rerun, hate speech, and party loyalty (the last still inside its window).
⚠ Note the precision drop. The two newest results are reported to one decimal, not two. That materially weakens the finding 13 reconstruction, because a single decimal admits far more denominators: 78.6/21.4 fits any multiple of 14 (smallest 11–3), and 73.7/26.3 any multiple of 19 (smallest 14–5). Against the military-service round’s 224 views, the 15–25% turnout band observed elsewhere would favour n≈42 (18.8%) over n=14 (6.3%) — but that is an inference from the band, not from the figures, and is flagged as such.
(Turner vs. Bourdeau still inside its window as of 2026-08-12.)
| Prediction | Made | Status |
|---|---|---|
| Owlish advances on delivery, not substance | R1 §7 | ✅ CORRECT — 74.2%, the second-largest margin on the board |
| Garcia keeps losing on legibility wherever he appears | R1 OQ4 | ✅ so far — 25.8%. Stays open for future appearances |
| Bracket selects for legibility over correctness | R1 OQ4 | ❌ REFUTED — and this is the bracket’s most valuable single result. The AI round is exactly the case finding 2 said was needed: legibility (finding 6) and audience priors both pointed at Bunn, and only argument quality pointed at Charsky. Charsky took 79.05%. Argument quality beat legibility and sympathy simultaneously and decisively |
| Round 2’s winner “should go far” | R2 §6 | open — Ruelas took R2 at 60.5%; resolves in round 2 |
| Charsky wins round 4 | R4 §12, §16 | ✅ CORRECT — 79.05%. The test resolved the informative way. Three predictors, two directions: argument quality → Charsky; delivery/legibility (finding 6) → Bunn; audience priors (AI skepticism is the popular position) → Bunn. Charsky won by 58 points, so argument quality beat both legibility and sympathy at once. The single most load-bearing result in the bracket |
| Kung Fu Joe (Aff) wins, Joe/Vic | R14 discussion | ❌ WRONG by 62 points — Kewl Vic took it 81–19, the second-largest margin in the bracket, against a call giving the Aff both better arguments and better presentation. Its value: the round holds argument quality nearly constant (the two agreed), leaving posture as the visible variable — Joe’s closing is about his opponent, Vic’s about the question. Best specimen for finding 16, with the confound that Vic is both the larger channel and the acquaintance. Original registration: registered blind. Chris: “Aff has the better arguments, and he presented well, I think he wins.” Convergent (better debater and predicted winner), and made with a personal connection to the opponent — Chris knows Vic — which is the first call in the ledger where the affiliation channel runs against the prediction rather than with it |
| McPherson (Neg) wins, McPherson/Thunder | R13 discussion | ❌ WRONG — Thunder took it 62.5–37.5. The better debater lost to an opponent closing on “satanic pedophilic overlords” and “global enslavement” — and lost while arguing the side Chris disagrees with, so Chris and the room reached the same winner-side by opposite routes. McPherson’s own post-mortem became finding 16. Original registration: registered blind, and a call against Chris’s own position on the resolution, which is what makes it worth logging. He sides with the Aff on the merits and still calls the round for the Neg: “the neg performed better here as he had a plan and he has experience dealing with these types of debates… Despite all of the ‘bad’ arguments, I think this goes to the Neg.” So this is the second clean instance of the Garcia pattern — the side the vault reads as correct losing on execution — and it refines the demoted finding 5 in a useful direction: not the format punishes correctness, but a correct position with no plan loses to an incorrect position with one |
| Tareyak (Neg) wins, David/Tareyak | R12 discussion | open — registered blind, and notable for being convergent where the previous call was split. Chris: “the Neg did well to steer the debate away from the very narrow framing… I think the neg pulls this as he was more straight-forward with his arguments and was more confident.. the Aff’s closing almost sounded like a concession speech.” The Aff himself models the pool correctly on air (“I do know sort of how the audience probably skews here”) and then trashes the debate in his exit interview, drawing the moderator’s “we’ll see if the strategy of trashing the debate quality is rewarded by the audience” — so a David win would be the bracket’s first case of a procedural/burden-play affirmative beating a substantive negative before a lay ballot, which is the direct test of finding 6’s extension |
| Lawrence (Aff) likely wins, Lawrence/Brunet — but the argument goes to Brunet | R11 discussion | ❌ WRONG — and it is the most important result in the bracket, because it was the designed test and it failed cleanly. Brunet took it 78.6 – 21.4. The call was the ledger’s first explicit split: argument to Brunet, ballot to Lawrence on rhetoric. The side named as having the better argument won by 57 points — the second-largest margin on the board. So the pattern the demoted finding 5 was built on does not merely lack support; the one round staged to test it produced the opposite, decisively. Note what makes it strong evidence: the Aff was a USAF veteran arguing a veterans-first proposition to an audience that skews patriotic, i.e. audience prior and rhetoric both pointed at Lawrence, and he still lost by 57. Together with Charsky (+58) this is now two decisive rounds where argument quality beat rhetoric and sympathy simultaneously. Original registration below. Chris: “the argument goes to Chris [Brunet], but I don’t think he was aggressive enough… Glenn is likely to take this based on appeal and rhetoric alone. He was the better ‘debater’ even though he had the worse argument.” This is the exact pattern the demoted finding 5 was built on, now stated as a prediction rather than noticed afterwards — so it is the cleanest available retest of whether the format rewards rhetoric over argument |
| Smith (Aff) wins, Smith/Ouedrago | R10 discussion | open — registered blind, and explicitly a reach call rather than a merits call. Chris: “I think JS takes this… JS has a channel called ‘JS Urban Adventures’, and has a decent sized following. I think this will carry him.” This is the first prediction in the ledger made on the affiliation channel alone, and it is testable against the merits read: the Aff conceded his own causal claim and the Neg — in his first formal debate — had a defensible case. A Smith win is consistent with finding 2; an Ouedrago win would be the bracket’s first clear instance of an unknown beating a following |
| Turner (Aff) wins, Turner/Bourdeau | R9 discussion | ❌ WRONG — Bourdeau took it 73.7 – 26.3, and this is the informative failure the registration anticipated. The call rested on execution: the Neg conceded the data, the mechanism, and the causal principle (“they’re going to be a product of whatever that environment is”), and the Aff never cashed any of it. The page said in advance that a Neg win “would say the concessions were invisible to the audience, which is itself worth knowing.” It says exactly that, by 47 points. Concessions extracted but never converted are worth nothing on a lay ballot — an audience scores what a debater does with an admission, not the admission. Pairs with the Ouedrago round, where the same failure to collect went unpunished. Original registration below. Unlike most calls in this ledger it rests on execution rather than sympathy: the Neg conceded the data, the mechanism, and the causal principle, and the Aff still declined to cash it — so a Neg win would say the concessions were invisible to the audience, which is itself worth knowing |
| Summerhays/Gilkison — no confident call | R8 discussion | ✅ the abstention was correct, and that is the point. Chris read Neg (Gilkison) as winning on substance — “does this rise to the level of ‘doesn’t work’.. I don’t think it does” — while explicitly declining to predict the vote: “this might be a victim of small sample size :)”. The vote went the other way: Summerhays 71.4–28.6. So the substance read and the ballot diverged by 43 points, exactly the gap the abstention anticipated. Finding 13 earning its keep: a round turning on ~21–49 ballots cannot be forecast from the merits, and the correct move was to say so rather than guess |
| Hamm (Aff) advances, Cruz/Hamm | R7 discussion | ❌ WRONG — Cruz took it 56.2–43.8. But Chris called the shape exactly: “might be close.” At +12.4 it is by far the tightest margin in the bracket, on an inferred n of roughly 16–73 — i.e. a swing of a few votes. This was flagged in advance as the informative outcome, and what it informs is finding 13 rather than anything about argument quality: on the tape the Aff had the cleaner case (the severance held; the Neg conceded the inside standard in his own closing), the predictors all agreed, and it still went the other way by single digits of ballots. At this sample size, “the better case lost” and “four people voted differently” are the same sentence |
| Tejada (Aff) advances, Guptill/Tejada | R6 discussion | ✅ CORRECT — 94.55%, the largest margin in the bracket. Tejada spent his opening and a third of the crossfire arguing his opponent’s case and was corrected by that opponent rather than the moderator. Qualified by Chris after the result: the second half was genuinely good — the struggle/duty case is organised and he closes well — so the margin is partly earned, and this page’s original “whatever the ballot measures, it isn’t performance” is withdrawn as too strong. An 89-point margin still exceeds what a strong second half explains |
| Calvin (Aff) wins round 5 | R5 discussion | ❌ WRONG — and instructively so. The Aftermath took 71.88%. Chris’s two premises: (a) audience priors favour the luck thesis because “people who are not rich want to believe the topic”; (b) the Neg didn’t make his case. (b) still looks right on the tape. (a) is backwards for this audience — see finding 10 |
✅ = reviewed · ◐ = captured, not reviewed · ◻ = candidate · blank = not planned
Note: the channel title for the Boomers round bills “Miriah Headbangs”; the moderator introduces her as Mariah Meyers (master herbalist, co-host of Poppy’s Field Project). Rex Jones is introduced as a former Infowars reporter/contributor “before 2016.”
| Resolution | Contenders | Why / vault hook | |
|---|---|---|---|
| ✅ | Is Masculinity in Crisis? | Owlish vs. Gabriel Garcia | → review; produced Constructed ≠ Arbitrary |
| ✅ | Is Feminism Helping Modern Relationships? | Martae Ruelas vs. Miles Anton | → review; scope+time, survey-reason≠cause |
| ✅ | Is Feminism Helping Modern Relationships? | David S. vs. Tareyak | → review; moderator Ryan Mullally (his third). The rerun — same resolution as round 2, different pair, still the designed test of finding 2. Decided in the prompt negotiation: only feminism is defined, so the Aff runs a squirrel (one instance satisfies him; the Neg must defeat every interpretation) — an affirmative annexing the negative’s easy burden, the mirror of Guptill forfeiting it. Also asymmetric latitude, and the unbundling pincer. Result pending |
| ✅ | Should Military Service Be a Prerequisite for Voting? | Glenn Lawrence vs. Chris Brunet | → review; 1:13, moderator Ryan Mullally (his second). The bracket’s cleanest inflation specimen — the Aff widens military service to national service in his opening and is named for it in the Neg’s first thirty seconds — and the inflation refutes him, since granting that firefighting demonstrates commitment concedes commitment isn’t military-specific. Also the round where the Aff’s answer to the electorate-capture charge (“most lefties don’t serve”) is the charge. Both vault suffrage pages pre-adjudicate it and neither debater reaches the scope move. Result pending |
| ✅ | Is Political Party Loyalty Illogical? | Kung Fu Joe (Aff) vs. Kewl Vic (Neg) | → review; ~1:10, moderator Ryan Mullally (his fourth). Both debaters admit on air they didn’t research why the two-party system exists — the one fact that decides the resolution, and the vault has it on the shelf (Duverger). Largely a verbal dispute written into the Neg’s own definition of allegiance (conditional, “not a blood oath”), which the Aff eventually names himself; underneath it a dual-standard failure — two live definitions of logical (instrumental vs. internal-coherence), each side scoring under his own. Finding 9 at 12 of 12 |
| ✅ | Is It Impossible to Get Rich Without Luck? | Calvin Constable vs. The Aftermath | → review; the bracket’s first universal quantifier resolution, and the Aff wins by emptying “luck.” Harvest: irreducible ≠ decisive, buy the ticket, the determinism pincer |
| ✅ | Is AI Making People Dumber? | Luke Charsky vs. Dr. Philip Bunn | → review; the control case for finding 4. Extended Mind vs. phronesis; the vault’s three-layer rule cuts between them |
| ✅ | Should Free Speech Protect Hate Speech? | Chase McPherson (Neg) vs. Octavius Thunder (Aff) | → review; 1:09, moderator Kyla Turner (her sixth). The resolution gets vacated from both ends — the Neg specifies down into existing doctrine, the Aff concedes that doctrine, and the live fight migrates to hate-crime enhancement, which is sentencing rather than speech protection. Confirms finding 9 at 11 of 11; supplies the finding-1 etymology pair (Kyla blocks a definitional move without owning it, where Pisco owned his); and the two cases that would have settled it — R.A.V. and Mitchell — go unnamed. Result pending |
| ✅ | Is the Algorithm Fair? | Silvio Cruz vs. Ryan Hamm | → review; 0:57, shortest round. The pre-registered hook — “‘fair’ is a weighting question” — was confirmed independently by Chris’s compression argument. Aff wins on a severance (algorithm = automated program, so all moderation evidence is off-topic); Neg names the winning inside/outside frame in his first answer and abandons it. Result pending at review time |
| ✅ | Are Boomers Responsible for Today’s Economy? | Rex Jones vs. Mariah Meyers | → review — analyst read, not discussed (Chris opted out of this one). Answers open question 3 with layer 0; full-tape confirmation of Pisco on finding 1; already spawned Generational Attribution, which now has it as a specimen |
| ✅ | Is Modern Western Culture Making People Fragile? | Nic Turner vs. Rowan Bourdeau | → review; 1:11, moderator Sam Tripoli (the fifth). The bracket’s contrast case: the matched pair to Guptill/Tejada — same question, another degree word — but here the term is pinned mid-round and the debate immediately acquires shape. Neg concedes the data, the mechanism, and “environment makes the man,” then severs culture from circumstances. Result pending |
| ✅ | Is Therapy Culture Making People Weaker? | Josh Smith vs. Haji Ouedrago | → review; 1:10, moderator Kyla Turner (her fifth). The bracket’s best definitional work — both load-bearing terms pinned, and the Neg catches a question-begging definition in real time (refusing “therapy culture := the overuse of therapy” as smuggling the conclusion), which supplies a third manipulation type for The Load-Bearing Word. Aff wins ground by severance (CBT ≠ therapy culture) and then concedes the causal claim — “I’m not saying it’s the direct cause” — which the Neg fails to collect. Result pending |
| ✅ | Does Western Therapy Work for Men? | Spencer Summerhays vs. Michael Gilkison | → review; 1:14. Two defective definitions of works (provider-carries-everything vs. any-learning-counts); the Aff’s market argument has a hole neither man finds. Moderated by Ryan Mullally — answers finding 1’s open successor. Result pending |
| ✅ | Are We Too Sensitive as a Society? | Forrest Guptill vs. Jose Tejada | → review; 1:20, longest round so far. The Aff argues the Neg’s case for ~20 minutes before the opponent corrects him — and after the correction they go on agreeing. The bracket’s first verbal dispute |
14 of 16 first-round matchups visible on the channel as of 2026-08-05; two not yet listed.
Production note — 2026-08-17, and it is a genuine anomaly. The same two matchups are still unlisted twelve days later, while round 2 is already being filmed — reported to Chris directly by contenders, not visible from the channel. Two rounds of Thunder 32 have therefore not aired, not been queued, and not been announced, yet the tournament has advanced past them.
Sourcing: contender word-of-mouth via Chris, uncorroborated against the site or channel. Recorded because the coverage policy already flags this information channel as part of the series’ evidence base, and because it is the kind of fact that never appears in a transcript.
The release cadence, which is what makes it odd. Rounds have been posted on a 48-hour cycle, and on that schedule the two missing matchups were due today. So this is a schedule slip against an otherwise regular pattern, not a gap of unknown length.
Why it matters to the findings, as hypotheses rather than conclusions:
Coverage status as of 2026-08-17: all 14 aired rounds reviewed ✅ — round 1 coverage is complete up to the two matchups that have never been listed. Only those 2 remain.
Ana Kasparian vs. Pearl Davis · Michael Rectenwald vs. Shabbos Kestenbaum · Ryan Mullally vs. Pisco.
Update (2026-08-11): WW1 has happened. Moderating the therapy round, Mullally introduces himself as “the winner of the first ever Word War Debate in Atlantic City” — so the event this hub logged as upcoming is in the books, and Mullally won. Stated by him on air; the site’s /ww1-debate-results page would confirm the full card and is not yet pulled.
Chris: “let’s wait for more results and then we can break apart why people won or lost.”
The bracket is four results short of complete (therapy culture, feminism rerun, hate speech — all closed and unposted — plus party loyalty, closing ~08-18). Once all fourteen are in, run a cross-round causal analysis rather than another per-round review. The materials are already assembled and this is the payoff pass:
Harvest to run in the same pass — three things the intros and outros give up cheaply. Chris’s addition: “note if we can catch the real names vs internet handles from the rereads.” All three are read straight off the moderator’s scripted segments, so they cost nothing extra once a transcript is open:
@_vigilante_tv; Octavius Thunder → Eli Halpern). Harvesting all 28 makes the reach analysis runnable for the first time, because follower counts become addressable. Build it as a table: billed name · handle · platform · channel, with unknowns marked rather than guessed.Do all three in the same read; they come from the first ninety seconds and the last two minutes of each tape.
Status 2026-08-18: 12 of 14 in hand; only therapy culture and the feminism rerun are outstanding, and both closed before the two that just posted — so the pipeline publishes out of order and further waiting may not produce them. The autopsy is runnable on 12 and should not block on the last two; code them as unknown rather than dropping them.
Chris: “this should help me make better predictions for the 16 and up brackets. I have not been able to understand the audience.”
That reframes the pass. A retrospective why did each side win produces fourteen stories, each plausible and none testable. What is wanted is a model that pays out on round 2, which means the fourteen rounds are a training set, not a set of anecdotes — and the discipline that follows is: score candidate predictors against all fourteen outcomes at once, keep the ones that beat the base rate, and discard the ones that only explain in hindsight.
The leading candidate is now finding 17’s two-factor model — default to the audience’s prior on the resolution; override only on a large execution gap; rhetoric moves the margin, not the winner — which fits 11–12 of 12 retrospectively and must now be tested prospectively rather than refined further. The autopsy’s job is to try to break it, not to decorate it. Remaining candidate predictors, each cheaply codable per round:
| Predictor | Coding | Currently supported by |
|---|---|---|
| Argument quality (vault verdict) | Aff / Neg / neither | Charsky +58, Brunet +57 |
| Delivery & legibility | which side was easier to follow | Owlish +48, Garcia’s 25.8 |
| Uncollected concessions | did the winner leave admissions on the table? | Bourdeau +47, Ouedrago |
| Audience prior on the resolution | which conclusion flatters this pool | Tejada +89, Aftermath +44 |
| Affiliation / turnout | personal network size, not follower count | finding 13’s ballot counts |
| Structural edge | last word, opener, moderator’s rule | finding 9’s corrections |
| Conceded the room | did a debater signal he expected to lose? | David S., pending |
The one calibration already in hand — and the honest starting point for “I don’t understand the audience” — is finding 15: the misses are one-directional. Every wrong ballot call but Hamm came from expecting rhetoric, sympathy, affiliation, or a flattering conclusion to beat the better argument, and this pool keeps declining. So the first correction is not a new variable, it is a weighting change: stop discounting argument quality. The audience has been more legible than the model applied to it — the difficulty was never that they are unpredictable, it is that the prior was cynical.
Success test, set in advance: the model earns its keep if it beats 3-of-7 on round 2, and it should be registered before the round-2 tapes are watched, or it is not a prediction.
Chris: “this new format needs lots of polish!”
Consolidated because the observations have accumulated across a dozen reviews and are more useful as a list than scattered through findings. Ordered roughly by how cheaply each could be fixed.
| # | Defect | Evidence | Cheap fix |
|---|---|---|---|
| 1 | Speaking order is unannounced and contradicts the intro order. The Neg is introduced first in 12 of 12; the Aff usually speaks first. Two signals, opposite directions, neither flagged | Tejada argued his opponent’s case for a fifth of a round | Say which side is which before the clock starts — the bios already state the position |
| 2 | Closing order follows two different rules, one positional (Kyla) and one side-based (Mullally), and the side-based one has run in both directions | finding 9’s correction table | Pick one and put it in the format sheet. The last word is a real edge |
| 3 | Load-bearing terms are not negotiated pre-round — only the topic noun is | The feminism rerun was decided by this before it started | Require an agreed definition of the resolution’s operative term, or announce on air that none was reached |
| 4 | No live bracket. Staggered 48-hour windows plus image-only results mean nobody outside the organizers can see the standings | The reason this hub maintains one | A single standings page — the data already exists |
| 5 | Results not published when windows close. Three rounds closed with no card | therapy culture, feminism rerun, hate speech | Publish on close |
| 6 | Matchups unlisted while later rounds film | two of sixteen, never announced | Publish the full bracket at seeding |
| 7 | Moderators are drawn from the competitor pool | Pisco moderates and competes; Mullally moderates after winning WW1 | Not obviously wrong — but it makes disclosure load-bearing |
In fairness, three things the format does well, and they should survive any polish: the moderated crossfire round is the most productive segment in almost every round (it is where definitions actually get pinned); the 48-hour window lets voters watch the whole tape rather than reacting live; and disclosure norms are emerging on their own — Mullally opens the party-loyalty round with “candidly, I know both these guys.”