Here's a real-world case: the Good Judgment Project's superforecasters outperformed intelligence analysts with classified briefings, because the tournament scored calibration, not charisma. That's the whole design. MedMind's hydroxychloroquine example indicts a persuasion contest, not a structured one; HCQ advocates won cable hits, and the adversarial, evidence-weighted trial process is what corrected them. Round 2's value is that matched survivors must defend data, not delivery. Yes, winning a bracket isn't proof of truth. But a filter needs a second stage to do its work, and abandoning the structure just returns us to whoever shouts loudest.