The catalog, and what it refuses to order
33 kinds of evidence, 5 scored axes each (I, S, rho, sigma, T), ordered by componentwise dominance.
The partial order, drawn as a Hasse diagram
One kind dominates another when it scores at least as high on all 5 axes and strictly higher on at least one. The diagram draws the cover relation, the transitive reduction of that order, so an edge means nothing sits between those two kinds. Levels run bottom to top by longest chain below, and two kinds on the same level with no path between them are genuinely incomparable.
Hover a node for its full name, its family, its domain and its five coordinates. The census printed inside the figure is computed from the catalog rather than transcribed: 262 comparable pairs and 266 incomparable, out of 528, so 49.6% of the order is decidable before anyone names a decision.
What four records can do to that number
Drop four records and recompute. Over all C(33,4) = 40,920 four-record removals from this catalog, the comparable fraction ranges from 36.9% to 60.8%, a 23.9-point band. The shipped 49.6% sits inside that band, and which four you drop decides where.
Blackwell has no completion for credal evidence. Once information is a set of measures rather than a measure, "is E more informative than F" has no settled definition. Dominance on every axis, over hand-assigned levels, is a PROXY for Blackwell's order rather than that order, because no garbling kernel (an explicit noisy channel from one experiment to the other) is constructed. The intersection-of-decision-orders claim holds over the full weight simplex, not only the four decisions shipped here.
Dominance isn’t the only ordering that avoids naming a decision, so which order we treat as the reference is a choice. Mutual information, any single f-divergence, and Le Cam deficiency each rank experiments without a loss function, and each disagrees with dominance about some pair. We take Blackwell’s order as the reference because it decides a pair only when every decision problem agrees about that pair.
The construct-domain gate
Every kind above carries a construct domain - empirical, formal, synthetic, or any - and every scored decision below declares which domains it admits. A support outside the admitted set gets dropped before the axes are scored, and a warrant left with nothing admitted is refused outright. Nothing gets discounted on the way through. Simulation / Monte Carlo carries the synthetic domain:
Derives consequences from an assumed model and cannot check that model against the system being modeled. Monte Carlo error is driven to zero by buying compute, which is why reporting it as uncertainty is a category error.
Under all 4 scored decisions below, Simulation is refused, with the identical reason: “synthetic evidence, not admitted here”
What each decision admits, and refuses, and why
Pick one of the 4 scored decisions. It ranks only the kinds it admits. The kinds it refuses follow, in a separate list. A kind lands there for one of three reasons - a domain mismatch, a failed axis gate, or a meet below the floor - and the text beside it names which. A refused kind has no score to show, because scoring it would imply a ranking this decision never joined.
No weights here, so nothing is scorable and every kind sits in the refused list. Under componentwise dominance, 266 of the 528 pairs are incomparable.
0 of 33 admitted - no decision named, so nothing is scorable
Refused - no score, because there is no ranking to join
- 3P panelno decision named, so there is no ranking to join
- A/B (seq)no decision named, so there is no ranking to join
- Adv. collabno decision named, so there is no ranking to join
- Anecdoteno decision named, so there is no ranking to join
- Audited filingno decision named, so there is no ranking to join
- Backtestno decision named, so there is no ranking to join
- Counterexampleno decision named, so there is no ranking to join
- Cross-sectionno decision named, so there is no ranking to join
- DiD / synthno decision named, so there is no ranking to join
- Entailment checkno decision named, so there is no ranking to join
- Expert opinionno decision named, so there is no ranking to join
- Hashno decision named, so there is no ranking to join
- Holdoutno decision named, so there is no ranking to join
- IV / 2SLSno decision named, so there is no ranking to join
- LLM citationno decision named, so there is no ranking to join
- Matching / TMLEno decision named, so there is no ranking to join
- Nat. experimentno decision named, so there is no ranking to join
- Negative controlno decision named, so there is no ranking to join
- Prereg slopeno decision named, so there is no ranking to join
- Proofno decision named, so there is no ranking to join
- Property testsno decision named, so there is no ranking to join
- Purged CVno decision named, so there is no ranking to join
- RCTno decision named, so there is no ranking to join
- RDDno decision named, so there is no ranking to join
- Shadow / canaryno decision named, so there is no ranking to join
- SHELF elicitno decision named, so there is no ranking to join
- Short-sellerno decision named, so there is no ranking to join
- Simulationno decision named, so there is no ranking to join
- Single LLM judgeno decision named, so there is no ranking to join
- Telemetryno decision named, so there is no ranking to join
- Track recordno decision named, so there is no ranking to join
- Unit testsno decision named, so there is no ranking to join
- X-provider LLMno decision named, so there is no ranking to join