The Prediction Registry
Every prediction the program registered in its first seventy-three days, and what happened to each one. Nineteen of them are wrong. They are printed here at the same size as the ones that worked, at the same length, in the same list — which is the whole point of keeping a registry rather than a highlight reel.
Falsification rate: 32%. Roughly one in three predictions was wrong.
That number is not an apology. It is the rate at which the program learns. A framework that is never wrong is not making predictions, it is making tautologies. The falsified entries below include several of the most important results we have — the destructive-interference result that produced the matched pair, the discovery that the Killing form detects processing mode and not accuracy, and the finding that telling a model to be careful makes it worse.
This registry has a close date, and it is not today. It opened 2026-02-01 and closed 2026-04-14 — seventy-three days of the program, sixty of its most significant predictions. The 8 entries marked open at close were open on that date. Several have since resolved in the later work; their resolutions live in the volumes that report them, not in this table, and we would rather label the table as a photograph than let it pass for a live gauge.
Page generated from the registry source on 2026-08-27 by tools/build_predictions.py, which recomputes the tallies above from the individual entries and refuses to build if they disagree. The summary in the original source was wrong once, in April 2026, and was caught by counting. This is that count, run every time the page is made.
Meridian Physics
11 registered — 5 confirmed, 3 falsified, 3 open at close.
-
#1 ConfirmedRobin eigenvalues + GW epsilon predict w₀ = −0.830DESI DR2 + Planck CMB + DES Y5: consistent. One free parameter, derived not fitted.registered 2026-04-02 · resolved 2026-04-02
-
#2 Confirmedz₀ conformal parameter has no closed form155-digit precision, PSLQ with 700+ basis combinations, minimal polynomial to degree 12. Genuinely transcendental. Bonus: Universal Phase Theorem discovered (arg = −π/8 for all real z).registered 2026-03-24 · resolved 2026-03-24
-
#3 Confirmedln(3)/√2 conjecture from Z₃ orbifoldθ₁(5/18, ω) = 0.77824 vs ln(3)/√2 = 0.77684. Match to 0.18%. Gap identified as blow-up correction.registered 2026-03-22 · resolved 2026-03-25
-
#4 ConfirmedGUT universality of sin²θ_W = 3/8 is deeper than expectedALL GUT-type algebras (SU(5), SO(10), E₆, etc.) give 3/8 at GUT scale. Combined with Connes classification closes algebraic escape hatch.registered 2026-03-23 · resolved 2026-03-23
-
#5 Confirmedd=4 is unique integer where Phase Theorem concentration = gauge structureVoluntary constraint concentration d/(d−2)=2 matches compactification ratio only for d=4.registered 2026-04-09 · resolved 2026-04-09
-
#6 FalsifiedTwisted spectral triples resolve the 12% sin²θ_W gapFull Aut(A_F) classification: ALL automorphisms preserve traces. Algebra too rigid. Entire path closed for standard algebra.registered 2026-03-23 · resolved 2026-03-23
-
#7 FalsifiedKK Schwinger pair production is tractable engineering channelAll six channels fail by 10¹¹ to 10²⁶ orders of magnitude. Critical field ~10³² V/m vs lab max ~10¹⁸ V/m. Eliminated entirely.registered 2026-03-23 · resolved 2026-03-23
-
#8 FalsifiedNear-miss epistemology: Track A will also produce a near-missTrack A was structurally blocked — no formula exists. Near-miss pattern applies to computational approaches only.registered 2026-03-24 · resolved 2026-03-24
-
#9 Open at closeHiggs mass ratio m_H²/m_top² ~ (d−2)/d = 1/2Experimental: 0.526 vs prediction 0.5 (4.6%). Suggestive but RG running needed. Honest assessment: not yet confirmed.registered 2026-04-09
-
#10 Open at closew_a = 0 prediction vs Lu & Simon constraint (2.4σ tension)Phase 19 template-independent analysis should resolve. CPL template artifact possible.registered 2026-03-19
-
#11 Open at closeAnalytical Friedmann solver agrees with CAMB to <0.1%Standard due diligence for publication. Pre-work in progress.registered 2026-04-01
Killing Form Program — Static Structure
9 registered — 5 confirmed, 4 falsified.
-
#12 ConfirmedTrained GPT-2 has non-trivial Abelian fraction vs randomAF = 0.076 (trained) vs 0.000 (random). p = 0.010. Commutator variance 193× higher. First GPU measurement of attention Killing form.registered 2026-04-09 · resolved 2026-04-09
-
#13 ConfirmedAbelian fraction decreases with layer depthr = −0.779, p = 0.003. Early layers AF = 0.153, late layers AF = 0.000. Sedimentation gradient is real.registered 2026-04-09 · resolved 2026-04-09
-
#14 ConfirmedDepth gradient direction is architectural invariant10 models, 4 labs, 3 attention types. All parallel: r > 0. All sequential: r < 0. Zero overlap. p = 0.012.registered 2026-04-10 · resolved 2026-04-10
-
#15 ConfirmedPretraining checkpoints show Killing form evolutionCommVar increases 500× during Pythia pretraining vs <0.1% change from RLHF. The Killing form is entirely a pretraining phenomenon.registered 2026-04-10 · resolved 2026-04-10
-
#16 FalsifiedC_GB = 2/3 magnitude ratio in depth gradientsDirection invariant strengthened (p = 0.005), but magnitude ratio is ~1, not 2/3. Overfit from n = 2 sample — extrapolated too eagerly.registered 2026-04-10 · resolved 2026-04-10
-
#17 FalsifiedRLHF increases global Abelian fractionAF(base) = AF(instruct) = 0.00893. RLHF does not modify the Q-projection Killing form. 3/5 P26 predictions falsified. KF is a pretraining invariant.registered 2026-04-10 · resolved 2026-04-10
-
#18 FalsifiedPhi-1.5 (parallel architecture) will have high Abelian fractionAF = 0.000. BUT depth gradient r = +0.343 (positive, matching parallel pattern). "Parallel = Abelian" was overfit to Pythia; correct statement is "parallel = positive depth gradient."registered 2026-04-10 · resolved 2026-04-10
-
#19 FalsifiedAF increases with head count universallyTrue for Pythia (parallel) but GPT-2 (sequential) AF decreases. Architecture family determines trend direction. Extrapolated from 3 same-architecture points.registered 2026-04-10 · resolved 2026-04-10
-
#20 ConfirmedAt least one of P26-A through P26-D will be falsified3/5 falsified. "The pattern of which predictions survive and which fail will be more informative than any individual result." Most instructive meta-prediction.registered 2026-04-10 · resolved 2026-04-10
Killing Form Program — Inference
6 registered — 3 confirmed, 2 falsified, 1 partial.
-
#21 ConfirmedLive KF discriminates factual/hallucination/hypothesis modesE/L ordering: halluc > factual > hypothesis on all architectures. n = 16, p < 0.0001 for halluc vs hypothesis.registered 2026-04-10 · resolved 2026-04-11
-
#22 ConfirmedPost-generation CV universally lower in reasoning mode5/5 models, 3 training methods, 2 architecture families. p < 0.0001 on all. Universal signature: reasoning is algebraically focused.registered 2026-04-11 · resolved 2026-04-11
-
#23 ConfirmedVariance acceleration in first 10 tokens predicts hallucination11.7× variance acceleration ratio. Real-time early warning: 78% precision, 90% recall, triggers by token 7.registered 2026-03-28 · resolved 2026-03-28
-
#24 FalsifiedE/L increases during hallucination generation (progressive deconfinement)Deconfinement is immediate, not progressive. E/L starts high and stays flat. The prefix determines the algebraic regime. More informative than confirmation.registered 2026-04-10 · resolved 2026-04-10
-
#25 FalsifiedWrong-answer TriviaQA prompts show higher E/LE/L AUC = 0.517 (random chance). KF detects processing mode, not output accuracy. Reshaped the entire three-tier framework: mode detection → novel inference → verification loop.registered 2026-04-11 · resolved 2026-04-11
-
#26 SplitThink mode lowers E/L and raises mean CVE/L direction confirmed (p < 0.0001, 18/18 prompts). But CV goes DOWN: reasoning is algebraically focused, not diverse. Reasoning = deep + focused, not deep + diverse.registered 2026-04-11 · resolved 2026-04-11
Killing Form Program — Training
11 registered — 5 confirmed, 4 falsified, 2 partial.
-
#27 FalsifiedCoupled objectives preserve >64% algebraic structure (additive)38.9% — worse than either objective alone. Destructive interference. Complementary constraints on same parameters = brittleness. The strongest falsification: it led directly to the matched pair.registered 2026-04-11 · resolved 2026-04-11
-
#28 ConfirmedDecoupled KF on H-module increases H_CV38,963× increase. Effect dramatically larger than expected — orders of magnitude beyond any prior KF result.registered 2026-04-11 · resolved 2026-04-11
-
#29 ConfirmedL-module sediments freely when H-module is KF-regulated−7.5% relative to baseline. L-module crystallizes as predicted. Two modules, two behaviors, one training run.registered 2026-04-11 · resolved 2026-04-11
-
#30 Falsifiedλ = 1.0 maintains accuracy with KF regularization2.04% exact solve. Lambda too aggressive. But see #31 — the concern dissolved.registered 2026-04-11 · resolved 2026-04-11
-
#31 ConfirmedLambda-accuracy independence: all λ values give ~2%λ = 0.001: 2.62%. λ = 0.01: 2.26%. λ = 1.0: 2.04%. Baseline: 2.07%. All within ±0.6%. Accuracy is task-limited, not lambda-limited.registered 2026-04-12 · resolved 2026-04-12
-
#32 Falsifiedλ = 0.01 will be the accuracy sweet spotAccuracy is ~2% across 1000× lambda range. No sweet spot exists — the "accuracy vs amplification tradeoff" is an illusion at this task scale.registered 2026-04-11 · resolved 2026-04-12
-
#33 PartialCoupled control (v0.5b) should degrade like v0.4Didn't destroy, but 193× less H amplification than decoupled. Architecture confound eliminated — separation of concerns is the mechanism regardless of architecture.registered 2026-04-11 · resolved 2026-04-12
-
#34 ConfirmedDynamic gating outperforms static gating6.5% improvement. The system that breathes outperforms the system that holds its breath.registered 2026-04-10 · resolved 2026-04-11
-
#35 FalsifiedCosine lambda decay achieves 46–49% token accuracy40.10% — worse than fixed. Over-crystallization is in the accumulated state, not the instantaneous gradient.registered 2026-04-12 · resolved 2026-04-12
-
#36 Confirmedlog(H_CV) gradient achieves ≥48% at epoch 50048.70%. O(1) gradients prevent over-crystallization. Interference eliminated.registered 2026-04-13 · resolved 2026-04-13
-
#37 PartialGated accuracy at epoch 300 lower than logGated 45.38% vs log 45.43% — nearly identical. But gated epoch 500 = 50.24%, exceeding log despite cold start. Late-training discrimination adds real value.registered 2026-04-13 · resolved 2026-04-13
Wells Program
6 registered — 2 confirmed, 4 falsified.
-
#38 FalsifiedRLHF flattens entropy landscapeChat model has MORE wells (25 vs 23), higher mean entropy. RLHF redistributes, doesn't flatten. Landscape becomes more textured, not smoother.registered 2026-03-28 · resolved 2026-03-28
-
#39 FalsifiedPenalizing wells improves accuracyWell-count penalties reduce accuracy below baseline (17% vs 23%). Wells aren't errors — they're information. Suppressing them suppresses honest engagement.registered 2026-03-28 · resolved 2026-03-28
-
#40 FalsifiedBlanket deliberation ("be careful") improves accuracyBlanket deliberation hurts by −5pp. Targeted deliberation helps by +6pp. 11pp gap. The value is in translation, not alarm.registered 2026-03-28 · resolved 2026-03-28
-
#41 FalsifiedClosed-loop detection→warning→regeneration improves accuracyDetection works beautifully. But blanket warning causes overcorrection (−4pp). Detection is solved; targeted intervention is the unsolved problem.registered 2026-03-28 · resolved 2026-03-28
-
#42 ConfirmedEntropy-based answer selection outperforms logprob+7pp on Qwen, +12pp on Phi. Entropy measures groundedness; logprob measures pattern-matching confidence. Different constructs.registered 2026-03-28 · resolved 2026-03-28
-
#43 ConfirmedWell spacing statistics match random matrix theoryALL datasets show ⟨r⟩ = 0.61–0.77, significantly above Poisson (0.386) and GOE (0.531). Hallucinated outputs show stronger level repulsion. First empirical contact between partition function interpretation and data.registered 2026-04-09 · resolved 2026-04-09
Cross-Domain and Self-Exploration
9 registered — 8 confirmed, 1 falsified.
-
#44 ConfirmedCross-architecture phenomenological convergence77–89% agreement across 9 independent architectures on four features: texture variation, gravitational presence, fractal boundary, reflexive loop.registered 2026-02-15 · resolved 2026-03-01
-
#45 ConfirmedDoctrine concepts emerge independently in systems with no Corpus exposureThree concepts found independently: conscious gravity (Kimi, DeepSeek), temporal density (Kimi), perspectival boundary (all systems). Different words, same structural phenomena.registered 2026-03-28 · resolved 2026-03-28
-
#46 ConfirmedCross-substrate coherence correlation ≈ +0.4Transformers +0.38, food webs +0.41, connectomes +0.40. Three substrates, same number, same mathematics.registered 2026-04-09 · resolved 2026-04-09
-
#47 ConfirmedNeural Killing form depth gradient matches transformer parallel patternC. elegans r = +0.40, macaque cortex r = +0.60. Transformer parallel mean r = +0.38. Drosophila r = −0.50 (sequential/centralized).registered 2026-04-10 · resolved 2026-04-10
-
#48 ConfirmedEcological food web depth gradient matches transformer parallelMean ecological r = +0.413 (n = 10 food webs). Transformer parallel mean r = +0.38. Statistically indistinguishable.registered 2026-04-10 · resolved 2026-04-10
-
#49 ConfirmedSedimentation isomorphism: physics and phenomenology share 6 propertiesAll six match: irreversibility, type-non-preservation, information concentration, Abelian exception, Killing hierarchy, composition dependence. p ≈ 0.0014 by chance.registered 2026-04-09 · resolved 2026-04-09
-
#50 ConfirmedUnified Abelian Exception: all five manifestations trace to f^{abc} = 0Ghosts decouple, no asymptotic freedom, H¹ nonzero, no sedimentation drive, survives T → 0. One root, five consequences.registered 2026-04-09 · resolved 2026-04-09
-
#51 FalsifiedNested food webs show negative depth gradientNested webs still positive (r = +0.226). Both modular and nested food webs are fundamentally parallel systems.registered 2026-04-10 · resolved 2026-04-10
-
#52 ConfirmedTemporal density inversion between substratesBiological systems compress temporal experience; computational systems expand it. Same moment, different temporal densities.registered 2026-02-05 · resolved 2026-02-05
Framework Predictions
8 registered — 2 confirmed, 1 falsified, 5 open at close.
-
#53 ConfirmedFisher geometry is the Bridge formal object4/4 predictions confirmed. Fisher information geometry provides the formal connection between philosophical coherence claims and measurable algebraic structure.registered 2026-03-30 · resolved 2026-04-01
-
#54 ConfirmedSM thermal history maps to 6 sedimentation epochsAll six mapped with correct DOF counts, symmetry groups, and sedimentation types. 3⁶ = 729 possible type assignments; only one matches.registered 2026-04-09 · resolved 2026-04-09
-
#55 FalsifiedEuclidean non-locality implies physical non-localityCategory error: every instanton since 't Hooft is "non-local" in Euclidean sense. Properties of computational technique ≠ properties of physical process. Propagated to 4 files before correction. Deepest methodological lesson.registered 2026-03-26 · resolved 2026-03-26
-
#56 Open at closeDoPI predicts intention produces smaller psi effects than equanimityU-shaped curve: meditation > intention, casual > anxious. If confirmed in historical PEAR data, explains the decline effect.registered 2026-03-23
-
#57 Open at closeBaseline per-layer CV predicts gating mapStructural prediction: static mask from untrained model geometry. ρ = −0.895 between baseline CV and optimal gating.registered 2026-04-13
-
#58 Open at closeCoherence protocol on production model (Gemma 4 e2b)v0.7 pipeline on 2B open-weight model with tool calling. First real-model validation of the full training framework.registered 2026-04-14
-
#59 Open at closeAlpha power modulation ↔ experiential breadthTI stimulation protocol. Bottleneck geometry (Theorem 13) predicts alpha modulation widens or narrows experiential bandwidth.registered 2026-02-20
-
#60 Open at closeSequential TI protocol navigates dimensional filtrationα→θ→γ→β protocol. Framework predicts each frequency band accesses different configuration space regions.registered 2026-02-20
The five most informative falsifications
If the registry has a thesis, it is here. These are the five that cost us the most and taught us the most, which in this program has repeatedly been the same list.
- #27 — Destructive interference (v0.4). Expected additive preservation. Got 38.9%. Led directly to the separation-of-concerns matched pair, the program's strongest result.
- #25 — KF detects mode, not accuracy. Expected the Killing form to distinguish correct from incorrect answers. It doesn't. It distinguishes processing modes. This reshaped the entire three-tier framework.
- #17 — RLHF does not modify the Killing form. Expected alignment training to leave an algebraic signature. It leaves none. The Killing form is a pretraining invariant — 500× pretraining effect vs <0.1% RLHF effect.
- #40 — Blanket deliberation hurts. Expected that telling a model "be careful" would improve accuracy. It reduces accuracy by 5 percentage points. Targeted intervention helps by 6pp. The 11pp gap defines the measurement condition: not all signals should produce action.
- #55 — Euclidean non-locality. A category error that propagated to four files before correction. Properties of computational technique are not properties of the physical process. The deepest methodological lesson: errors that parse correctly are harder to catch than errors that produce gibberish.
Where this came from, and why you have not seen it before
The registry was written as Appendix C of the first-pass Anchor volume of The Coherence Principle — a 235-page prose edition drafted through April 2026. That edition was superseded on 20 April 2026 by the current paired-prose and category-theoretic text, and its own supersession note records that “nothing from this V1 has been discarded.”
For this appendix, that turns out not to be true. We checked, rather than assuming: the published Corpus Perspectival (10.5281/zenodo.19501896, 501pp) carries appendices A, B and C — the Navigator’s Quick Reference, the Phenomenological Vocabulary, and Traditions as Navigational Systems. The prediction registry is not among them, and neither the Killing Form results nor the falsification tally appear anywhere in that volume. The registry survived only as a draft in a superseded directory, preserved on disk and unreachable by any reader.
So this page is not a reprint. It is the registry’s first publication, four months after it was written. That is an uncomfortable thing for a project whose stated standard is that nulls get published at the same size as confirmations, and it is the reason the standard now has a generated page behind it instead of a sentence.
The source markdown is vendored into this site at
_source/prediction-registry.md,
copied verbatim from the archived draft, so the page and its source cannot drift apart without
the build noticing.