LOSS MANIFEST - every objective this program has back-propagated, rated
#TAG:loss_manifest #TAG:accumulation #TAG:losses #TAG:rubric
[CENSUS-VERIFIED 2026-07-25 - three independent censuses: canon+MANIFEST, history+sessions, full code tree]
Living registry (SCALE_MANIFEST discipline): entries are NEVER deleted - demoted with a
reason and a date. Rubric v1; every rating below is RECOMPUTABLE from its six sub-score
digits and the lookup table in THE RUBRIC. Machine-readable sidecar:
inventory/loss_manifest.json. Views + lint: python skill/loss_view.py card|ladder|lint.
Taxonomy and format laws: canon/loss_accumulation_forms.md.
THE RATING ANSWERS: "how much would I stake on this term in a NEW bed" - NOT effect size. InfoNCE is the most potent term in the census and rates 2.
THE STRUCTURAL FINDING (read before any row)
Only three primitives ever receive a backward pass in this program: F.cross_entropy,
squared error, and F.kl_div (plus fp64 Cayley-Menger determinants in gauges). No margin,
triplet, hinge, contrastive-margin, or load-balancing loss has ever existed in the tree.
The four objectives that look composed are those same primitives under different accumulation formats: InfoNCE is CE over an N x N grid; blob-LP-x0 is squared error dose-coupled and band-routed; HP/LP band roles are squared error under a cosine crossfade; CV is a determinant statistic accumulated in fp64.
So the loss surface is three primitives wide and eleven accumulation formats deep, and every loss discovery this program has made lives on the second axis. Two receipts:
- The addr_head cure changed accumulation only, primitive held fixed. 5.6650 bpb with
usage perplexity 1.88/64 (collapsed) -> 2.47 by going slot-parallel. Same
F.cross_entropyon both sides. (L-112 -> L-012) - Chunked CE and plain CE are identical mathematics and a 5x memory law. 22.8GB dedicated + 42.8GB silently spilled to shared memory -> 8.8GB peak at 1.03 s/step. The law lives entirely in the reduction schedule. (L-004)
A third receipt sits in the forbidden class. A10 - accumulation that carries state across
steps (EMA codebooks, commitment counters, k-means centroids) - contains exactly two
entries, and both are prohibitions (L-105 VQ/commitment/EMA, L-134 load-balancing). Not
one working objective in the program's history has ever used that format. It is precisely
where a newcomer would naturally write code, and the empirical warrant for not doing so is
L-083: the codebook stays 125+/128 axes alive at div_weight = 0, with no diversity term
at all.
A9 (raw sum, no denominator) is empty for a duller reason: the scale rides on batch and
sequence length, so the learning rate stops transferring between beds.
THE RUBRIC (how every number below was produced)
Six sub-scores, summed to S_raw (0-13), then a lookup table - not arithmetic, so
anyone can re-derive a rating without trusting a float:
| sub-score | 0 | 1 | 2 | 3 |
|---|---|---|---|---|
| R replication | designed, never run | one seed, one bed | >=2 seeds OR >=2 beds | >=2 seeds AND >=2 beds/substrates |
| P potency vs its own gauge's noise | null / within noise / wrong direction | directional but <1% or below seed spread | 1-10%, or an ordinal win above seed spread | >10%, categorical enable, or 50-200x separation |
| D doctrinal fit | violates a MANIFEST law | permitted but antagonistic | aligned/reconstructive, or a gated exception with a written statute | - |
| C cost | recorded catastrophic-cost event | material but budgeted | free or near-free | - |
| I instrument risk | headline rests on a DISTRUSTED gauge, or the term can silently become zero with no assert | aggregate-only; gauge-blindness plausible | dedicated non-blind gauge AND a silent-zero assert | - |
| +L law bonus | - | encoded as a MANIFEST law and enforced in code (assert, refusal, or guard) | - | - |
S_raw: 0 1 2 3 4 5 6 7 8 9 10 11 12 13
rating: 1 1 2 3 3 4 5 6 7 8 8 9 10 10
The nine traceability rules (applied in this order, after the lookup):
- UNRUN CEILING -
R=0=> rating <= 3. A beautiful design does not score on paper. - BLIND-GAUGE DEMOTION - headline from a DISTRUSTED gauge, not re-measured => -2.
- SINGLE-SEED CAP -
R<=1=> rating <= 6. - SUB-1% CAP -
P<=1=> rating <= 5. Four-for-four directional at 0.05-0.2% cannot outrank a measured 5.9%. - DOCTRINE-ZERO CEILING -
D=0=> rating <= 2. - RETRACTION FLOOR-BREAKER - a dated retraction, demotion, or REFUTED verdict anywhere in the record => rating = 1, unconditional. Sub-scores are still printed so you can see why it once looked good. Retracted entries never compete on the ladder.
- CONTRADICTION - unreconciled records get a RANGE or a marked sub-claim, never an average.
- PROVENANCE - every sub-score digit is backed by the row's citations. An uncited digit forces the unrun ceiling.
- SCOPE / GAUGE SPLIT - a term whose standing differs by scope or by judging instrument gets TWO adjacent rows, never one averaged row.
Calibration - three pairs that prove the rubric works
- L-048 rates 8 and L-123 rates 4. Same objective, same seeds, same data. The only difference is which gauge judged it: the role-aligned in-bed gauge, or aggregate eps-MSE. This pair is the concrete price of gauge blindness.
- L-047 rates 6 and L-111 rates 1. Same operator - Procrustes. As a regularizer it tightens CV to .19; as a training force it produced R@1 = 0.000 and P_cos stuck at .094 for thirty epochs. Campaign law 2 restated in loss terms: placement decides load-bearingness.
- L-113 rates 2 and L-001 rates 10. The most potent measured objective in the census against the plainest one. If that ordering is surprising, the rubric is working: potency and standing are different axes.
THE GRID - PRIMITIVE x ACCUMULATION FORMAT
| accumulation | CE | SQ | KL | DET | BREG | - | row |
|---|---|---|---|---|---|---|---|
A0 uniform-mean |
28 | 29 | 2 | 3 | 4 | 22 | 88 |
A1 chunk-sum-renormalize |
1 | . | . | . | . | . | 1 |
A2 per-sample-then-weighted |
4 | 3 | . | 1 | . | 2 | 10 |
A3 band-crossfade |
. | 3 | . | . | . | 6 | 9 |
A4 masked-denominator |
5 | . | 1 | . | . | . | 6 |
A5 dose-coupled |
. | 2 | . | . | . | 1 | 3 |
A6 paired-difference |
. | . | . | . | . | 3 | 3 |
A7 grid-pairwise |
11 | . | 1 | . | . | . | 12 |
A8 fp64-accumulate |
. | 1 | . | 3 | . | 4 | 8 |
A9 sum-no-norm |
1 | . | . | . | . | . | 1 |
A10 EMA / external cross-step state |
1 | 1 | . | . | . | . | 2 PROHIBITIONS ONLY |
NA not an accumulation |
. | 1 | . | . | . | 25 | 26 |
Rating distribution (n=169, mean 5.95): 1x31 2x2 3x4 4x3 5x32 6x23 7x8 8x21 9x26 10x19
Glyphs: † retraction floor-breaker | ✖ doctrine ceiling | ⟂ split sibling (scope or gauge) | ⚠ unreconciled contradiction
THE FORMULA MAP - every entry's mathematics and where it runs
Extracted 2026-07-25 (operator directive: the process is the ROSTER). 'NOT RECORDED' is itself information - those formulas were never written down anywhere and their numbers cannot be reproduced without re-derivation. Machine copy in the sidecar.
| ID | FORMULA | IMPLEMENTATION |
|---|---|---|
| L-001 | L = mean((dec(z) - x)^2); convergence endpoint = bitwise-exact recon |
geolip_svae (installed pkg) . loss_forms::prim_sq+a0_mean |
| L-002 | M_hat = sum_k sinh(u_k)A_k / sum_k cosh(u_k), u = cos(x,A)/tau; L = mean((dec(M_hat)-x)^2); codebook grad ONLY via num@A and cos=M@A^T |
amoe.core.address::m_hat + geolip_svae aleph_model.py . loss_forms::prim_sq |
| L-003 | L = mean(-log softmax(W h)[y]) |
pod/v35_exp001_placement.py:170 . loss_forms::prim_ce+a0_mean |
| L-004 | L = sum_chunks CE_sum(h[i:i+512]) / n_live_tokens (ONE global denominator) |
loss_forms::a1_chunked_ce . pod/v35_exp003_termination.py:128-138 |
| L-005 | CE(logits[:,:-1], y[:,1:], ignore_index=-100) |
pod/v35_exp006_math.py:180 . loss_forms::prim_ce+a4_masked |
| L-006 | model(**batch, labels=y).loss (masked shift-CE inside HF; vision tower fires) |
pod/v35_exp004_caption.py:175 (chunking unavailable on this path) |
| L-007 | same CE; trainable set = per-block dispatch key matrices ONLY |
closeout amoe/train/aligner.py:41-67 |
| L-008 | CE(logits, y) on CIFAR-10 |
tools/addressed_distillation_v1.py:301 |
| L-009 | x_t = sqrt(abar_t)x0 + sqrt(1-abar_t)eps, t~U{0..999}; L = mse(unet(x_t,t,c), eps); CFG drop p=.1 |
pod2/dexp006_sd15core_relay.py:132-146 . amoe.diffusion.train.objectives |
| L-010 | s = warp(u; shift=2.5); x_t = (1-s)x0 + s*eps; L = mse(pred, eps - x0); x0 = x_t - s*v EXACT/LINEAR |
pod2/dexp001_sd15_relay.py:151-164 . amoe objectives.py:51-58 |
| L-011 | `M_hard = sign(cos[argmax | cos |
| L-012 | feats = concat_p M_hat^(p)(slots); logits = W feats; CE. P=4/16/32/64 dose |
tools/ar_differentiation_bed.py (addr_msl* arms) |
| L-013 | reads at multiple tau; concat -> logits; CE (stroboscope) |
tools/ar_differentiation_bed.py (addr_3tau arm) |
| L-014 | logits = head(M_hat) directly (reconstructive read consumed in AR); CE |
tools/ar_differentiation_bed.py (addr_mhat arm) |
| L-015 | shift-CE(-100) on stepwise-CoT target sequences vs direct-answer targets |
pod/v35_exp013_experts.py:211 |
| L-016 | L = mean_B[ mse_vec + lam*w_HIGH(s01)*blob_lp ]; blob_lp = sum(blob*(LP(x0h)-LP(x0))^2)/(sum(blob)*C); x0h = x_t - s*v_hat |
closeout amoe/diffusion/train/objectives.py:63-69 [REFUSES on eps] |
| L-017 | sym CE over sims = za@zb^T/0.07 with in-batch labels |
loss_forms::a7_grid_infonce . tools/qwen_exp011_newanchors.py:445-453 |
| L-018 | `InfoNCE1.0 + Procrustes_SVD0.3 + | CV(bank)-0.20 |
| L-019 | InfoNCE + MSE + BCE + Procrustes + CV + spread (6-term, never ablated) |
external (Expert Soup, Form 2) |
| L-020 | mse(norm(pred), norm(tgt)) + (1 - cos(pred, tgt)) on (B,77,768) |
external (Form 4) . pieces in loss_forms |
| L-021 | w = min(SNR,5)/(SNR+1) velocity-adjusted; L = mean(w * mse_vec) |
external (KSimplex trainer) . loss_forms::a2_weighted |
| L-022 | L = CE + a*KL(log_softmax(student), mean_k softmax(teacher_k).detach()), a<=0.25, never founders |
tools/exp014_genetic_distillation.py:249-255 . loss_forms::prim_kl+kd_guard |
| L-023 | fact rows: KL(teacher) ONLY (CE masked off); clean rows: CE - row-routed channels |
tools/exp019_content_retention.py:158-176 |
| L-024 | GPA: mean shape after per-teacher Procrustes to consensus (delta<1e-8); student anchors init from it |
tools/exp014-016 (dif line) |
| L-025 | projective ICP: iterate sign-aligned Procrustes on RP; germline overwrite |
tools/exp015_content_heredity.py (dif line) |
| L-026 | prox = exp(-(cv-target)^2/2sig^2); L = (1+boost*prox)*mse + pen*(1-prox) |
geolip_core.core.distinguish (installed pkg) . loss_forms::soft_hand_weights |
| L-027 | conv := fold(m_hat(unfold(x))); no plain filter, no ReLU; CE on head |
geolip-aleph-classification (external repo, cls line) |
| L-028 | `score(v) = logP(x_i=v | x_masked) - logP(x_i=WT |
| L-029 | CE; trainable = aleph MslRelay adapters on frozen GPT-2 (<1%) |
tools/exp013_augmentation_bed.py track C |
| L-030 | CE on frozen CLIP-L token-AR (matched transforms + shared vocab proj) |
tools/exp013_augmentation_bed.py track A |
| L-031 | SOURCE RECOVERED 2026-07-25: attraction (1-s_y)^2 + 0.5*repulsion sum_{c!=y}(s_c^2) + 0.5*margin relu(max_{c!=y}s_c - s_y + .3) + 0.1*range clamp; siblings: prototype (cos-to-learned-prototypes + diversity) and hierarchical (coarse+fine+consistency) |
geofractal.losses.pure_geometric (installed pkg, VERBATIM) . tools/geobasin_bed.py |
| L-032 | SOURCE RECOVERED 2026-07-25 (GBC head, geofractal/model/experiment_geometric_basin.py:118): compat = triadic (resonance-weighted cos to class signatures) x self-similarity (level-pair pattern match) x Cantor coherence exp(-(m - proto_c)^2/.1) x hierarchical early/late check - the 4-factor product |
geofractal.model.experiment_geometric_basin::GeometricBasinCompatibility (installed pkg) |
| L-033 | mask patches; L = mse(recon_from_antipode_read(masked), x) (law 2 ORIGINAL form) |
built-unrun in geolip-aleph-classification (cls line) |
| L-034 | w=.05 entropy-balanced alignment (exact form NOT fully recorded); endpoint M = +/-ref EXACT |
external (SVAE H2 arc) |
| L-035 | rectified-flow velocity mse + Min-SNR gamma=5 + CM terms (L-045/L-046) |
external (KSimplex 4.8M prior) |
| L-036 | `ArcFace cos(th+m) | CosFace cos(th)-m |
| L-037 | CE on x0 tokens from noise-level-t corrupted tokens, iterative |
q25-e008 bed (pod archive) |
| L-038 | L = mse(ea, norm(frozen_h @ fixed_proj)) + mse(eb, ...) (absolute target regression) |
tools/qwen_exp011_newanchors.py:449-453 |
| L-039 | standard contrastive; measured as a CV-compression force |
external (cv_band_origin sweep) |
| L-040 | V = sqrt(clamp(-det(CM(A[idx5]))/9216)); L += 1e-3 * std(V)/mean(V); fp64, fixed seed-0 subsets, S15 bank ONLY |
tools/exp017_aleph_constellation.py:154-186 . loss_forms::cv_band_loss (NEW, arm-gated) |
| L-041 | `M = F.normalize(M, dim=-1) ( | |
| L-042 | `per tower: g <- g * target/ | |
| L-043 | Sp = S * (1 + a*tanh(f)), a <= 0.2, init .024 - modulate never inject |
external (omega line) |
| L-044 | p = softmax(-(y-[.5,1.5,2.5])^2/.25); bit_k = p_R + 0.5*p_M; C = sum bit_k 2^-k |
tools/compartment_smoke.py::exact_cantor + soft_cantor_ungated |
| L-045 | L_CM = .01 * relu(eps - vol2(CM)) on first k+1 tokens (validity hinge) |
loss_forms::cm_validity_hinge (NEW) . external KSimplex |
| L-046 | `L_vol = -.005 * std(log | vol^2 |
| L-047 | `L = | |
| L-048 | low = base + .5*mse(HP3(pred),HP3(tgt)); high = base + .5*mse(LP7,..); composed by band windows [role-aligned gauge] |
closeout amoe/diffusion/train/objectives.py:74-80 |
| L-049 | dropout(anchors, p=.3) during alignment |
external (constellation) |
| L-050 | clip_grad_norm .5 on cross-attn ONLY; NEVER inside an LBFGS closure |
protocol |
| L-051 | Adam(params, lr, weight_decay=0.0) - the ONLY constructor |
closeout amoe/laws.py:26-28 |
| L-052 | zeros_(head.weight); zeros_(head.bias); gates = -3.0 |
closeout amoe relay.py:49-52 |
| L-053 | g_tang pass; g_radial *= (1-.01); g_collapse *= 1.0 (gradient gating) |
external (GM3 Form 12; ALEPH-OMITTED) |
| L-054 | q_comp = R (Hamilton) q_expert over 4 FiLM arms |
external (protein line) |
| L-055 | Q = (I-A)(I+A)^-1, A skew - det=1 by construction |
external (protein line) |
| L-056 | k-step MLP cascade in place of one dimensional jump |
structural, external |
| L-057 | w_route = f(phase-0 triangulation distances), softmax-free, geometry-derived |
external (Form 6) |
| L-058 | exact softmax within sorted equal-width same-bucket block; codebook grad via address-agreement bias |
design (Tree 3b pending) |
| L-059 | M_hard + (M_soft - sg[M_soft]) (STE over an ABSOLUTE reconstructive read) |
amoe.core.address / geolip_svae |
| L-060 | n_i_new = min(ceil(norm * n_i^0.5), 8, 1.25*top) (sqrt-damped repeats) |
external (anima bucketing) |
| L-061 | `assert | question_space |
| L-062 | on starvation strike: sampling_weight[starved] *= 2; 3 strikes abort - DATA, never a loss |
closeout amoe/train/aligner.py:68-84 |
| L-063 | with p=.1: cond <- 0 (zeroed, not empty-prompt) |
pod2/dexp006:132-146 |
| L-064 | NOT RECORDED (role-weighted pentachoron regularization; rose_w=1e-4, temp .07) |
external (David classifier) |
| L-065 | ramp(x)=.5-.5cos(pi*(clamp(x/XF,-1,1)+1)/2); low=1-up1; mid=up1(1-up2); high=up1*up2; edges(.35,.75) XF=.06 |
closeout amoe/diffusion/core/multiband.py:23-35 . compartment_smoke::compartment_windows |
| L-066 | L = base + lam * w_route * aux, lam~1 (3-pt dose curve) |
loss_forms::a5_dose_coupled |
| L-067 | reduce in float64, autocast off (gauges) |
loss_forms::a8_fp64_gauge . geolip_vitals |
| L-068 | acc = mean_fp64(res_arm(row,noise,t) - res_ref(row,noise,t)), triples FIXED per row |
loss_forms::a6_paired . dexp001:167-185 |
| L-069 | w = f(pi_frozen_ref); L = sum(w*ce_tok)/sum(w) |
loss_forms::pwa_weights+a2_weighted [gate: 0.014-0.083] |
| L-070 | v = norm(feats)@R^T/t; L = mean(cosh(clamp(v - c_y*mu, -4, 4)) - 1) |
loss_forms::fac_loss . tools/fac_bed.py |
| L-071 | CV = std(V)/mean(V), V = CM 4-volumes over 200 random 5-subsets, fp64 - READOUT |
tools/geolip_vitals.py::pentachoron_cv |
| L-072 | `drift = arccos(<norm(a), norm(a_init)>); binding_frac = mean( | drift-.29154 |
| L-073 | bpb = mean CE / ln(2) per byte |
beds (ar_differentiation, fac_bed) |
| L-074 | `tax = exp(mean CE_512) | adapted - exp(mean CE_512) |
| L-075 | F1 = 2PR/(P+R) over token multisets vs GT captions |
pod v35 judges |
| L-076 | `precision = | pred inter GT |
| L-077 | sep_L = mean_ij inter-register Ham(code_i,code_j) - mean intra (diagonal KEPT, +4% bias, comparability) |
pod2/aleph_diffusion_core.py:257-290 . canon/register_probe_gauge.md |
| L-078 | `Spearman(Hamming(c_i,c_j), arccos | <a_i,a_j> |
| L-079 | HIGH-band foreground-masked LP-x0 error (fp32 judged) |
closeout amoe/diagnostics/diffusion.py::foreground_gauge |
| L-080 | assert torch.equal(logits_all_off, logits_base) |
closeout amoe/testing/invariants.py:76-97 |
| L-081 | ratio = damage(own band lesion) / damage(cross band lesion) per gauge |
closeout amoe/diagnostics/diffusion.py::lesion_report |
| L-082 | excess = metric(real keys) - metric(SAME key repeated); + matched-vs-mismatched delta |
pod2/dexp015_router_v2.py:159-191 |
| L-083 | usage = mean oriented-softmax row; ppl = exp(H(usage)); alive = usage > eps/2K |
tools/geolip_vitals.py::axis_aliveness |
| L-084 | `read ppl = exp(H(mean read weights)); commitment = | cos(read, nearest atom) |
| L-085 | `ratio = mean | delta |
| L-086 | exact-match on two-step composite prompts vs single-step controls |
pod v35 composite judges |
| L-087 | `effect = mean | loss_off - loss_on |
| L-088 | `orders = log10(max group | |
| L-089 | CV at step 1000 -> band {<.30 LOW / .35-.50 MID / >.80 HIGH} + stability + freeze-survival |
tools/geolip_vitals.py::CVScreen |
| L-090 | `ripser H2 on d(a,b)=arccos | <a,b> |
| L-091 | S0/S_D spectral ratio; erank = exp(-sum p ln p), p = sigma/sum sigma |
external (structural attractors) |
| L-092 | var(recon MSE) across patch grids 81..4096 - flatness IS the pass |
external (omega line) |
| L-093 | guarded subprocess: restricted builtins, len cap, timeout, no net; exact-match out |
pod v35 exec judge (protocol) |
| L-094 | json.parse validity + IoU(xywh) + pair-order + termination-within-window |
pod v35 judges |
| L-095 | torch.cuda.max_memory_allocated + s/step at an early step (WDDM spill tell) |
telemetry; pod/run_g6b_rider.py:34-48 |
| L-096 | drift_g = arccos(<consensus_g, consensus_prev>); stationarity = no acceleration |
tools (dif line) |
| L-097 | held-out byte acc under substitution cipher; variant-format recall |
tools/exp019/exp020 |
| L-098 | NN symbol-Hamming between stored and recomputed keys; match@theta=.25 |
tools/exp021_persistent_store.py |
| L-099 | mean cos(book_epoch, book_init) across the bank sweep |
external (192-bank sweep) |
| L-100 | ` | V_fast - V_geovocab2 |
| L-101 | gate_mean = mean sigmoid(g); band [.012,.03] ADVISORY |
tools/geolip_vitals.py::gate_stats |
| L-102 | mean mse over all sigma - BLIND to band structure (distrusted for bands) |
prim_sq+a0 (the point is what it hides) |
| L-103 | cos(recon, x) - WRONG instrument for addressed systems (address = lookup key) |
distrusted; do not implement as a judge |
| L-104 | recall@k under Hamming(code_query, code_bank) |
tools/qwen_exp011 (q25 line) |
| L-105 | ` | |
| L-106 | selection event = argmax/softmax-over-roster in the compute path |
prohibited class (the EVENT is the violation, not one formula) |
| L-107 | alphabet learned by task gradient (vs fitted-frozen) |
prohibited class (dif line evidence) |
| L-108 | direct task-gradient descent on pentachoron vertices |
prohibited (geovocab evidence) |
| L-109 | gap = x.mean(dim=spatial) in a geometric encoder |
loss_forms::forbidden_gap (NEW, force-gated control) |
| L-110 | CV term injected in backward, or weight > 1e-3 |
prohibited placement of L-040 |
| L-111 | same as L-047 - AS THE TRAINING FORCE (placement retracted) |
loss_forms::procrustes_sq [placement warning] |
| L-112 | logits = W u, u = single-slot coefficients at hard tau (K=32) |
tools/ar_differentiation_bed.py (addr_head) . fac_bed p4 cells |
| L-113 | a7_grid_infonce INTO an address path |
loss_forms::a7 [placement ban in docstring] |
| L-114 | prim_kl at alpha=1.0 from near-parity teachers in a selection loop |
loss_forms::prim_kl+kd_guard [guard warns] |
| L-115 | L-016 with x0h = (x_t - sqrt(1-abar)eps_hat)/sqrt(abar) - divides by vanishing sqrt(abar) |
closeout amoe trainer [REFUSES unless force_blob_on_eps] |
| L-116 | keep-or-kill on 1-epoch MSE rank |
retracted protocol; replaced by L-089 |
| L-117 | logits = tied(M_hat) with U=M_hat, S=Omega, Vt=I |
dif-e018 bed (tools) |
| L-118 | route experts by frozen text keys (raw/pooled/M-hat-slot) vs repeated-key null |
pod2/dexp014/dexp015 beds |
| L-119 | a trainable anchor trained to orchestrate others (composite prereg >= +.15) |
pod v35_exp010/exp020 beds (refuted) |
| L-120 | solo specialist stack attached always-on |
configuration class (v35 e018 evidence) |
| L-121 | frozen solo-trained experts composed under dispatch |
pod v35_exp014 beds (demoted) |
| L-122 | transplant proj+codebook onto a fresh trunk |
NA protocol (dif line, demoted) |
| L-123 | L-048 judged by aggregate eps-MSE |
same impl; the GAUGE is the difference |
| L-124 | append frozen byte-trigram address beside full text cond |
pod2/dexp002 bed |
| L-125 | null = shuffle keys across rows (measures diversity, not correctness) |
retracted instrument; use L-082 |
| L-126 | CE on generated question sets with space < 3x draws |
retracted claim; guard = L-061 |
| L-127 | cross-seed comparison across DIFFERENT instruments |
NA (claim-level error) |
| L-128 | bands nested within bands on a Cantor axis |
prohibited configuration |
| L-129 | re-enter measure space per layer/step |
prohibited configuration |
| L-130 | soft_cantor_ungated(x) used as a band COORDINATE (non-monotone) |
tools/compartment_smoke.py::soft_cantor_ungated (as the regression) |
| L-131 | repeat count = ceil(max_bucket / n_i) (alpha=0 equalize-to-largest) |
NA data policy (prohibited) |
| L-132 | conv re-weighted by convex sum a_k=1 over a filter bank (hull-bounded mean) |
geolip-aleph-classification (cautionary control, cls line) |
| L-133 | argmax decoding in an iterative denoiser |
NA decoding policy |
| L-134 | aux = alpha * N * sum_i f_i * P_i (switch-style balance) |
loss_forms::forbidden_load_balancing (NEW, force-gated control) |
| L-135 | CE on the big-JSON composite format |
dropped by operator ruling (never built) |
| L-136 | learned SVD rotation transform in the dual-penta head |
retracted (convergence failure) |
| L-137 | coefficients->logits at ONE hard tau, any dim |
class of L-112; cure = L-012/L-013 |
| L-138 | L-070 on the addr_head collapse configuration |
tools/fac_bed.py (p4_addr_head_fac) |
| L-139 | sparsemax_loss on addr_head logits (K=32, hard tau) - only the loss differs from the certified collapse cell |
tools/deviant_bed.py (p4_sparsemax) . loss_forms::sparsemax_loss |
| L-140 | L = -z_y + 0.5*sum_{j in S}(z_j^2 - tau^2) + 0.5 (sparse support S) |
loss_forms::sparsemax_loss . tools/deviant_bed.py |
| L-141 | L = T*logsumexp(ce_tok/T) - T*log(N), T=0.5 |
loss_forms::dev_softmax_accum . tools/deviant_bed.py |
| L-142 | L = mean(log(ce_tok + 1e-3)) |
loss_forms::dev_geomean_accum . tools/deviant_bed.py |
| L-143 | CE to (1-eps) smoothed targets == (1-eps)CE + eps*uniform-KL |
F.cross_entropy(label_smoothing=.1) . tools/deviant_bed.py |
| L-144 | L = sum((1-p_y)^2 * ce_tok) / sum((1-p_y)^2), p_y detached from the live model |
tools/deviant_bed.py (focal) |
| L-145 | L = sum(ce_tok * [pi_ref > .6]) / count, pi_ref from the frozen ce_s0 checkpoint |
tools/deviant_bed.py (anti_curr) |
| L-146 | L = mean(1 - tanh(v) * c) |
loss_forms::fac_loss_link(link='tanh') . tools/deviant_bed.py |
| L-147 | L = mean(log(1 + (v - c*mu)^2)) |
loss_forms::fac_loss_link(link='cauchy') . tools/deviant_bed.py |
| L-148 | L = CE - 0.1*H(p) |
tools/deviant_bed.py (conf_penalty_CONTROL) |
| L-149 | CE over logits = cos(normalize(feats), normalize(A_c)) * 10 |
tools/geobasin_bed.py (geo_ce_scores) |
| L-150 | attraction (1-s_y)^2 + 0.5*sum_{c!=y}s_c^2 + 0.5*relu(max_{c!=y}s_c - s_y + .3) + 0.1*range |
geofractal.losses.pure_geometric (VERBATIM) . tools/geobasin_bed.py |
| L-151 | same loss; A registered as a frozen buffer |
tools/geobasin_bed.py (geo_pure_frozen) |
| L-152 | L = (1 - s_y)^2 + 0.1*(relu(s-1)^2 + relu(-s)^2) - no other-class terms at all |
tools/geobasin_bed.py (geo_pure_norep) |
| L-153 | cos(proj(scores), prototypes) pulled/pushed + prototype-diversity term |
geofractal.losses.pure_geometric::GeometricPrototypeLoss (VERBATIM) |
| L-154 | coarse (superclass sums to target) + fine + consistency, sigmoid-weighted |
geofractal.losses.pure_geometric::HierarchicalGeometricLoss (VERBATIM, FastHierarchical parity wrapper) |
| L-155 | 0.5*CE(cos*10) + 0.5*PureGeometricLoss(scores) |
tools/geobasin_bed.py (geo_hybrid) |
| L-156 | mean((z_s - z_t)^2), targets L2-normalized |
tools/dist_bed.py loss_of feature_mse (loss_forms.prim_sq) |
| L-157 | 0.5*[CE(sim/t, diag) + CE(sim.T/t, diag)], sim = z_s @ z_t.T |
tools/dist_bed.py loss_of infonce (loss_forms.a7_grid_infonce) |
| L-158 | mean(BCE(sim*t' + b, I)) per pair, no batch denominator |
tools/loss_forms.py siglip_pairwise_loss |
| L-159 | `KL(softmax(sim_t/t) | |
| L-160 | `a7_grid_infonce + 0.3procrustes_sq(fp64 R) + 1e-3 | CV16-0.20 |
| L-161 | mean((z_s - gpa(z_t1..t5))^2), GPA cached iter-5 align .9401 |
tools/dist_bed.py loss_of consensus_gpa (target build_gpa) |
| L-162 | a7_grid_infonce(zs, zt_gpa, t=.07) + prim_sq(zs, zt_gpa).sum(-1).mean() |
tools/dist_bed.py loss_of consensus_nce_mse |
| L-163 | `1.0expert_agreement + 1.0rotation_ortho + 0.5anchor_spread + 0.1anchor_entropy + 0.3cross_expert_var + 0.3 | bank_cv-t |
| L-164 | a7_grid_infonce + 0.5*prim_sq.mean() + 0.3*BCE(frozen_head) + 0.5*centered_cos + 1e-3*x3_cv + EA(tang .01, sep 1.0) |
tools/dist_bed.py loss_of x3_full |
| L-165 | 0.3 * BCE_logits(frozen_linear80(zs), labels) |
tools/dist_bed.py loss_of x3_bce + _fit_teacher_head |
| L-166 | backward: g <- (g - radial) + tang*radial; g <- g - sep*(g.nearest_anchor)_+ ; forward identity |
tools/loss_forms.py EmbeddingAutograd |
| L-167 | `1e-3 * | CV_16sets(emb) - cv_target |
| L-168 | a7_grid_infonce + prim_sq.mean() + 0.1*x3_cv_loss; AdamW(3e-4, wd .01) + LinearLR(1000) + cosine(1e-6) + clip 1.0 |
tools/dist_bed.py loss_of cbert_full + train_arm sched branch |
| L-169 | a7_grid_infonce(zs, zt_ref, t=.07) + prim_sq(zs, zt_ref).sum(-1).mean(); zt_ref = normalized centroid of teachers whitened-Procrustes-mapped into a REFERENCE MEMBER's frame |
tools/captionbert_8192_v2.py (stage 1 maps + stage 3 loss) |
PRIMARY OBJECTIVES
Terms that were the training signal, not an addition to one.
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-001 | 10 | 33222+1 | SQ.A0/bytes |
MSE -> bitwise reconstruction (SVAE H2, 16 noise types) | 16 noise types converge simultaneously; bitwise text recon; the two-year survivor | MANIFEST.md:133-139; MANIFEST.md:144-145 |
| L-002 | 10 | 33222+1 | SQ.A0/codebook |
recon gradient through M-hat (the aleph's ONLY codebook pressure) | cos .992-.997 hard-mode, 125-126/128 axes alive, ZERO collapse, div_weight=0 | canon/aleph_core.md:16-18; canon/discovery_catalog.md:40-44 |
| L-162 | 10⟂ | 33222+0 | CE.A7/consensus-embedding |
consensus NCE+MSE composite on a MEAN consensus (CaptionBert Stage-A form, vision; rotation required) | THE CROWN AT SCALE: consensus composite + rotation beats direct deployment-teacher distillation on EVERY task gauge 2/2 (CC12M 11x data); local COCO verdict was BETWEEN - scale promoted it | tools/dist_runs/consensus_nce_mse_s0.jsonl; tools/dist_runs/vitbank_s0.jsonl |
| L-169 | 10⟂ | 33222+0 | CE.A7/consensus-embedding |
consensus NCE+MSE composite on a REFERENCE-MEMBER consensus (captionbert-8192-v2/-b, text; NO rotation) | THE SIBLING OF L-162 AND THE CHEAPER DEPLOYMENT: same objective, but the consensus is aligned to a REFERENCE MEMBER (bert-base) instead of a running mean, so the frame gap collapses monotonically to +.0010 and NO rotation ships. Student beats its best teacher (.6077 vs distilbert .5883) at 13% of the teachers' combined params; replicated on a second trunk (-b .6031, a null on 19% more data) | repos/captionbert-v2.md; repos/captionbert-b.md |
| L-012 | 9 | 33212+0 | CE.A2/coefficients |
addr_msl slot-parallel read (P parallel D=4 slots, shared K=64) | THE ACCUMULATION CURE: 5.6650 (collapsed) -> 2.47 with the primitive held FIXED | history/timeline.md:124-138; history/timeline.md:1098-1104 |
| L-015 | 9 | 33212+0 | CE.A4/logits |
derived-steps expert supervision (stepwise-CoT target vs direct target) | +0.79 vs direct +0.63; held-out ceiling 1.00/1.00; seeds matched to 4 decimals (+0.7917 / +0.7916) | inventory/SCALE_MANIFEST.md:144-154; history/timeline.md:1640 |
| L-157 | 9 | 33212+0 | CE.A7/similarity-grid |
symmetric InfoNCE distillation (student<->teacher, t=0.07) | R@1 .0580/.0544 = 2.2x feature_mse 2/2; zs-C100 .0939/.0847 wins 2/2 (+3.2/+2.2pt); erank 184; cv16 lands the S^15 band NATIVELY | tools/dist_runs/; canon/clip_mini_distillation.md |
| L-004 | 9 | 23122+1 | CE.A1/logits |
chunked masked CE (512-token slices, sum-then-renormalize) | 22.8GB dedicated + 42.8GB SILENTLY SHARED -> 8.8GB peak @ 1.03 s/step | MANIFEST.md:50-59; pod/v35_exp003_termination.py:128-138 |
| L-138 | 9⟂ | 23222+0 | BREG.A0/coefficients |
FAC on the partition-collapse configuration (the P4 loss-swap cell) | the certified addr_head collapse DECOMPRESSES under a loss swap alone, 3/3 seeds: usage ppl 1.0-2.7 -> 60.6-61.1 of 64; decoded acc 0.05-0.20 -> 0.45-0.47; the win | cos |
| L-139 | 9 | 23222+0 | CE.A0/logits |
sparsemax on the collapse configuration (the coupling-axis probe) | THE DOSE-RESPONSE: usage 1.85 (CE, full coupling) -> 23.5 (sparsemax, partial) -> 60.9 (FAC, none); win | cos |
| L-152 | 9 | 23222+0 | SQ.A0/basin-scores |
PureGeometric ABSOLUTE-ONLY (attraction + range; comparative terms dropped) | MORE THAN DOUBLES the full set: acc 0.349 vs 0.157, bpb 3.75 vs 7.52, 3/3 seeds - the comparative terms are the poison | tools/geobasin_runs/; history/timeline.md 2026-07-25 geobasin block |
| L-010 | 9 | 32212+1 | SQ.A0/x0-linear |
flow v-MSE (rectified flow, SHIFT-warped sigma) | x0 = x_t - sigma*v is EXACT and LINEAR at every sigma - asserted, not assumed | pod2/dexp001_sd15_relay.py:151-164; closeout_2026-07-19/amoe/src/amoe/diffusion/train/objectives.py:51-58 |
| L-003 | 8⚠ | 33121+0 | CE.A0/logits |
plain full-sequence cross-entropy (packed labels) | the workhorse; also the documented geometry antagonist - CE drove the Oct '25 geometric collapse | pod/v35_exp001_placement.py:170; tools/ar_differentiation_bed.py:423 |
| L-017 | 8⟂ | 33111+0 | CE.A7/similarity-grid |
InfoNCE as an alignment force (OFF address paths) | NECESSARY + SUFFICIENT for alignment: swap it in -> R@1 .999 | canon/geometric_memory_blueprint.md:7-8; canon/geometric_memory_blueprint.md:20 |
| L-005 | 8 | 32121+0 | CE.A4/logits |
shift-CE with ignore_index=-100 (prefix-masked instruction rows) | the standard instruction-tuning form across the v35 and q25 lines | pod/v35_exp006_math.py:180; pod/v35_exp013_experts.py:211 |
| L-009 | 8 | 32211+0 | SQ.A0/eps |
eps-MSE (epsilon prediction, stock schedule) | relay -2.5% over frozen, 2 seeds; relay >= matched LoRA 2-for-2 across substrates | pod2/dexp006_sd15core_relay.py:132-146; history/timeline.md:1337-1350 |
| L-011 | 8 | 32122+0 | CE.A0/logits |
sign-code head addr_mslh64 (fully discrete forward, STE backward) | bpb 2.4711 vs soft 2.4685 - parity certified 3 seeds; a ~2.8% gap opens at 4x budget | history/timeline.md:249-283; inventory/SCALE_MANIFEST.md:64-71 |
| L-159 | 8 | 32221+0 | KL.A7/similarity-grid |
affinity-KL distillation (TinyCLIP row-softmax mimicking) | FRAME AMBIGUITY PROVEN 2/2: agree_cos -0.021/-0.002 (orthogonal to teacher frame, zs floor .109/.111, erank 175 healthy - the tell) yet ONE fitted rotation recovers agree .501/.497 and zs .240/.218 - geometry learned, gauge free, both seeds | tools/dist_runs/; canon/clip_mini_distillation.md |
| L-161 | 8 | 32221+0 | SQ.A0/consensus-embedding |
consensus-GPA distillation (MSE to projective-ICP/GPA mean of 5 CLIP towers) | P7 REFUTED at the letter (raw zs .154/.161 loses to single-teacher 2/2) BUT post-rotation agree->LAION .7132/.7116 EQUALS feature_mse's in-frame .714/.718 both seeds: the GPA mean shape carries the geometry, loses only the gauge | tools/dist_runs/; canon/clip_mini_distillation.md |
| L-168 | 8 | 32212+0 | CE.A7/consensus-embedding |
CaptionBert Stage-A VERBATIM replication on the vision student (elem-MSE 1.0 + CV 0.1 + AdamW/warmup/cosine/clip) | THE REPLICATION ANCHOR 2/2 (prototype-replication law's first yield): the faithful recipe is WORSE than the adaptation on nearly every gauge — agree .4113/.4120 vs .672/.670, zs_rot .213/.221 vs .232/.264, probe .390/.386 vs .401/.394 — so the bank refutation was NOT adaptation damage; the adaptation is the stronger Stage-A | tools/dist_runs/x3_queue.log; tools/dist_runs/vitbank_s0.jsonl |
| L-149 | 8 | 22222+0 | CE.A0/basin-scores |
CE on the cosine-anchor basin head (the geobasin control) | the head itself costs +0.13 bpb under CE (2.607 vs 2.477 linear, 3 seeds; acc .498 vs .505) - small, so every geometric-arm deficit is THE LOSS, isolated | tools/geobasin_runs/; history/timeline.md 2026-07-25 geobasin block |
| L-156 | 8 | 31222+0 | SQ.A0/teacher-embedding |
feature_mse distillation (MSE to LAION-B/16 embedding) | zs-C10 .2693/.2657 (matrix best) and agree_cos .714/.718 - but C100 .062 LOSES to infonce (+3.2pt) and R@1 .025 is 2.2x below, 2/2 seeds | tools/dist_runs/; canon/clip_mini_distillation.md |
| L-158 | 8 | 31221+0 | CE.A0/similarity-grid |
SigLIP pairwise sigmoid distillation (decoupled contrastive) | zs SPLIT (-4.6pt s0 / +2.3pt s1 vs infonce) but LOSES retrieval 2/2 (R@1 .043/.040 vs .058/.054): coupling wins retrieval at batch 256 | tools/dist_runs/; canon/clip_mini_distillation.md |
| L-028 | 7 | 23111+0 | CE.A4/logits |
masked-marginal variant scoring (protein VEP) | WT unmasked marginal rho 0.10 -> masked marginal ESSENTIAL; final rho .993 / .309 unseen | canon/ryan_spearman_protein.md:41-42; canon/ryan_spearman_protein.md:6-8 |
| L-006 | 7 | 32111+0 | CE.A4/logits+pixels |
HuggingFace out.loss (VLM labels= path, vision tower fires) | required wherever the vision tower must fire - chunking bypasses it | pod/v35_exp004_caption.py:175; pod/v35_exp004_caption.py:17-19 |
| L-029 | 7 | 22211+0 | CE.A0/logits |
GPT-2 frozen-trunk relay objective (dif-e013 Track C) | frozen 38.648 -> aleph 26.53 vs param-matched zero-init MLP 27.26; beats matched 2/2 seeds | history/timeline.md:349-363 |
| L-164 | 7 | 22112+0 | CE.A7/consensus-embedding |
x3 full stack on the vision student (NCE + elem-MSE .5 + BCE .3 + centered-cos align .5 + CV 1e-3 + EmbeddingAutograd) | WEAK-ANCHOR SET-POINT 2/2: agreement lands at .4957/.4958 (cross-seed 4e-4) — .17 BELOW the anchored composite — while task gauges stay in-noise and probe drops 1.1pt. Its faster NCE descent (~8%) is UNANCHORED-optimizer speed: the inert elem-MSE leaves the contrastive term unopposed (diag: base agree .593 at step500 vs .432; x3diag_*.jsonl) | tools/dist_runs/x3_queue.log; tools/dist_runs/vitbank_s0.jsonl |
| L-013 | 6 | 13212+0 | CE.A2/coefficients |
addr_3tau multi-tau stroboscope | 4.2884 no collapse (usage ppl 7.9, 117/128 alive) against addr_d4's 5.3698 | history/timeline.md:1093-1095 |
| L-027 | 6 | 13212+0 | CE.A0/logits |
antipode-conv objective (the address AS the convolution operator) | CIFAR-10 87.23% @ 861,450 params with NO ReLU/GELU anywhere; none -> mag +21.8 | inventory/SCALE_MANIFEST.md:8-30; canon/discovery_catalog.md:49-61 |
| L-037 | 6 | 13111+0 | CE.A0/logits |
denoiser objective (tokendiff iterative image-token denoise) | beats identity at every level; t=1.0 gives 0.378 vs 0.002 (189x) | sessions/2026-07-13.md:35-43 |
| L-018 | 6 | 22111+0 | CE.A7/mixed |
blueprint composite (InfoNCE 1.0 + Procrustes_SVD 0.3 + | CV-0.20 | 0.05) |
| L-014 | 6 | 12212+0 | CE.A0/coefficients |
addr_mhat reconstructive read consumed in AR | 5.1300 bpb but the HEALTHIEST cultivation on the bed (ppl 11.0, binding_frac .234) | history/timeline.md:1095-1097 |
| L-035 | 6 | 12211+0 | SQ.A0/x0-linear |
rectified-flow velocity objective (KSimplex / Form 7 bottleneck) | loss .1749 beat the 268M skip's .1757; the model routed 88% through the 768 dims | canon/ksimplex_diffusion_prior.md:26-28; canon/constellation_forms.md:59-63 |
| L-038 | 6 | 12211+0 | SQ.A0/embeddings |
recon_target (absolute MSE to a fixed frozen-trunk projection) | recall@1 0.264 - real (5x frozen) but HALF of InfoNCE's 0.494 at matched budget | sessions/2026-07-13.md:122-126; tools/qwen_exp011_newanchors.py:445-453 |
| L-020 | 5⚠ | 12111+0 | SQ.A0/sequence-embeddings |
SequenceReconstructor loss: MSE(normed) + (1 - cos) | CLIP-L ep5 m_acc .957 / s_cos .734; Meridian bigG s_cos PLATEAUS at .425 | canon/geometric_memory_blueprint.md:56-63; canon/constellation_forms.md:44-48 |
| L-146 | 5 | 21222+0 | BREG.A0/address-code |
FAC tanh-Hamming link (bounded tails) | cosh beats tanh 3/3: 4.349 vs 4.1285 (+0.22) - the bounded link loses within the family | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-147 | 5 | 21222+0 | BREG.A0/address-code |
FAC Cauchy link (sub-quadratic tails) | cosh beats Cauchy 3/3: 4.360 vs 4.1285 (+0.23) - robust-statistics tails lose within the family | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-007 | 5 | 11221+0 | CE.A0/logits |
dispatch-keys-only CE (aligner; adapters frozen as anchors) | trainable set is ONLY the per-block key matrices; reference-grade, never seed-replicated | closeout_2026-07-19/amoe/src/amoe/train/aligner.py:41-67 |
| L-030 | 5 | 20211+0 | CE.A0/logits |
val_ce on a frozen semantic substrate (CLIP-L token-AR) | MLP WINS frozen-substrate token-AR (penult 5.245 best); aleph tax ~ +0.09 | history/timeline.md:335-348 |
| L-031 | 5 | 20122+0 | SQ.A0/geometry |
pure geometric-basin loss set (coherence/separation/discretization/geometry) | the program's ONE attempted CE replacement - NOW PROPERLY TESTED: refuted on the byte bed (acc 0.157 vs control 0.498, 3 seeds); the absolute-only variant doubles it (L-152) | canon/cantor_empirics.md:20-22 |
| L-070 | 5⟂ | 20221+0 | BREG.A0/address-code |
FAC as a PRIMARY sequential objective (cosh-Bregman, replace CE) | REFUTED AS PREREGISTERED, 3/3 seeds: fac_lsh 4.13 bpb vs ce 2.48; ce_fixedcode 3.81 beats it; fac_none 3.95 beats it | tools/fac_runs/; history/timeline.md 2026-07-25 pass-3 block |
| L-140 | 5 | 20221+0 | CE.A0/logits |
sparsemax as a full-bed objective | REFUTED as a general objective: bpb 7.43 / acc 0.331 vs ce 2.4769 / 0.505 (3 seeds) | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-141 | 5 | 20222+0 | CE.A9/logits |
soft-max / worst-position accumulation (trained) | REFUTED: bpb 4.24 / acc 0.276, 3 seeds - the 0.911-novelty champion chases irreducible entropy exactly as flagged | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-142 | 5 | 20222+0 | CE.A2/logits |
geometric-mean accumulation (trained) | REFUTED decisively: bpb 9.03 - the anti-focal starves hard positions and the distribution never forms (3 seeds) | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-143 | 5 | 20122+0 | CE.A0/logits |
label smoothing eps=.1 (trained on the byte bed) | bpb 2.587 vs ce 2.4769 (+0.11, 3 seeds) - payer-class novelty (0.479), mildly WORSE outcome | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-144 | 5 | 20122+0 | CE.A2/logits |
focal gamma=2 (trained, live-model weights) | bpb 2.597 (+0.12 vs ce, 3 seeds) - payer-class trained novelty (0.337), mildly worse outcome | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-145 | 5 | 20222+0 | CE.A4/logits |
anti-curriculum (train only where the frozen reference is confident) | REFUTED as an objective: bpb 6.74 (3 seeds) - abandoning 72% of the distribution buys nothing on the rest | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-148 | 5 | 20122+0 | CE.A0/logits |
confidence penalty (trained as the GATE-VALIDATION CONTROL) | CLOSEST TO CE OF ALL DEVIANTS: bpb 2.520 (+0.043, 3 seeds) - the gate's refusal correctly predicted 'CE plus nothing' | tools/deviant_runs/; history/timeline.md 2026-07-25 deviant-matrix block |
| L-150 | 5 | 20122+0 | SQ.A0/basin-scores |
PureGeometricLoss, learned anchors (the Oct '25 arm, properly tested) | REFUTED on this substrate: acc 0.157 vs control 0.498 (-68% relative, 3 seeds) - far below the historical -12% trade | tools/geobasin_runs/; history/timeline.md 2026-07-25 geobasin block |
| L-151 | 5 | 20222+0 | SQ.A0/basin-scores |
PureGeometricLoss, FROZEN anchors (the L-108 cell) | learned BEATS frozen by +8 acc points (0.157 vs 0.076, 3 seeds) - the L-108 falsifier FIRED for cosine anchors | tools/geobasin_runs/; history/timeline.md 2026-07-25 geobasin block |
| L-153 | 5 | 20112+0 | SQ.A0/basin-scores |
GeometricPrototypeLoss (verbatim, own projector) | WORST of the family: bpb 8.12, acc 0.008 (3 seeds) - the extra indirection buys total failure | tools/geobasin_runs/; history/timeline.md 2026-07-25 geobasin block |
| L-154 | 5 | 20122+0 | SQ.A0/basin-scores |
HierarchicalGeometricLoss on the nibble hierarchy (16x16) | CATASTROPHIC: acc 0.0003 - below chance (1/256) - hierarchy-in-class-space destroyed fine structure entirely (3 seeds) | tools/geobasin_runs/; history/timeline.md 2026-07-25 geobasin block |
| L-155 | 5 | 20222+0 | SQ.A0/basin-scores |
CE + PureGeometric hybrid (0.5/0.5) | the geometric set POISONS CE rather than riding it: bpb 4.53 vs control 2.61 (+1.9, 3 seeds) - P4 bar (within 0.15) missed by 12x | tools/geobasin_runs/; history/timeline.md 2026-07-25 geobasin block |
| L-160 | 5 | 20112+0 | CE.A7/similarity-grid |
blueprint composite on image distillation (infonce + 0.3*procrustes + 1e-3 CV band) | INERT vs plain infonce 2/2: zs .248/.253 vs .267/.231, R@1 .057/.051 vs .058/.054, erank 184.6 vs 184.5, cv16 .232/.239 vs .221/.245 - every delta within seed noise | tools/dist_runs/; canon/clip_mini_distillation.md |
| L-036 | 4 | 11111+0 | CE.A0/logits |
margin losses ArcFace / CosFace / SphereFace (RoseFace dual-norm) | the ZANA innovation - and it hit a 60% single-stream ceiling | canon/vit_lineage_beatrix_zana.md:17-28 |
| L-008 | 4 | 10121+0 | CE.A0/logits |
image-classification CE (CIFAR-10, aleph-dispatched MoE vs dense) | MoE 58.52% TIES param-matched dense 58.52% exactly; 6x params bought nothing | history/timeline.md:1127-1138; tools/addressed_distillation_v1.py:301 |
| L-032 | 3 | 00220+0 | -.NA/geometry |
GBC - 'cross-entropy can be replaced entirely' (roadmap claim) | classification via triadic compatibility, self-similarity, Cantor coherence, hierarchical basin checks | canon/geovocab_chunking.md:21-22 |
| L-033 | 3 | 00212+0 | SQ.A0/pixels |
masked-recon / generative arm (campaign law 2 in its ORIGINAL form) | BUILT, NEVER RUN - predicted to be where the SIGNED read finally beats magnitude | inventory/SCALE_MANIFEST.md:19-21; history/open_questions.md:144-145 |
AUXILIARY TERMS, REGULARIZERS, AND STRUCTURAL SUBSTITUTES
Includes the structural devices that REPLACED a loss term - sphere normalization, gradient equalization, bounded alpha. Their presence here is the point: this program's strongest results come from making geometry structural rather than penalized.
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-052 | 10 | 33221+1 | -.NA/adapters |
zero-init output heads (WEIGHT and bias) - the inertness contract | makes the toggle law bit-exact (max | dlogit |
| L-065 | 10 | 33212+1 | -.A3/sigma-axis |
band crossfade windows as STRUCTURAL positional gating | own-band damage 50-200x cross-band, 3/3 both seeds - specialists manufactured with NO routing loss | MANIFEST.md:90-96; plans/2026-07-17_multiband_stage_plan.md:18-35 |
| L-044 | 10 | 23222+1 | -.NA/cantor-measure |
Devil's Staircase alpha-normalization (bit_k = p[RIGHT] + alpha*p[MIDDLE], alpha=0.5) | WITHOUT the alpha term the measure COLLAPSES to {0, .333, .667} - the FractalDavid bug | MANIFEST.md:30-31; canon/vit_lineage_beatrix_zana.md:40-49 |
| L-061 | 10 | 23222+1 | -.NA/data |
question-space guard (training-question space >= 3x draws) | caught TWO memorized experts that had PASSED the answer-diversity guard (spaces 480 and 248 vs 800 draws) | MANIFEST.md:169-170; inventory/SCALE_MANIFEST.md:152-154 |
| L-051 | 10 | 32222+1 | -.NA/optimizer |
pure Adam, weight_decay = 0 (the anti-regularizer law) | Adam+gates .731 vs AdamW(3e-4, wd .01) .667 - 'weight decay is uniform damping that destroys the geometric harmonic' | MANIFEST.md:15; canon/discovery_catalog.md:18 |
| L-068 | 9 | 33212+0 | -.A6/any |
paired (row, noise, t) triples - the variance-killing accumulation | the noise-pair floor is ~0.988 - without pairing the effects this program measures are invisible | history/timeline.md:1220-1227; pod2/dexp001_sd15_relay.py:167-185 |
| L-042 | 9 | 23212+1 | -.NA/gradients |
gradient equalization across heterogeneous geometric towers | without it spreads hit 20 ORDERS of magnitude (fibonacci dead at 2.25e-21 under helix) | MANIFEST.md:43-45; canon/fibonacci_systems.md:14-20 |
| L-016 | 9 | 32212+1 | SQ.A5/x0-linear |
blob-LP-x0 structural supervision on FLOW (lambda ~ 1) | -5.9% / -3.7% two seeds on flow vs +0.03% / -1.0% on eps: a ~125-200x effect ratio | MANIFEST.md:65-69; history/timeline.md:1481-1491 |
| L-041 | 9 | 32222+0 | -.NA/activations |
sphere normalization (M = F.normalize(M); | M | |
| L-043 | 9 | 32221+1 | -.NA/spectra |
bounded multiplicative alpha (S*(1 + alpha*tanh), alpha <= 0.2, init .024) | unbounded alpha POISONS the spectrum; bounded modulation costs 2,272 of 16.9M params (0.013%) | MANIFEST.md:40; canon/omega_tokens_fresnel_johanna.md:5-9 |
| L-050 | 9 | 22222+1 | -.NA/gradients |
gradient clipping discipline (0.5 on cross-attn ONLY; NEVER inside an LBFGS closure) | unclipped LBFGS closure DIVERGED to G-MSE 7.4e26; safety is line_search_fn='strong_wolfe' | MANIFEST.md:46-47; canon/tri_band_omega_arc.md:55-57 |
| L-048 | 8⟂ | 32212+0 | SQ.A3/band-routed |
HP/LP band-role objectives [judged by the ROLE-ALIGNED gauge] | multiband beats the matched monolith ~10% on HIGH-band foreground, BOTH seeds | history/timeline.md:1465-1479 |
| L-059 | 8 | 32122+0 | -.NA/codebook |
straight-through estimator on the aleph HARD read | forward fully DISCRETE oriented code, backward soft: hosted books hold cos .992-.997, 112-122/128 hard axes, zero collapse | canon/aleph_core.md:20-23; MANIFEST.md:135-138 |
| L-166 | 8 | 32122+0 | SQ.NA/gradient-path |
EmbeddingAutograd gradient operator (radial damping + nearest-anchor separation), isolated | THE SEED-LOTTERY COLLAPSER (candidate, 2 seeds): pins rotated zs at .2539/.2534 (cross-seed 5e-4) where the baseline swings .232-.264 — means unchanged, VARIANCE collapsed. Mildly retards its own descent (~.03 EMA, it deletes gradient components) | tools/dist_runs/x3_queue.log; tools/dist_runs/vitbank_s0.jsonl |
| L-040 | 8 | 22221+0 | DET.A2/S15-bank |
1e-3 CV bank loss (arm-gated, S^15 bank ONLY, never the aleph codebook) | holds CV .295-.305 at zero-to-positive task cost where bare CE drifts it to .31-.34 | tools/exp017_aleph_constellation.py:154-186; history/timeline.md:692-696 |
| L-066 | 8 | 22212+1 | -.A5/x0-linear |
lambda dose coupling (3-point curve on the blob term) | 0.5 -> -5.9% | 1.0 -> -8.3% (in bound) |
| L-067 | 8 | 22212+1 | -.A8/gauges |
fp64 gauge accumulation (autocast disabled in the reduction) | fp32 determinants lose up to ~4% on near-degenerate pentachora - 'fp32 det only' now means fp32 MINIMUM | MANIFEST.md:32-39; tools/geolip_vitals.py:60-103 |
| L-023 | 7 | 23111+0 | KL.A4/logits |
kd_facts (fact rows supervised ONLY by teacher logits, alpha=1.0 legal here) | recall 0.953 vs direct 0.871; held-out RULE induction 0.264/0.279 >= the teacher itself | history/timeline.md:758-767; history/timeline.md:783-806 |
| L-024 | 7 | 23111+0 | SQ.A0/shape-mean |
dual-teacher Procrustes consensus distillation | teachers .699/.649 -> student .761 EXCEEDS BOTH, still accelerating at E30 | canon/geometric_memory_blueprint.md:90-98; canon/gm3_parity/03_teachers_and_distillation.md:7-13 |
| L-055 | 7 | 22211+0 | -.NA/rotations |
Cayley orthogonality constraint + Newton-Schulz whitening | Q = (I-A)(I+A)^-1 guarantees pure rotation: det = 1.000 throughout, wins 76/84 unseen assays | canon/ryan_spearman_protein.md:19-23 |
| L-026 | 6 | 13211+0 | SQ.A2/recon |
soft-hand loss (proximity REWARD, not penalty) | V256 D24: MSE 0.034 at 400ep - 37% better than the best unconstrained run (.054) | canon/structural_attractors.md:38-56 |
| L-054 | 6 | 13211+0 | -.NA/arms |
quaternion composition as a structural regularizer (Hamilton product) | GeoQuat 0.916 -> 0.993 over 100 epochs vs best baseline 0.903 | canon/ryan_spearman_protein.md:12-18; canon/ryan_spearman_protein.md:26-29 |
| L-056 | 6 | 13221+0 | -.NA/architecture |
cascade as a regularizer (multi-step MLP instead of a direct dimensional jump) | 9-step 256->64 gives 84.6% vs a direct jump's 29.6%; a 27-step r=.95 cascade EXCEEDS the root | canon/geometric_memory_blueprint.md:38-40 |
| L-057 | 6 | 13201+0 | -.A3/routing |
Cantor router (soft weights derived FROM triangulation distances) | cos .9818 at 8 layers vs relay-alone .6533; geometry IMPROVES with more tokens | canon/constellation_forms.md:53-58 |
| L-022 | 6 | 22111+0 | KL.A0/logits |
logit-KD at alpha <= 0.25 with founder exemption | mlp_kd lineage 2.4106 -> 2.3707 -> 2.3662 -> 2.3594 monotone ascent; replicates at s1 | history/timeline.md:396-410; history/timeline.md:570-583 |
| L-034 | 6 | 12221+0 | SQ.A0/codebook |
entropy-balanced alignment cultivation (w = .05) | produced the emergent basin M = +/- ref EXACTLY - the sign-code convergence endpoint | canon/svae_h2_class.md:23 |
| L-045 | 6 | 12212+0 | DET.A0/simplices |
L_CM - Cayley-Menger validity hinge (lambda = .01) | CM validity 100% across the lineage table | canon/ksimplex_diffusion_prior.md:18-23 |
| L-046 | 6 | 12211+0 | DET.A0/simplices |
L_vol - volume-spread REWARD (-std(log | vol^2 | ), lambda = .005) |
| L-047 | 6⟂ | 12211+0 | SQ.A0/shape-mean |
Procrustes_SVD as a REGULARIZER (x 0.3 alongside a real force) | tightens CV (.19 vs .25) when it rides alongside InfoNCE | canon/geometric_memory_blueprint.md:22-24; canon/constellation_forms.md:26 |
| L-049 | 6 | 12221+0 | -.NA/anchors |
anchor dropout (30%) | prevents collapse: 508/512 anchors active | canon/constellation_forms.md:27 |
| L-060 | 6 | 12221+0 | -.NA/data |
data-level dampening (sqrt damping alpha=0.5, max_repeats=8, cap 1.25x) | NEVER equalize-to-largest: alpha=0 repeats 5 images ~50x/epoch | canon/subject_bucketing_anima.md:17-19 |
| L-064 | 6 | 12221+0 | SQ.A0/crystals |
rose loss (role-weighted pentachoron regularization, rose_w = 1e-4) | 74.87% CIFAR-100 @ 393,216 params vs ~65% zero-shot and ~70-72% linear probe | canon/fibonacci_systems.md:29-33; canon/geovocab_chunking.md:24 |
| L-039 | 5 | 12111+0 | CE.A7/similarity-grid |
contrastive dynamics as a CV-compression force | 100 clusters / 200 steps at d=128 -> CV .2451 (in band); 10 clusters -> .94 | canon/cv_band_origin.md:14-17 |
| L-053 | 5 | 12111+0 | -.NA/gradients |
geometric autograd / gradient gating (Form 12 tangential-radial split) | gradients split tangential (pass) / radial (attenuate) / collapse-direction (attenuate) | canon/constellation_forms.md:83-86; canon/gm3_parity/02_optimizer_and_cv.md:7-18 |
| L-025 | 5 | 21212+0 | SQ.A0/shape-mean |
projective-ICP / GPA consensus operator (germline overwrite) | recovers planted truth | cos |
| L-062 | 5 | 21222+1 | -.A2/data |
usage / starvation reweighting (drives DATA sampling, NEVER a loss term) | the program's ONLY answer to load balancing: x2 upweight the starved anchor's DATA, 3-strike abort | closeout_2026-07-19/amoe/src/amoe/train/aligner.py:68-84; pod/v35_exp007_collective.py:248-262 |
| L-063 | 5 | 21221+0 | -.NA/conditioning |
CFG dropout 0.1 (conditioning zeroed, not empty-prompt) | standard in every diffusion bed; never ablated in this program | pod2/dexp006_sd15core_relay.py:132-146 |
| L-021 | 5 | 11221+0 | SQ.A2/eps |
Min-SNR gamma=5 weighting + velocity adjustment | part of a working recipe (1 ep, 10k synthetic, ~7 min on an L4); never ablated | canon/ksimplex_diffusion_prior.md:28 |
| L-167 | 5 | 30222+0 | SQ.A8/pentachoron-volumes |
CV force at the 1e-3 ceiling on the anchored composite (x3 cv form) | INERT 2/2, third independent confirmation of the P4 shape: trajectory rides the baseline curve to ~.01 EMA the whole way; gauges in-noise (zs seed-spread if anything wider) | tools/dist_runs/x3_queue.log; tools/dist_runs/vitbank_s0.jsonl |
| L-163 | 5 | 20112+0 | SQ.A2/bank-geometry |
AlignmentBank 8-term geometry loss (CaptionBert Stage-B, vision port) | PROBE-REFUTED with FULL CONTROLS 2/2: trained bank BELOW untrained (gap -.0038/-.0045) AND a fully-random bank (random frames+anchors) matches the real-frames control (.4159/.4114 vs .4149/.4107) - the entire +1.4-1.6pt enrichment over raw is generic random-feature capacity; the geometric content contributes ~zero and geometric training subtracts | tools/dist_runs/vitbank_s0.jsonl; tools/dist_runs/vitbank_s1.jsonl |
| L-165 | 5 | 20222+0 | CE.A0/frozen-head-logits |
BCE through a frozen teacher-space head (x3 supervised force, isolated) | INERT as an add-on 2/2: every gauge within noise of the anchored composite (agree +.008 consistent-sign whisper); supervision distilled through a frozen head does not behave like supervision | tools/dist_runs/x3_queue.log; tools/dist_runs/vitbank_s0.jsonl |
| L-058 | 3 | 01211+0 | -.NA/codebook |
address-agreement bias (BUCKET - making a hard address differentiable) | exact softmax within sorted equal-width blocks masked to the same bucket; argmax alone is gradient-dead | canon/mother_threads.md:35-37; canon/constellation_forms.md:118-120 |
| L-069 | 3 | 00221+0 | -.A2/sequence-axis |
predictability-weighted accumulation (PWA) | DESIGNED 2026-07-25: make the PREDICTABILITY PRINCIPLE a loss geometry instead of a discovered side effect | history/timeline.md:1062-1066; inventory/SCALE_MANIFEST.md:144-154 |
GAUGES - measure only, never in the gradient path
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-071 | 10 | 33212+1 | DET.A8/anchors |
CV as a READOUT (never a force) | the historical CV 'loss' was GRADIENT-FREE all along - .item() stripped the graph | canon/discovery_catalog.md:9-10; canon/aleph_core.md:40-41 |
| L-073 | 10 | 33222+0 | -.A0/bytes |
bpb (bits per byte) - the AR line's verdict currency | certified band 2.469-2.499; addr_msl64 beats the unrestricted head 7/7 across seeds and budgets | inventory/SCALE_MANIFEST.md:53-58; history/timeline.md:139-147 |
| L-074 | 10 | 33222+0 | -.A0/logits |
perplexity tax ladder (wikitext ppl delta, one shared gauge) | one always-on stack +9.23/+9.87 | monolith +3.66 |
| L-075 | 10 | 33222+0 | -.A0/text |
token-F1 (caption distribution-match delta) | 0.408 -> 0.706/0.704 (+0.30, | s0-s1 |
| L-080 | 10 | 33222+1 | -.NA/logits |
toggle law - all anchors off is BIT-EXACT to the base model | max | dlogit |
| L-083 | 10 | 33222+0 | -.A0/codebook |
usage perplexity / axis aliveness (read-only) | 125+/128 axes alive WITHOUT regularization - the standing refutation of load-balancing auxiliaries | canon/discovery_catalog.md:41-42; tools/geolip_vitals.py:108 |
| L-085 | 10 | 33222+1 | -.A0/anchors |
blend-escape ratio (threshold 1.5) and damping ratio (target >= 3x) | specialists damped 5-11x but caption ESCAPES undamped at 0.1004 - the corollary that became the regime law | closeout_2026-07-19/amoe/src/amoe/laws.py:37-39; history/timeline.md:1961-1964 |
| L-089 | 10 | 33222+0 | DET.A8/anchors |
CV@1000-batches early screen + the 3-tier filter | CV at 1000 batches PREDICTS the final band; turnaround ~2h -> ~7 min per config | canon/tri_band_omega_arc.md:21-27; canon/discovery_catalog.md:11-13 |
| L-092 | 10 | 33222+0 | -.A0/patches |
resolution-invariance flatness (the debugging canary) | 4.5% MSE variance from 81 to 4096 patches; ~1% across a 36-config sweep - ANY shift means an upstream break | canon/discovery_catalog.md:20-22; canon/tri_band_omega_arc.md:67-70 |
| L-095 | 10 | 23222+1 | -.NA/hardware |
peak_mem + s/step (the WDDM sysmem-spill tell) | the tell is ~100W/450W at '100% util' with no step prints - 42.8GB observed spilled to shared memory | MANIFEST.md:50-59; pod/run_g6b_rider.py:34-48 |
| L-077 | 9 | 33221+0 | -.A0/codes |
register probe (sign-code inter-minus-intra Hamming separation) | THE PREDICTOR of the two-regime law: registers ~0.2-0.3 blend, domains ~0.35-0.5 specialize | canon/register_probe_gauge.md; inventory/SCALE_MANIFEST.md:38-42 |
| L-081 | 9 | 33212+0 | -.A3/bands |
band-lesion surgical test (own vs cross damage) | surgical 3/3 both seeds at 50-200x; on a DiT edge bands hit cross-damage EXACTLY 0.0 | MANIFEST.md:70-76; history/timeline.md:1851-1864 |
| L-091 | 9 | 33212+0 | -.A8/spectra |
spectral gauges: S0/S_D ratio, effective rank, the universal attractor | critical ratio ~6.5 triggers DISCHARGE; universal attractor S0 ~5.1, erank 15.88 +/- 0.04 across 48+ measurements | canon/structural_attractors.md:24-30; canon/discovery_catalog.md:5-7 |
| L-094 | 9 | 33221+0 | -.A0/structured-text |
structured-task validity judges (JSON validity, IoU, pair-order, termination) | bbox 0 -> 0.6875 valid (0.894 IoU); the FORMAT TRAMPLING signature: 9/12 truncated_no_json | inventory/SCALE_MANIFEST.md:253-255; sessions/2026-07-14_pod.md:378-394 |
| L-076 | 9 | 23222+0 | -.A0/text |
precision + invented-attribute rate (the hallucination decomposition) | precision 0.356 -> 0.694/0.705 and invented-attribute rate 0.200 -> 0.136/0.101, BOTH seeds | inventory/SCALE_MANIFEST.md:233-238; history/timeline.md:1679 |
| L-078 | 9 | 23222+0 | -.A8/codes |
sign_fidelity (Spearman of code-Hamming vs true angular distance) | PROMOTED: separates heredity from lottery where bpb CANNOT - heirs lock at .9555-.9558, spread < .001 | history/timeline.md:476-521; inventory/SCALE_MANIFEST.md:67-69 |
| L-086 | 9 | 23222+0 | -.A0/text |
composition score (the controller prereg gauge) | the chaining wall: components >= 0.96 solo, composite 0.0 for EVERY config | inventory/SCALE_MANIFEST.md:261-270; history/timeline.md:1657-1660 |
| L-088 | 9 | 23222+0 | -.A0/gradients |
grad_norm_spread (gradient democracy monitor) | reference failure it exists to catch: 20 orders of magnitude across unequalized towers | tools/geolip_vitals.py:160-173; canon/fibonacci_systems.md:14-20 |
| L-097 | 9 | 23222+0 | -.A0/text |
held-out byte accuracy (rule induction) and variant-format recall (the format lock) | teachers memorize at 1.000 train but induce at 0.270/0.245 held-out; memorized content is BOUND to surface form | history/timeline.md:783-806; history/timeline.md:794-800 |
| L-098 | 9 | 23222+0 | -.A0/codes |
key-durability gauge (nearest-neighbour symbol Hamming + key drift) | sign-code keys disagree on ~91% of symbols; match rate at theta=0.25 is 0.000 EVERYWHERE | history/timeline.md:868-894 |
| L-099 | 9 | 23222+0 | -.A0/codebook |
basin mean_cos (BASIN SET AT INIT) | 192-bank sweep: epoch_1 .8632 / best .8635 / final .8615 - delta 0.0017 BELOW the within-phase std | MANIFEST.md:124-127; canon/tri_band_omega_arc.md:71-75 |
| L-100 | 9 | 22222+1 | DET.A8/simplices |
cv_reference_check (fp64 parity against the source of truth) | exact parity (relative 0.0) at fp64 against geovocab2, at ~260x the speed | MANIFEST.md:32-39; tools/geolip_vitals.py:78-103 |
| L-079 | 8 | 23212+0 | SQ.A3/band-routed |
role-aligned in-bed gauge (HIGH-band foreground-masked LP-x0) | PROMOTED: found a ~10% multiband win that EVERY aggregate comparison hid | history/timeline.md:1465-1479 |
| L-082 | 8 | 23212+0 | -.A6/routing |
repeated-key null + matched-vs-mismatched deltas | the instrument that falsified address-as-key: routing excess 2.5e-06 over the null | inventory/SCALE_MANIFEST.md:203-208; history/timeline.md:1611-1629 |
| L-096 | 8 | 22222+0 | -.A0/anchors |
consensus drift / stationarity gauge | ROBUST for structured genomes (0.003 drift by g2, both seeds) but SEED-DEPENDENT for a lone flat book | inventory/SCALE_MANIFEST.md:79-86; canon/gm3_parity/01_inherited_object.md:77-85 |
| L-090 | 7 | 22211+0 | -.A8/codebook |
void topology beta_2/axis (persistent homology on RP^(D-1)) | within the D=4 cohort every GEOMETRIC signal collapses while VOIDS rise; beta_2 vs recon MSE | rho |
| L-084 | 6 | 13222+0 | -.A0/codebook |
read perplexity + | cos to nearest atom | (the quantizer gauge) |
| L-087 | 6 | 13222+0 | -.A0/any |
adapter_effect_mean - the VACUOUS guard | returns VACUOUS instead of a false PASS when the stack barely moves the loss | sessions/2026-07-21_comfyui_amoe_packaging.md:363-371 |
| L-093 | 6 | 12222+0 | -.NA/code |
exec judge (guarded subprocess: restricted builtins, length cap, hard timeout, no network) | the write-0.0 floor was verified GENUINE off-pod, not a judge artifact | history/timeline.md:1634-1635; plans/2026-07-16_math_night_plan.md:46-51 |
| L-104 | 6 | 12222+0 | -.A0/codes |
sign-code Hamming retrieval | 0.359 @1 against the continuous head's 0.494 - ~73% of its power from raw 64-symbol Hamming | sessions/2026-07-13.md:116-121 |
DISTRUSTED GAUGES - do not rate anything on these alone
Each was load-bearing until it was caught. Any entry whose headline rests on one takes the -2 blind-gauge demotion (rule 2) until it is re-measured on a promoted gauge.
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-072 | 8 | 32221+0 | -.A0/anchors |
anchor drift -> 0.29154 rad + binding_fraction | the binding constant recurs across 5 architectures and 3 paradigms - but the drift-based fraction is a STAGE statistic | MANIFEST.md:115-116; history/timeline.md:231-241 |
| L-101 | 5 | 31221+0 | -.A0/gates |
gate-mean band 0.012-0.03 (advisory, NOT universal) | held across 6 architectures and 2 optimizers - then MISSED on a 7th at 0.051-0.061 | MANIFEST.md:121; history/open_questions.md:31-33 |
| L-102 | 5 | 30220+0 | SQ.A0/eps |
aggregate eps-MSE as a band-behaviour gauge | DISTRUSTED: moved 0.2% against +0.089 grounding effects in image space, and HID a ~10% multiband win | history/timeline.md:1402-1422; history/timeline.md:1465-1479 |
| L-103 | 2✖ | 20020+0 | -.A0/addresses |
recon cosine as a judge for ADDRESSED systems | DISTRUSTED: an address is a LOOKUP KEY, not a compressor - judge drift and crushed CV instead | MANIFEST.md:115-116; history/open_questions.md:198-204 |
RETRACTED / DEMOTED / FORBIDDEN - the doctrine's proof set
Each row below is the EVIDENCE for a standing law. This is not a graveyard and not an appendix - these are the controls that turned results into laws. Cite them.
Grouped by the law each row proves, not by entry.
absolute-beats-relative
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-105 | 1† | 30012+1 | SQ.A10/codebook |
VQ / commitment / EMA codebook losses | THE NAMED PROHIBITION - and unnecessary: the codebook stays 125+/128 alive at div_weight = 0 | canon/aleph_core.md:16-18; MANIFEST.md:132-134 |
| L-106 | 1† | 30012+1 | CE.A0/roster |
comparative / relative selectors (argmax anchors, softmax-over-roster, STE one-hots, k-means alphabets) | roster-dependent; the gradient HOMOGENIZES - 14x path collapse, width attenuation, BN-on-padding, same disease | MANIFEST.md:135-138; canon/aleph_core.md:53-58 |
| L-107 | 1† | 20011+1 | CE.A0/alphabet |
gradient-learned alphabets (CAMPAIGN LAW 3) | fitted-frozen alphabets differentiate (1,594 unique paths); gradient-learned alphabets COLLAPSE (116) | MANIFEST.md:157-158; canon/acd_campaign.md:18-20 |
| L-108 | 1† | 10011+0 | SQ.A0/crystals |
direct gradient descent on pentachora | collapses them to zero - as FROZEN anchors the same crystals retain full cohesion and stay backtrackable | canon/geovocab_chunking.md:16-18 |
| L-118 | 1† | 30012+1 | SQ.A0/routing |
comparative routing on diffusion (state+sigma, raw address, M-hat address-as-key) | FALSIFIED THREE WAYS, 2 seeds: routing excess 2.5e-06 over the repeated-key null; match advantage -0.0 | MANIFEST.md:70-76; inventory/SCALE_MANIFEST.md:203-208 |
regime-law
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-119 | 1† | 20212+0 | CE.A0/logits |
the controller hypothesis (a trainable anchor that orchestrates the others) | prereg required >= +0.15; measured -0.417 / -0.167. The passenger role is an ATTRACTOR | inventory/SCALE_MANIFEST.md:261-270; sessions/2026-07-14_pod.md:610-646 |
| L-120 | 1† | 30112+1 | CE.A0/logits |
always-on solo specialist stacks | MUTUALLY DESTRUCTIVE at n=48: the depth stack drives caption F1 to 0.0014 with termination 0.0 | MANIFEST.md:160-166; inventory/SCALE_MANIFEST.md:216-232 |
| L-121 | 1† | 20212+0 | CE.A0/logits |
frozen solo-trained expert collectives under aleph dispatch | no surgical independence (own-drop 0.04/0.00), NO damping (all five blend-regime, 0.86-1.6), composite 0.0 | inventory/SCALE_MANIFEST.md:169-178 |
instrument-design
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-125 | 1† | 10222+0 | -.A6/routing |
shuffled-key null | CONFESSED INSTRUMENT FAILURE: it measures diversity, not correctness - the null scored like the real thing | history/timeline.md:1539-1548 |
| L-127 | 1† | 20221+0 | -.NA/anchors |
the exp021 seed-inversion claim for the trainable anchor | RETRACTED WITHIN HOURS: the claim compared DIFFERENT INSTRUMENTS across seeds | history/timeline.md:1697; hfstage/exp021_s1battery/README.md |
| L-133 | 1† | 10222+0 | -.NA/sampling |
deterministic (greedy) decoding in an iterative denoiser | collapses to the global mode: diversity 0.0, conditional == shuffled EXACTLY | sessions/2026-07-13.md:35-43 |
coordinate-law
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-128 | 1† | 10021+1 | -.A3/cantor-measure |
hierarchical refinement in Cantor space | HARMFUL (-10%); parallel ADJACENT NON-OVERLAPPING bands are +3% | canon/cantor_empirics.md:5-18; MANIFEST.md:122-123 |
| L-129 | 1† | 10021+1 | -.A3/cantor-measure |
repeated boundary crossing in a measure space | KILLS gradients (catastrophic -> random). Enter and exit the measure space ONCE | canon/cantor_empirics.md:5-7; MANIFEST.md:122-123 |
| L-130 | 1† | 20022+0 | -.A3/cantor-measure |
the SOFT devil's staircase used as a BAND COORDINATE | NEW 2026-07-25: measured NON-MONOTONE - min slope -0.13 to -0.49 at EVERY level count on EVERY grid tested | .venv/Lib/site-packages/geofractal/model/positional/staircase.py:49-67; MANIFEST.md:90-96 |
aggregation-destroys
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-109 | 1† | 30022+1 | -.A0/patches |
global average pooling in geometric encoders | 70% -> 29% collapse, REPLICATED independently in the protein line | MANIFEST.md:24; canon/discovery_catalog.md:17 |
| L-131 | 1† | 10021+0 | -.NA/data |
equalize-to-largest data balancing (alpha = 0) | repeats the 5-image bucket ~50x per epoch - 'the textbook way to overfit the long tail you were trying to protect' | canon/subject_bucketing_anima.md:17-21 |
consumption-law
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-112 | 1† | 20022+0 | CE.A0/coefficients |
addr_head - coefficients to logits at a single hard tau | 5.6650 bpb COLLAPSED: usage ppl 1.88/64, TWO unique winners, win | cos |
| L-137 | 1† | 20022+0 | CE.A0/coefficients |
single hard-tau coefficient heads at ANY dimension | DEMOTED on the standing registry: collapse, and low-D was falsified as the fix | inventory/SCALE_MANIFEST.md:137-139 |
inheritance
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-114 | 1† | 20112+0 | KL.A0/logits |
logit-KD at alpha = 1.0 from near-parity teachers | INVERSE EVOLUTION, compounding downward: 2.4301 -> 2.5046 -> 2.5603 | history/timeline.md:383-395; canon/gm3_parity/03_teachers_and_distillation.md:14-24 |
| L-122 | 1† | 20222+0 | -.NA/weights |
organ-only inheritance (projection + book transplanted onto fresh trunks) | BELOW random init, 2/2 lineages - sixteen random draws beat organ heredity | inventory/SCALE_MANIFEST.md:132-134; history/timeline.md:396-437 |
regime-transfer
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-117 | 1† | 20222+0 | CE.A0/coefficients |
tied M-hat readout (U=M_hat, S=Omega-token, Vt=I) in an AR head | +1.0 bpb BOTH seeds and it STARVES the codebook (drift 0.02, binding 0) | history/timeline.md:715-736; canon/gm3_parity/01_inherited_object.md:77-85 |
| L-136 | 1† | 10211+0 | SQ.A0/rotations |
SVD-rotation transform in the dual-pentachoron head | DROPPED for convergence failure; reduced to scale + shift | canon/cayley_menger_foundations.md:36-38 |
alignment-force
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-019 | 1† | 12111+0 | CE.A7/mixed |
Expert Soup composite (InfoNCE + MSE + BCE + Procrustes + CV + spread) | RETRACTED 2026-07-31 with its source system: the Form-2 / Bertenstein numbers (mAP .84, 1000/1000 positive volumes, eff-dim 76.9) were never independently audited, and the system carrying them was falsified — R@1 was a shared-forward-pass leak, Procrustes cos_after sat at/below a no-relationship null, CV .20 at d=1024 was collapse | canon/constellation_forms.md:35-46 (retirement annotation); repos/bertenstein.md |
cv-pressure
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-110 | 1† | 20011+0 | DET.A0/anchors |
CV loss as backward injection / above the 1e-3 ceiling | MUST be a forward loss; above ~.001 the CV term dominates CE and trades discrimination for regularity | canon/constellation_forms.md:85-86; canon/geometric_memory_blueprint.md:87 |
force-vs-regularizer
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-111 | 1†⟂ | 10112+0 | SQ.A0/shape-mean |
Procrustes as a training FORCE | as a training loss: R@1 = 0.000, P_cos stuck at .094 for THIRTY EPOCHS | canon/geometric_memory_blueprint.md:22-24 |
infonce-placement
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-113 | 2✖⟂ | 33011+0 | CE.A7/similarity-grid |
InfoNCE into ADDRESS paths | BANNED despite R@1 .999 - it is the LOUDEST gradient and the bank learns IT instead of the useful signal | MANIFEST.md:135-137; tools/exp013_augmentation_bed.py:18 |
conditioning-law
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-115 | 1† | 30012+1 | SQ.A5/eps |
blob structural supervision on the EPS objective | +0.03% / -1.0%, two seeds - the x0 recovery divides by a vanishing sqrt(alpha_bar) EXACTLY in the supervised band | MANIFEST.md:65-69; history/timeline.md:1465-1500 |
screening
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-116 | 1† | 20220+0 | SQ.A0/recon |
MSE-first single-epoch keep-or-kill screening | DEAD: the lowest-MSE config was a HIGH-band false candidate | canon/tri_band_omega_arc.md:23-27 |
band-roles
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-123 | 4⟂ | 31210+0 | SQ.A3/band-routed |
HP/LP band-role objectives [judged by AGGREGATE eps-MSE] | 4/4 directional both seeds at 0.05-0.2% margins - 'nearly collinear with the base objective' | history/timeline.md:1394-1435 |
redundant-in-context
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-124 | 5 | 21222+0 | SQ.A0/conditioning |
frozen-address conditioning injected beside full text | real vs deranged -0.0009 beside full text; but ALONE the address steers at +0.0287 | inventory/SCALE_MANIFEST.md:209-210; history/timeline.md:1305-1318 |
question-space
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-126 | 1† | 10222+0 | CE.A0/logits |
the sequences / baseconv expert gains | SELF-RETRACTED: question space 480 and 248 against 800 training draws per tier = MEMORIZED | history/timeline.md:1640 |
law2-placement
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-132 | 1† | 10122+0 | CE.A0/conv-filters |
addr_conv - the decorative address (convex re-weighting of a filter bank) | DECORATIVE: a convex sum a_k = 1 is a hull-bounded perturbation of a MEAN; the 1x1 address is CONSTANT on grayscale (variance 4e-16) | history/timeline.md:1890-1896; sessions/2026-07-24_antipode_conv_decouple.md:24-30 |
no-balancing
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-134 | 1† | 30022+1 | CE.A10/routing |
load-balancing / auxiliary router losses | BANNED and replaced by architectural equality; ZERO instances exist in the tree | history/progression_plan_2026-07-08.md:126-137; closeout_2026-07-19/amoe/src/amoe/core/dispatch.py:1-12 |
operator-ruling
| ID | RATE | RPDCI+L | CELL / SUBSTRATE | NAME | HEADLINE NUMBER | CITE |
|---|---|---|---|---|---|---|
| L-135 | 1† | 00201+0 | CE.A0/logits |
the big-JSON objective | FORMALLY DROPPED by operator ruling - too costly; 3-5 task adapters deliver more per GPU hour | history/timeline.md:1977-1981; sessions/2026-07-14_pod.md:813-817 |
MECHANISM NOTES
Keyed by ID so grep L-115 finds both the row and its reasoning. Full notes live in
inventory/loss_manifest.json; python skill/loss_view.py show L-115 prints one.
OPEN CONTRADICTIONS (never averaged, per rule 7)
- CE vs the CV band. One record has Form-1 CORE holding CV .2045 under cross-entropy; the repo results table shows CV drifting .238 -> .124 by ep50 against CE at weight .01. Bears on L-003's doctrinal standing. Settled by: a same-config re-run, or a provenance split showing the two runs are not the same configuration.
- The Meridian s_cos plateau at .425. Two recorded explanations - a 1280->1024 dimensional mismatch, or pooled-bank gradient loudness. Bears on L-020. Settled by: re-running the bank at matched dimension.
- soft-vs-mag is depth-scoped. Magnitude leads +10.3 at 3 layers; soft ran AHEAD at 6 layers with cosine schedule. SCALE_MANIFEST already forbids stating the modulus finding as a law. Settled by: both arms run to completion at depth.
CROSS-REFERENCES
canon/loss_accumulation_forms.md- the taxonomy, the eleven format laws, the binning procedure, the A10 statute.inventory/SCALE_MANIFEST.md- answers "what scales". This file answers "what is allowed to pull, and on what evidence". A mechanism in both carries the other's ID.canon/discovery_catalog.mdentry 26 - the three-primitive finding as a program-wide law.inventory/EXPERIMENT_LINES.md- resolves the line-qualified IDs used in citations.
AMENDMENTS (append-only, dated)
- 2026-07-25 - created. Rubric v1, 137 entries, from three independent censuses. Two findings NEW to the program are folded in as rows: the soft devil's staircase is non-monotone and therefore inadmissible as a band coordinate (L-130), and alpha=0.5 is the unique expectation-matching value for the staircase (L-044).
- 2026-07-25 - A0 AUDIT (Phil: "is A0 correct?"). Partially no, three ways: (1) two PRIMARY misbins fixed - L-006 (HF labels= path is masked CE inside) and L-015 (derived-steps trains shift-CE with ignore_index=-100) both move A0 -> A4 per binning rule 4; (2) the fallback bucket had swallowed 25 rows with NO residual-reduction semantics (structural devices, optimizer/data disciplines, assertion protocols, decoding policies, claims) - a category error, now an explicit NA class; (3) what remains at A0 (71 of 138) is honestly dominant: plain means ARE the program default, and the aleph's only pressure lives there. GAP stays at A0 deliberately - it IS a uniform-mean accumulation, the format's cautionary instance.
- 2026-07-25 - pass 3 measured (same day). L-070 (FAC) moved from UNRUN to a scope-split pair: as a PRIMARY objective it is REFUTED as preregistered, 3/3 seeds (loses to ce, to ce_fixedcode, and to fac_none on the bpb-of-record); on the partition-collapse configuration (new L-138) a loss swap alone decompresses the certified addr_head catastrophe (usage ppl 1.0-2.7 -> 60.6-61.1 of 64, decoded acc 0.05-0.20 -> 0.45-0.47, 3/3 SEEDS, spread < 0.5) and dissolves the win|cos| saturation (0.9995+ -> 0.13) - amending L-112's recorded geometry attribution. The collapse follows the loss.
- 2026-08-02 - L-019 RETRACTED (rule 6). The Expert Soup composite's sole source (Form 2 / GEOLIP-Bertenstein) was retired 2026-07-31: its R@1 was a shared- forward-pass leak (flat 1.0000 from gallery N=64 to 4096; 0.0007 masked), its Procrustes cos_after (.377-.440) sat at or below a no-relationship null at the same fit size, and its CV .20 at d=1024 was collapse (null .0207, effective rank 19.7). The composite's own numbers (mAP .84, eff-dim 76.9) were never independently audited and no ablation isolates any of its six terms. Rating 5 -> 1 dagger, role PRIMARY -> RETRACTED. WHAT SURVIVES is the form, not the composite: whitened-Procrustes pre-alignment WITH held-out fit discipline (N/d 31.3 retains 95% out-of-sample; N/d 4.9 retains 3.8%) and per-expert projectors into a REFERENCE MEMBER's frame. Its validated descendant is L-162, measured 2/2 on two modalities. repos/bertenstein.md · canon/shared_pass_contrastive_leak.md · repos/captionbert-v2.md