diff --git a/docs/design/class-v6-rotating-family.md b/docs/design/class-v6-rotating-family.md index 10494e6f2..9ba8d2f1b 100644 --- a/docs/design/class-v6-rotating-family.md +++ b/docs/design/class-v6-rotating-family.md @@ -150,7 +150,7 @@ Every test the generator applies to a candidate program is today a function of t |---|---|---|---|---| | (c''') the distinct-item ratio floor | 0.995 over the 64 units at the genesis width and mixer | the same floor evaluated with the era's width (a 16-byte load touches one item too) and dataset size; 2.435 percent of candidates under it today, attempts +3.4 percent | class-v5 section 14 (measured census of 4,600 candidates) | a candidate below 0.995 on the exemplar seed 100767 is refused | | F8's uniformity (the largest 64-line bucket within 6 sigma over 2^28 derivations; the top 0.1 percent of items within 1.2x of the window model over 2^24 nonces on 64 seeds) | a gate run by hand per class on the attack board | run by the census tool per era draw at genesis (the band's corners plus 64 random eras) and by the node's acceptance as the per-site version below; a draw whose corner fails is excluded from the band | f8-uniform.md section 7: +4.84 sigma against the control's +4.18, PASS; the tail p4, p8, p10, p34 attributed this morning as per-site bucket concentration at a narrow-window site | the `quarter-lines` and `const-item` plants fire at +75.97 and +92,682 sigma (f8-uniform.md 2.1) | -| The per-site largest-bucket bound (this morning's attribution) | named for a next class | per load site, the largest 256-item bucket over the units' addresses, stated in SIGMA against its own window's Poisson expectation, never as a ratio: ratios of 2.2 to 2.4 at full-window sites are the CLEAN maximum (65,536 Poisson(16) buckets read +4.4 sigma), so a ratio bound would refuse clean programs; lane D's rebuild carries the sigma column | AP-F8-1's tail: p10 1.50x, p8 1.38x, p34 1.25x, p4 1.22x, each a narrow-window site's bucket; the sigma figures from lane D's family-gate.md (17:00 UK). **Two F8-256 attributions in hand (the hash lane, 14:0x UK, labelled estimate): p212 (the attack-pass lane, the gate's line) is site 9's bucket concentration at 1.75x its window expectation with 0.019 bits short, and p225 (the hash lane's run on the 1c420786 crate) is a value-level concentration at a mad-written site with the bucket at expectation and full entropy; so a per-site bucket bound at about 2x, which catches all four of the morning's tail (3.1x to 5.6x), catches neither: p212 sits under 2x and p225 has no bucket signature. Layer 4 carries both tests: adv-cache-2's value-level bias test (the item histogram's top against the window model per site) is the one that sees p225, and a bucket bound would need to sit near 1.5x to take p212 at a clean-seed cost now put at 3 to 6 percent rather than 1 to 3 (the clean spread's buckets read up to about 1.5x on the sites read today); the generalised era-level uniformity test is unchanged** | the four tail seeds must be refused on sigma; the 60 passing seeds accepted; p212 and p225 the two cases the bucket bound alone misses | +| The per-site largest-bucket bound (this morning's attribution) | named for a next class | per load site, the largest 256-item bucket over the units' addresses, stated in SIGMA against its own window's Poisson expectation, never as a ratio: ratios of 2.2 to 2.4 at full-window sites are the CLEAN maximum (65,536 Poisson(16) buckets read +4.4 sigma), so a ratio bound would refuse clean programs; lane D's rebuild carries the sigma column | AP-F8-1's tail: p10 1.50x, p8 1.38x, p34 1.25x, p4 1.22x, each a narrow-window site's bucket; the sigma figures from lane D's family-gate.md (17:00 UK). **Two F8-256 attributions in hand (the hash lane, 14:0x UK, labelled estimate): p212 (the attack-pass lane, the gate's line) is site 9's bucket concentration at 1.75x its window expectation with 0.019 bits short, and p225 (the hash lane's run on the 1c420786 crate) is a value-level concentration at a mad-written site with the bucket at expectation and full entropy; so a per-site bucket bound at about 2x, which catches all four of the morning's tail (3.1x to 5.6x), catches neither: p212 sits under 2x and p225 has no bucket signature. Layer 4 carries both tests: adv-cache-2's value-level bias test (the item histogram's top against the window model per site) is the one that sees p225, and a bucket bound would need to sit near 1.5x to take p212 at a clean-seed cost now put at 3 to 6 percent rather than 1 to 3 (the clean spread's buckets read up to about 1.5x on the sites read today); the generalised era-level uniformity test is unchanged. **RESOLVED by lane D's measurement (build-2, 14:2x UK, the family harness at the shipped parameters, every attempt equal to the chain's class v5 draw; family-gate.md section 6.8 in the 16:30 landing): p212's largest 4,096-word bucket is 1.72x its window expectation at +5.75 sigma at the 2^20 sample, inside the clean spread's own maximum (+5.5 median, +6.8 p99), so no bucket band under a 10 percent clean cost takes it; its index-bit read is -512 sigma at site 9, bit 8 = R (the product's bit 0 at P = 1/4, z = -512 exactly). p225 reads a bucket at expectation and a bit bias of -567 sigma at site 4, bit 2 = R. The morning's tail is the same mechanism one window up: p4, p8, p10 at -511 to -513 sigma at bit R (13, 22, 17) with buckets +49, +22, +42 sigma because R lands inside the bucket's bits on a quarter window; p34 +10.8 and p15 -58 at the bit level, buckets clean. The bucket bound at +8 sigma catches p4, p8, p10 and neither of the two; at 1.5x it would take p212 at the biased population's cost (14.5 percent of eras over +8), not a clean 3 to 6; the bit read at 6 sigma catches all seven at the standing 48 to 52 percent of eras; at 300 sigma it catches five (every one at or above 511) at 16 percent of epochs redrawn, and nothing clean at any band (the clean maximum over 26 bits and 16 sites is 5.5 sigma at p99); the bucket's independent catch beyond the bit read is 0.07 to 0.23 percent of eras. What ships: ONE value-level test, the per-site index-bit one-count at the acceptance's 2^20 sample, as a per-era record (its z at bit R names the product writer) and as the known-failed set of the structural fix (the index fold of the product's low bits before the rotation in `load_index`, layer 1's rule), the bucket bound retired into it; if a refusal is wanted before the fold lands, the band is 300 sigma, never 6**, measured | the seven cases (p4, p8, p10, p15, p34, p212, p225) are the known-failed set of the fold; the bucket bound retired | | The value-level bias test (adv-cache-2; the research file's 20.2b) | named for a next class | per load site, the one-count of every index bit over the 64 units within 6 sigma of n / 2 (the per-load prototype's `BiasedIndexBit`, built last night, measured as the record); AS A TEST ONLY AFTER the index fold of layer 1 is in, because on today's `load_index` it would redraw about 40 percent of epochs (lane D: 7 of 16 drawn eras at abs z 130 to 511); with the fold in, the test is the guard that the fold holds | a product's low bits at P(bit 0) = 1/4 placed at address bit R by the stride rotation; 6 of 16 drawn eras over 1.04x, 14 of 17 eras flagged by the instrument on the pre-amendment generator | a program whose site is sourced by a product under an era with R under 28 is refused; the devnet era's R = 29 is not relied on | | The duplicate-lane test (the research file's 20.2a) | built for the per-load class only | kept as a per-load-only test unless a drawn block shape ever places shadow work between loads (layer 1 does not: the block shape is the size, the placement stays after instruction 63) | 1,482 duplicate lanes on the per-load record, 0 to 2 on every sound class | the per-load candidate 0 is refused | @@ -172,7 +172,7 @@ What the cut settles, each row measured unless marked: |---|---|---| | The op-mix band is settled by measurement | B = 4 with or, mul and mulhi free to rise exhausts the 256-attempt cap in 1.2 percent of eras (61 of 5,000 at the lossy corner); with or, mul and mulhi never raised above their base (section 2's band) r = 0.595, mean attempt 1.47, max 59, 0 exhausted in 3,000. The lossy-share curve, complete at 3,000 eras per point (13:5x UK): r = 0.80, 0.88, 0.92, 0.96 at +1 to +4 points on or, mul and mulhi; exhaustion 0, 0.07, 0.20, 1.10 percent of eras, against independent-attempt estimates of 4e-25, 3e-15, 9e-10, 1e-5, so the per-era correlation is 10^5 to 10^10 above the geometric figure and the band's edge is the measured +2, not the arithmetic's | the layer 1 op-mix row's known-failed test now reads 0 of 3,000 under the band (it owed a 0); the cap on the lossy families stays at the base, with +2 points the most the band could ever open to | | What breaks at the lossy corner | class v5's last-resort scan passes at its first or second candidate on every exhausted era seen, so the corner costs liveness time, not an unchecked program; the per-era exhaustion is about 1,000x the independent-attempt estimate because one era's attempts share its weight table, which is the independence the scan's 1e-300 assumes | section 5.2's bound: the attempts within an era are not independent draws; the bound is per era from the census, not r^256 | -| The (c''') floor needs a per-width calibration | at width 4 (expectation divided by the width) it refuses 4.5 to 6.5 percent of candidates against 0.6 to 1.0 percent at width 1 | ring B's row: with W pinned at 4 at genesis (section 10.3) it is one measurement, being taken now (the harness's next build records the pre-floor spread of every candidate at width 4 under the band over 3,000 eras plus the width-1 control); the 09:00 report states either a single width-4 floor at the shipped clean-refusal rate (about 2.4 percent of candidates) or the sigma-over-expectation form, whichever keeps the known-failed hot sets refused at the lower clean cost; the default here is the sigma form | +| The (c''') floor needs a per-width calibration | at width 4 (expectation divided by the width) it refuses 4.5 to 6.5 percent of candidates against 0.6 to 1.0 percent at width 1 | ring B's row: with W pinned at 4 at genesis (section 10.3) it is one measurement, being taken now (the harness's next build records the pre-floor spread of every candidate at width 4 under the band over 3,000 eras plus the width-1 control); RESOLVED at 14:2x UK: the class v5 floor at 0.995 refused both hot sets the live point-A census found under a band era's weights (p38 at 0.9932, p54 at 0.9945, the draw moving to the next attempt at 0.9999), so the floor stays at 0.995 at W = 4 at its measured cost (14.6 percent of candidates reaching the 2^20 pass, +0.3 attempts per epoch); the width-4 statistic has a real tail (10 percent of reaching candidates under 0.991 against 2 percent at width 1, with the expectation divided by the width), so neither a lower floor nor the sigma form keeps the width-1 cost, and the 0.990 to 0.995 band at width 4 is read live before any floor moves | | The shape axis moves the draw's cost, the mixer axis is invisible | 64 x 108: 1.8 attempts; 256 x 27: 3.8, through (a')'s fixpoint over the block; r 0.714 to 0.731 across m | m is closed per family by adv-mixer-3's ladder at m = 4 and the x16 verifier row, not by drawing eras (section 5.1a's reading stands, now measured); the block-shape row of layer 1 carries the attempt cost | | The era-stride bias at the bit level | 48 to 58 percent of accepted programs in every stratum carry one site biased at over 6 sigma at 2^20; 73 percent of those at address bit R exactly (12 percent at R+1, 5 at R+2: the product law's bits 0, 1, 2 through `rotl(x*M, R)`); a third over 100 sigma, worst z 1,024; under R in 28..31 the over-100-sigma share falls from 37.7 to 7.5 percent; the bucket statistic sees the same mechanism (p99 +6.8 sigma on bit-clean eras, +44 on biased ones) | one value-level test covers both; the remedy is structural, the index fold of the product's low bits in `load_index` before the rotation (layer 1's rule, section 2), not a per-epoch refusal that would redraw half the epochs; the live-dataset price per site (ring C) is being read now (the F8 census at 2^24 on 64 seeds at two band points; 31 of 64 seeds PASS so far at the shape-256 point, no test fired) | | The bound arithmetic on the counts | 10,000 random eras and 3,000 band eras passing bound the failing fraction on the ring-B tests at 3.0e-4 and 1.0e-3 at 95 percent; the floors' miss rates on the live-dataset classes rest on 9 and 5 known-failed cases (under 0.33 and 0.60), tightened by cases from the adversarial tails, not by eras | section 5.1a's sampling bound now has its measured n; the gate-record JSON per era lands with the full report |