Restructure: descriptive tier and experiment names, paper/manuscript

- paper/pnas -> paper/manuscript (venue-neutral)
- configs/layer1 -> configs/inheritance, src/knowledge -> src/inheritance
  (imported as `inheritance`), make layer1 -> make inheritance; layer2 alias dropped
- inheritance and trained-network bundles named after the manuscript figure
  they feed (fig2_grounding_sweep, figS3_rebaselining, ...), or descriptively
  where they feed none; configs keep their `experiment:` value so parquet
  hashes are unchanged, only output.dir moves
- figure scripts, SI figure sources, notebooks, REPRODUCING.md, README and the
  SI Methods/tables updated; make clean no longer deletes tracked manifests;
  reproduce.sh hashes the s{seed}/ layouts too

Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01Y64o8FKP7rCuXzC48pxpMm
This commit is contained in:
Giorgio Gilestro 2026-09-13 17:00:40 +01:00
parent 84124de143
commit ab3dc10587
240 changed files with 477 additions and 476 deletions

View file

@ -0,0 +1,34 @@
# E6 — Re-minting is irreversible; gate it on diversity
**Claim tested:** what happens if you "re-baseline" — declare the current model's output to be the new
ground truth and throw away the original? If you do this while the model is already collapsed, is the
damage permanent? And can a simple safeguard prevent it?
**Setup (Layer 1, pure math).** `K = 500`, `n = 200`, 400 generations, 100 repeats. **Re-minting**
periodically freezes the current distribution as the new grounding reference and *discards the
original truth* (it survives only as a yardstick for measuring drift). Four arms:
- **healthy re-mint** — generous grounding (`m = 60`), re-mint while still diverse;
- **collapsed re-mint (ungated)** — starved grounding (`m = 1`), re-mint anyway;
- **collapsed + diversity gate** — same starvation, but only re-mint if diversity `H ≥ 0.75`;
- **collapsed, no re-mint** — the baseline.
### Symbols
- **re-mint** — adopt the current model's output as the new "reality" and discard the original truth (a founder event).
- **diversity gate** — refuse to re-mint while `H` is below a threshold (here 0.75).
- **forward-KL to ORIGINAL truth** — how far the lineage has drifted from the *real* original, even after it changed its own reference.
### The two panels
1. **Lock-in.** Forward-KL to the *original* truth over generations; dotted verticals mark re-mint
events. Red (collapsed, ungated) **jumps up at each re-mint and never comes back** — once the
original tails are gone, re-baselining onto the impoverished distribution makes the loss permanent
(they can no longer be grounded back). Green (healthy) and blue (gated) stay low; grey (baseline)
is the reference.
2. **What the gate reads.** Diversity `H` over generations, same colour key, with the gate threshold
(`H = 0.75`, dashed). The gated arm simply **refuses to re-mint while below the line**, so it never
locks in a collapsed state; the ungated collapsed arm re-mints into the floor.
### Takeaway
Re-minting a collapsed model **crystallises** the collapse — it is a one-way door. A trivial
safeguard (only re-baseline when diversity is still high) preserves recoverability; re-minting a
healthy model is harmless. **Falsifier (not triggered):** if the collapsed lineage had recovered its
original tails after re-minting, the irreversibility claim would be overstated.