Code and data associated with "The evolution of sex for artificial intelligence - A population-genetic framework for multigenerational model populations". Gilestro, 2026
Find a file
Giorgio Gilestro 840b6b00b3 Layer 1.5: architecture-general neural existence proof
Re-scopes Layer 2 into a cheaper, architecture-general neural collapse proof
before the LLM rung. Realises the same Wright–Fisher abstractions in real trained
generative models on a fully-synthetic sandbox with an exact oracle, reusing
knowledge.metrics/truth/seeding and the output contract so neural curves overlay
the Layer-1 analytic curves.

  - src/neural/: synthetic token-grammar sandbox (lossless identity + stochastic
    style), ExactOracle, HistogramModel bridge, generation loop, experiment runner
  - HARD GATE passed: histogram lineage reproduces Layer 1 exactly (neutral decay,
    exact H_eq, tracks run_lineage) — tests/test_neural_validation.py
  - torch models: autoregressive RNN + MLP (VAE implemented, not yet fidelity-
    passing); determinism seeding derived from the SeedSequence stream
  - N0 bridge (neural g*=0.047 ≈ Layer-1 0.048), N1 collapse-in-weights, N2 phase
    boundary, N5 architecture-generality (collapse + grounding-rescue in histogram
    + RNN + MLP). Manifests/configs committed; parquet gitignored, hashes tracked
  - additive backward-compatible save_artifacts extension; Makefile neural targets

Finding: neural smoothing partially resists H-collapse, so forward-KL and tail
survival are the sharp neural collapse metrics (H is smooth, per Layer 1).

92 tests green. Remaining (tasks/todo.md): N4 merge, N2 refine, N3/N6, VAE
fidelity, MNIST tier, figures. LLM/LoRA rung and C3 deferred.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-04 21:02:49 +01:00
configs Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
figures Layer 1 complete: E3-E6 + E2 analysis add-ons 2026-07-04 18:54:42 +02:00
paper Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
results Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
src Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
tasks Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
tests Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
.gitignore Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
CLAUDE.md Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
Makefile Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
pyproject.toml Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00
README.md Layer 1 core: Wright-Fisher knowledge-transmission model with E1-E2 2026-07-04 18:10:18 +02:00
uv.lock Layer 1.5: architecture-general neural existence proof 2026-07-04 21:02:49 +01:00

The Lamarckian Society — Layer 1 (analytical core)

A parametric population-genetics model of knowledge transmission across generations of learning agents. Knowledge transmission is modelled literally as a WrightFisher process (not by analogy): a model's knowledge is a distribution p_t over K discrete items; a fixed true distribution p* has a rare tail; each generational step is "sample from the parent (drift) + mix in fresh real samples (grounding/immigration) + refit." Model collapse is the loss of rare alleles under drift.

See paper/blueprint.md (the normative build spec) and paper/the-lamarckian-society-v4.md (the perspective paper).

Reproduce

Environment is a uv venv built from the committed, hash-pinned uv.lock — that lockfile is the single source of truth for "it runs" (Layer 1 is pure NumPy/SciPy and bitwise-reproducible from a seed; no container needed).

# one-time: install uv (https://astral.sh/uv)
curl -LsSf https://astral.sh/uv/install.sh | sh

uv sync                 # build .venv from uv.lock
make test               # correctness + scientific-validation tests (the spine of trust)
make layer1             # run experiments E1E6
make figures            # regenerate figures from committed results

Layout

src/knowledge/   Layer 1 package (imported as `knowledge`)
configs/layer1/  one YAML per experiment (E1..E6)
figures/         plot_EX.py — read results.parquet only
tests/           test_correctness.py + test_scientific_validation.py (analytic checks)
paper/           blueprint.md, perspective paper, figure_manifest.md
results/         written artifacts (gitignored; hashes tracked in manifest.json)