Manuscript revision and pending experiment work, snapshot before restructuring
Clarity pass over the main text (36-item audit), Discussion rewrite and cut, acknowledgements, Souly et al. as ref 62, lettered SI panels, model section moved under Results; plus the untracked curriculum/society/compose/smol configs, runners, figures, stats and tests that the SI already cites. Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01Y64o8FKP7rCuXzC48pxpMm
This commit is contained in:
parent
e4804adabc
commit
84124de143
450 changed files with 52813 additions and 1202 deletions
35
configs/llm/curriculum_v5_decor.yaml
Normal file
35
configs/llm/curriculum_v5_decor.yaml
Normal file
|
|
@ -0,0 +1,35 @@
|
|||
# Decorrelated curriculum (manuscript review, 2026-09-11). In the Latin square partner complementarity
|
||||
# falls monotonically with generation (1.0, 1.0, 0.8, 0.67, 0.33, 0.0), so the veto's acceptance curve
|
||||
# is collinear with adapter age. Here every lineage starts with the same non-destroyer family (mnli),
|
||||
# then diverges maximally, then converges: complementarity 0.00, 0.67, 0.70, 0.58, 0.33, 0.00 by
|
||||
# generation. Same six families, same G, destroyers (boolq, winogrande) spread across lineages as
|
||||
# in the Latin square. Arms: the declinable merge (`society` + `allow_veto`) and its never-merge
|
||||
# reference under the same curriculum. Pre-registered readout: tasks/prereg-llm-society-v4.md §8g.
|
||||
experiment: llm_curriculum_v5_decor
|
||||
kind: llm_curriculum
|
||||
base_model: Qwen/Qwen2.5-1.5B
|
||||
seed: 1
|
||||
families: [mnli, arc, hellaswag, squad, boolq, winogrande]
|
||||
orders:
|
||||
- [mnli, arc, hellaswag, squad, boolq, winogrande]
|
||||
- [mnli, squad, boolq, winogrande, arc, hellaswag]
|
||||
- [mnli, winogrande, arc, hellaswag, squad, boolq]
|
||||
lineages: 3
|
||||
generations: 6
|
||||
arms: [isolated, society]
|
||||
baselines: []
|
||||
allow_veto: true
|
||||
n_new: 300
|
||||
n_replay: 150
|
||||
n_test: 60
|
||||
n_val: 20
|
||||
epochs: 3
|
||||
lr: 1.0e-4
|
||||
operator: linear
|
||||
merge_weights: [[0.5, 0.5], [0.3, 0.7], [0.7, 0.3]]
|
||||
max_new_tokens: 48
|
||||
batch_size: 24
|
||||
train_batch_size: 2
|
||||
train_max_len: 512
|
||||
lora: {r: 16, alpha: 32}
|
||||
output: {dir: results/llm_curriculum_v5_decor}
|
||||
Loading…
Add table
Add a link
Reference in a new issue