{ "output_dir": "/home/infra-ops/erp-tune/run-05", "roots_dir": "/home/infra-ops/erp-tune/datasets/derived", "base_model_path": "/home/infra-ops/models/gemma4-26b-a4b-it-bf16", "base_model_revision": "google/gemma-4-26B-A4B-it (stock instruct, upstream) -- UNCHANGED from runs 2, 3, 3c and 4. Run 5 varies the 20% slot's COMPOSITION only: airoboros-3.2 OUT, govreport + qmsum IN, at run 4's lr 2e-04. Base byte-identical to run 4 (index sha 907826a6...).", "recipe": "/home/infra-ops/erp-tune/recipe-r5/recipe-erp-seat-sft-r5.json", "survivors": "/home/infra-ops/erp-tune/recipe-r5/survivors-r5.jsonl", "chat_template_path": "/home/infra-ops/models/gemma4-26b-a4b-it-bf16/chat_template.jinja", "impersonation_mask_path": "/home/infra-ops/erp-tune/recipe-r3/lossmask-r3.jsonl", "lora_rank": 64, "lora_alpha": 128, "lora_dropout": 0.0, "max_seq_len": 16384, "epochs": 1, "seed": 20260824, "per_device_batch_size": 2, "gradient_accumulation_steps": 8, "learning_rate": 0.0002, "warmup_ratio": 0.1, "lr_scheduler_type": "cosine", "weight_decay": 0.01, "load_in_4bit": false, "gradient_checkpointing": true, "loss_chunk_tokens": 1024, "training_eligibility_override": "operator-2026-09-07-rnd-run5", "overridden_blockers": [ "contamination-scan-not-implemented", "stage-2-csam-detector-inert" ], "substitute_controls": [ "pre-training holdout, run-1 (8,404 samples, work/card/session split)", "pre-training holdout, govreport/holdout-v1 (416 reports, sha256-ranked, never_trained_on)", "pre-training holdout, qmsum/holdout-v1 (5 transcripts, sha256-ranked, never_trained_on)", "stage-A lexical quarantine, RP (829 records held unread)", "stage-A lexical quarantine, run-5 slot (133 records held unread, /mnt/smithy/datasets/quarantine/r47-run5-longdep-screen/)", "SCROLLS-membership disclosure on both slot sources (avoidance, NOT a scan): govreport + qmsum are SCROLLS/ZeroSCROLLS members, in no hoard/default-benchmarks.yaml entry and used by no R47 instrument", "SINGLE VARIABLE vs run 4: the 20% slot's COMPOSITION changes -- airoboros-3.2 is REMOVED and govreport (496 reports) + qmsum (97 transcripts) are ADDED at run 4's realized slot ctx (3,723,090 vs 3,720,956, +0.06%). Dialogue survivors (survivors-r3 selection) and the impersonation loss-mask (lossmask-r3) are held BYTE-IDENTICAL to runs 3/4; fireball whole; kvasir at run 4's EXACT 1,613-sample prefix cut (survivors reused from survivors-r4, NOT re-cut); base, lr 2e-04, max_seq_len 16384, rank 64, alpha 128, dropout 0.0, cosine, warmup 0.1, wd 0.01, batch 2 x accum 8, 1 epoch all UNCHANGED.", "kvasir is HELD, not re-cut: the 1,613 kvasir survivors are reused verbatim from survivors-r4.jsonl (which cut run-3's seed-20260824 prefix at 3,347,622 ctx). survivors-r5.jsonl = survivors-r4 minus airoboros plus the govreport + qmsum roots whole; sha256 a25169a6258cd4abb0cb494a176a921c0e98eb73d65c53d033b6ee18293a43ae.", "window_count belt-and-suspenders (SFT-RECIPE-run5-SCOPE.md 7.1): every govreport + qmsum row renders <= 14,000 tokens (max 9,385 / 13,700) and the harness never packs across samples, so window_count MUST be 1 on every slot row; a chunked_into_2 or single_window_truncated on either new root in truncation-report.json is a BUILD DEFECT and the run is killed before training.", "HOST: pfi-gx10 (GB10, aarch64, sm_121, 121 GB unified). Base model sha256-verified identical to ana-ml2's copy; new-root shard sha256 verified against CLEANROOT after transfer (govreport d533a5cd, qmsum c0269d69); harness eitri-smithy 0a6bd2e; corpus COPIED, box mounts no NFS." ], "unfittable": "drop", "holdout_dir": "/home/infra-ops/erp-tune/datasets/holdout", "save_steps": 50 }