← Complete research archive
Evaluation & auditsClosed / no-go51 lines

R12 CDRL Neural Optimization Result

Job: Newton 691750 on evc22, exit 0:0, elapsed 00:03:25. Decision SHA-256: ad94ac15ca17eaa2c5381aa0a3f94fc60a49dbbf2a528552a1212b3ecf1cabdb

R12_CDRL_NEURAL_OPTIMIZATION_RESULT.mdOpen original Markdown ↗

R12 CDRL Neural Optimization Result

Status: advance=false. Conjecture C is rejected on the frozen R12-CDRL-NEURAL-v1 board. Not a Shohin, ACW, or reasoning result.

Job: Newton 691750 on evc22, exit 0:0, elapsed 00:03:25. Decision SHA-256: ad94ac15ca17eaa2c5381aa0a3f94fc60a49dbbf2a528552a1212b3ecf1cabdb

Locked outcome

Median depth-OOD margins (core - control), required >= +0.05:

MarginMedianGate
core − full-0.7759FAIL
core − rand-0.0205FAIL
core − hard-0.7783FAIL

Depth-OOD exact state accuracy by seed/arm:

Seedcorefullhardrand
20260716010.04490.87740.82320.0586
20260716020.04200.81790.92480.0625
20260716030.03810.70410.75490.0835

Interpretation

Core-only supervision learns a residual predictor that never sees identity padding P in training, then fails when depth-OOD evaluation restores full distractor-laden histories. Random length-matched subsequences fail similarly. Full-history and hard-mined full-history arms both solve the board.

This rejects pure Nerode-core allocation as stated in Conjecture C under the frozen eval contract (train allocation may differ; eval is always on full histories). It does not reject mixture curricula, ACW/CGBR collision injection, or other residual-transport mechanisms.

Non-claims

  • No Shohin adapter was trained
  • No ACW Track S/C custody bytes were touched
  • No language or autonomous-reasoning claim

Next

Close Conjecture C. Preserve artifacts under artifacts/r12/cdrl_neural_v1/. Do not retune thresholds or re-run with altered eval. Any successor must be a new preregistration (for example a mixture core∪full arm with matched label budget).