One continuation step fails preservation
0.1 · Public research record
- Question
- Can one bounded DPO update improve evidence acquisition while preserving the incumbent?
- Changed variable
- One full-response DPO update on eight fixed preferences; learning rate 1e-6, beta 0.1, microbatch 1 and accumulation 8, fresh paged AdamW8bit, clipping 0.5.
- Key result
- Original matched accounting: 244→224/300 (−20). Historical Q1 was 242/300. Experiment 0051 subsequently established canonical Q1=243/300 and candidate=224/300 (−19). All 30 baseline behavior hashes were unchanged across the historical/current comparison. Candidate changed 14/30 benchmark behaviors; format failures rose 0→1. Heldout implicit acquisition rose 1/3→2/3, missing-case fabrication stayed 3/6, and grammar failures rose 1→3. Adapter L2 displacement was 0.006050562709, relative 0.01555298%; all 560 tensors moved.
- Verdict
- Rejected. A narrow acquisition gain did not offset preservation failures; no qualification or promotion.
- What it ruled down
- This exact single-update recipe is not preservation-safe on the tested panels.
Read full record · 0050
Question
Can one bounded DPO update improve evidence acquisition while preserving the incumbent?
Why it mattered
The protected 0035/Q1 incumbent is 243/300 under the explicit current canonical accounting. This experiment tested the bounded continuation whose failure motivated the subsequent forensic sequence. A useful successor must preserve behavior as well as improve the targeted boundary.
Setup
Granite 4.1 8B coordinator with a LoRA adapter; the experiment used the frozen incumbent or its stored artifacts. Results describe the recorded experimental runtime and panels, not universal model behavior.
Fixed inputs
Frozen incumbent identity, source-bound experimental inputs and the relevant evaluator/measurement protocol were preserved. The shared eight-pair preference set and later reviewed span masks were held fixed where applicable. Benchmark text, exact prompts, token identities, tool schemas and raw completions are withheld to protect evaluation integrity.
Changed variable
One full-response DPO update on eight fixed preferences; learning rate 1e-6, beta 0.1, microbatch 1 and accumulation 8, fresh paged AdamW8bit, clipping 0.5.
Measurements
Matched 30-case benchmark, separate 18-case synthetic heldout, adapter displacement, native grammar and source-grounding judgments.
Result
Original matched accounting: 244→224/300 (−20). Historical Q1 was 242/300. Experiment 0051 subsequently established canonical Q1=243/300 and candidate=224/300 (−19). All 30 baseline behavior hashes were unchanged across the historical/current comparison. Candidate changed 14/30 benchmark behaviors; format failures rose 0→1. Heldout implicit acquisition rose 1/3→2/3, missing-case fabrication stayed 3/6, and grammar failures rose 1→3. Adapter L2 displacement was 0.006050562709, relative 0.01555298%; all 560 tensors moved.
Verdict
Rejected. A narrow acquisition gain did not offset preservation failures; no qualification or promotion.
What this ruled out
This exact single-update recipe is not preservation-safe on the tested panels.
What remained unresolved
Historical loop-accounting disagreement; individual causal contributions of objective, optimizer and numerical execution; cross-seed and wider generalization.
Next experiment
0051: inspect stored update dynamics and resolve accounting without new model execution. This is chronology of the recorded research program, not authority to execute new work.
Reproducibility notes
Frozen policy/reference identities, training membership and singleton update were verified. Five equal numeric replays per model establish deterministic score reconstruction, not five fresh semantic judgments. Findings were independently reviewed within the experiment workflow. That review was internal and does not imply external peer review. The public aggregate data and analysis-only renderer support checking the displayed figures. Model artifacts and protected evaluation inputs remain withheld; any future artifact release requires separate license and disclosure review.
Public artifacts
This public research record includes the sanitized summary and the source-grounded aggregate figures/data used by the flagship article. Figure inputs, the analysis-only renderer and plotting dependencies are available as aggregate data, renderer and requirements. VEZRYN LLC retains all rights to its authored material; third-party components retain their own licenses. Public accessibility does not change the experimental verdict or imply model promotion. Raw reports, prompts, benchmark case identifiers, sensitive traces, model artifacts and private provenance are withheld.
Corrections / version history
-
Historical 0.1-draft, 2026-10-06: initial source-grounded summary.
-
Accounting note: preserve historical Q1=242, original 0050 matched Q1=244 and corrected canonical Q1=243 as separate labeled records. The 0050 candidate remains 224; the correction changed accounting of identical saved behavior, not model quality. Later diagnostic experiments did not reevaluate the incumbent.
-
0.1, 2026-10-06: first public release of the sanitized record; experimental findings and accounting unchanged.
Source attribution: VEZRYN LLC, experiment 0050 final report and final independent internal review; private manifest supplies exact bindings. Model, method and frozen-run software attribution appear in the flagship reproducibility section.