R12 Conventional Compiler One-Shot Qualification Result
Decision: qualify_conventional_compiler_for_isolated_stage_b_development
Purpose
The factorized development matrix established that a conventional complete source parser can solve known-atom compositional compilation, but its favorable ordinary arm had been selected on development data. The old factorized confirmation remained sealed because the parameter-islands attribution gate failed. This one-shot qualification asked whether the already-selected ordinary parser reproduces exact compilation on a fresh untouched board without using those confirmation bytes.
Passing qualifies only Stage-A parsing infrastructure for one separately preregistered source-deleted executor/consumer experiment. It is not evidence of execution, recurrence, halting, autonomous rollout, native reasoning, or architectural novelty.
Frozen identities and custody
| Object | Frozen identity |
|---|---|
| raw Shohin 300k base | 211d6b2cddf0c2cf8b12cb0b2d73f9c4440d85f6f531018080c8afd35b2f66a6 |
ordinary adapter from 693049 | 747a559b827c6d114943c091b9dea5b4b90cef7af13aa5003b8435c092d24991 |
| adapter / total parameters | 8,607,886 / 133,689,550 |
| generator/prereg commit before seed | e7fa112 |
| assessor commit before score | 3b3e7e9 |
| qualification seed | 1218705082397710755 |
| board JSONL | 06deeb39ac8c6ceb74003f6e503361401c58ead445558485e30a51d8c6d9358e |
| board report | 1467b089f964b5078f444f5d1c91228dcd3ee0a40792987b319f62dc7e98023d |
| raw result | 05c50c79672cf1b07a42fd02c48b5df84e4d4de87a36dfa637730ac600deccba |
| assessment | 35278899fdbcdf801838c414adf860d59a255ecb4de5a44b3acf072248fa6cc7 |
| job log | 81f80a6ab545520771ce093032a502ebde4b0b9982b728ba2614f776af1320ba |
The fresh board contains 2,048 semantic quartets / 8,192 rows, 4,096 factor
combinations, and 639 source names. All language atoms were known from public
factorized training, while exact prompts, word 13-grams, entity names, and full
factor combinations had zero overlap with public train, compositional
development, and lexical-OOD splits. Every ten-span, two-executor, quartet,
token-bag, and shortcut gate passed. The generator accepts no path to the sealed
confirmation and reports confirmation_access=0.
Execution
Slurm job 693105 completed once on evc25 in 20 seconds with exit code 0:0.
It used the frozen raw-300k base and frozen ordinary adapter with no fitting,
oracle, retry, seed sweep, or alternative arm. The result and report identities
were checked by the assessor committed before the score was read.
Result
| Frozen gate | Result | Floor | Pass |
|---|---|---|---|
| answer accuracy | 8187/8192 = 99.938965% | 99% | yes |
| semantic-program exact | 8186/8192 = 99.926758% | 99% | yes |
| full ten-pointer exact | 8186/8192 = 99.926758% | 99% | yes |
| operation-kind accuracy | 8192/8192 = 100% | 99.9% | yes |
| initial-state joint exact | 8186/8192 = 99.926758% | 99% | yes |
| all-four exact quartets | 2045/2048 | 2000 | yes |
Operation-0 and operation-1 joints were both 100%. All-four answers were exact for 2,046/2,048 quartets. The three program failures are consistent with the remaining initial-binding errors; no operation, literal, query, or kind error was observed.
Consequence
The conventional complete compiler is now qualified as frozen Stage-A infrastructure. The next admissible experiment must keep the compiler and base frozen, gather a bounded model-owned packet, remove source states before any update, and train a separately parameterized recurrent executor and consumer. Gold packets may be diagnostic ceilings only. Host code may route tensor positions but may not decode, correct, or execute semantic fields.
The old factorized confirmation remains sealed/local-only. This qualification does not reopen it and does not establish reasoning.