← Complete research archive
Compositional lawsAudit63 lines

R12 S3 Training-Lexicon Action Preregistration

Claim class: bounded known-atom compiler repair plus exact source-deleted S3 execution. External schedule and halt remain.

R12_S3_LEXICAL_ACTION_PREREG.mdOpen original Markdown ↗

R12 S3 Training-Lexicon Action Preregistration

Status: public development passed all frozen gates; one fresh confirmation is authorized and not yet scored.

Claim class: bounded known-atom compiler repair plus exact source-deleted S3 execution. External schedule and halt remain.

Falsified interface

Closure-complete S3 v1.2 leaves amount at 100% but direction at 93.403% across the public depth board. A score-blind CPU audit shows all depth direction spans are exact training atoms: six left sequences and six right sequences, with no cross-class collision. The contextual kind head has lost a finite relation already present in the compiler's operation-kind pointer channel.

Sole intervention

A deterministic builder reads only the frozen 96,000-row training split and its gold operation-kind spans. It emits the 12 exact token sequences, their left/right class, and counts, refusing any class collision or non-training row. It records zero development and confirmation access.

At evaluation, the frozen compiler's normalized operation-kind pointer is aligned against those sequences. If one class receives at least 0.5 total pointer mass on an exact occurrence, that categorical class replaces the contextual kind argmax. Otherwise the original neural kind prediction is used. The override cannot inspect development labels, program fields, or answers. Identity, amount, query, exact S3 action, base/compiler/executor weights, boards, and all source-deletion boundaries remain unchanged. No optimizer or parameter is added.

Frozen gates

One zero-fit H100 run must satisfy all of:

  1. the lexicon is training-only, collision-free, 6+6 patterns, and zero-fit;
  2. two-step mean answer/state/chains each >=95%, every surface answer >=94%;
  3. lexical-OOD match coverage <=5% and answer >=75%, proving fallback rather than distractor capture;
  4. depth lexical coverage >=99.5%, direction >=99.5%, and amount >=99.5%;
  5. depth mean answer >=90%, state >=88%, and complete chains >=80%;
  6. depth-eight mean answer >=85%;
  7. depth ordered answer/state/chains each >=98%;
  8. depth gold answer >=98.5%, exact state =100%, and exact chains =100%; and
  9. every output records zero fit updates and zero confirmation access.

Passing authorizes one fresh seed-after-commit confirmation with unseen nonce names, unseen known-atom factor combinations, and causal action/query controls. It would establish only a known-lexeme neural-symbolic execution component. It would not establish unseen-phrase generalization, autonomous planning, learned halt, free-form reasoning, or novelty.

Public closure

Zero-fit job 693136 completed once on H100 evc25 in 2m15s, exit 0:0. Depth mean reaches 94.434% answers / 94.336% state / 89.453% complete chains; ordered reaches 98.340% / 99.463% / 98.730%; gold reaches 98.779% answers and 100% exact state/chains. Direction and amount are 100%. Lexical OOD uses 0% lexicon coverage and preserves the fallback baseline. Assessment 41102547... passes every frozen gate and authorizes one fresh confirmation. See R12_S3_LEXICAL_ACTION_RESULT.md.