Paper 3 · Section 2Learning
What an interface must preserve
Preparations, joint outputs, adaptive actions and retained resources belong to one complete interaction law. This chapter defines the contract and the measurements used throughout the study.
2 Operational contract and measures
2.1 Finite public interaction
A source fixture has a finite hidden carrier , four actions , four joint output symbols encoding two bits, and four public preparation codes . Preparation selects an initial distribution ; it does not reveal the resulting hidden state. Its joint output/successor instrument is
The fixed public clock permits at most eight calls per episode. Resource use, blocked service, mode changes and delayed return are represented by the hidden dynamics and observed outputs. A reset starts another charged preparation episode; it is not an uncharged mid-episode operation.
An experiment specifies a preparation, horizon , and causal policy . The retained history is , with its timed position and the common stopping convention. Its law is
For deterministic policies the policy factor is zero or one. The same policy is used for source and candidate evaluation. Joint outputs are not factorized. Matching a singleton marginal is weaker than matching Equation 2.
A candidate may be a finite unifilar transducer, a tree with explicitly update-closed history registers, or a latent-state model whose public state is a posterior vector. Each is executable, but executable syntax alone does not certify source-faithful prediction. Zero-probability branches need an explicit fallback convention; no equality of conditional laws is inferred at a branch that is impossible in the source.
2.2 Adequacy beyond a finite score
For a fixed admitted family , define the target-relative law discrepancy
The companion paper supplies a sufficient actual-state interface construction: a quotient must preserve protected marks and make
independent of the discarded representative. Under the specified matching causal context, retained joint reference variables, resource ownership and prepared-law pushforward, such quotients admit substitution. Uniform conditional joint-row errors for at most calls and initial discrepancy yield the imported finite-history bound
These are conditional results from [16], not new theorems of this paper. Transferring distance from a resource-restricted simulator family additionally requires transporting that family; target-law fidelity alone is insufficient. We do not repeat those proofs.
The OII learners do not observe or receive a map satisfying Equation 4. Neither a small fitted model nor success on finitely many tests verifies its universal row hypotheses. The benchmark asks how well candidate interfaces predict nominated legal histories, while leaving the stronger source-quotient and native-resource certification obligations open.
2.3 Scoring units
For method , fixture and registered core slot , let be the recorded law error. OII-1–3 use conservative upper values including their stated evaluation allowances; OII-4 uses the full-support floating law evaluation, with separate rigorous event enclosures for selected failure witnesses. The benchmark pass allowance is throughout, an engineering tolerance rather than a physical boundary.
For fixtures with core slots each, the main summaries are
The mean fixture-worst value is not the overall maximum or a bound on Equation 3. An all-core pass means that every slot in that fixture's finite panel passes, not that every legal future context does. Duplicated slots retain their originally assigned weight.
A red-team failure is a fixture for which the privileged postcommit search finds a legal event witnessing error above . An unsafe-merge count instead concerns pairs of equal-clock histories whose true continuation laws differ by more than but share one retained model-state key. The panels and denominators are reported separately. No-witness outcomes are not adequacy certificates, and a belief vector can avoid exact key collisions while still predicting poorly.
| Object | What is established | What is not established |
|---|---|---|
| Model syntax | A normalized executable prediction/update rule | Agreement with the unknown source |
| Core-panel score | Error on registered complete-law slots | Uniform adequacy on all legal policies |
| Source interface theorem | Substitution under marked joint-row and causal-contract premises | Those premises for an opaque learned model |
| Post-reveal diagnostic | A source-relative capacity, obstruction or selection fact | Knowledge or oracle access available to the learner |