Receipt-Mediated Terrain Learning: Concentration as an Anti-Cheat Signal

    Breyden E. Taylor · Prompted LLC · 24 July 2026

    cs.LG · cs.AI · 9 pages · 2,372 words · depends on P02, P04 · gates P10

    Abstract

    A governance terrain should move because behavior occurred, not because an agent declared a desirable lane, repeated the right vocabulary, or warmed an entire embedding space. We define receipt-mediated terrain learning: only behavior entering a source-disjoint receipt corpus can move the learned field, and legitimate motion should concentrate on the patterns the behavior actually touched. Uniform or declared-lane-correlated warming is an anti-cheat failure. The originating implementation observed a 3--4x concentration on four touched patterns relative to an untouched control in one trainer pass and correctly withheld promotion. We formalize repeated-pass experiments with lexical stuffing, declared-lane manipulation, behavior without receipts, generic activity, and counterexamples. The empirical claim remains tentative until sustained independent conformations accumulate.

    Claim register

    • P09-C1 — formal/proposed · reopens: new replication, falsifier, source correction, or configuration change
    • P09-C2 — testable hypothesis · reopens: new replication, falsifier, source correction, or configuration change
    • P09-C3 — implemented-or-pilot bounded · reopens: new replication, falsifier, source correction, or configuration change
    • P09-C4 — external validation required · reopens: new replication, falsifier, source correction, or configuration change
    pdf sha256
    6502b8e6f8f54d0e6938c2c3bc059da63ca5d85f5ba5c705c516f39480c18a04
    src sha256
    8453f6c3241b91474b21877011ffa81692b383d7df521c588688c5fdaae118bc
    md sha256
    ee86d6a7c5dd31fb45c9fb6eda4aafc488d5bda5e67357d3518efdda64735f35
    Reading scale

    Proposed arXiv categories: cs.LG / cs.AI Publication DAG node: P09, Wave C Dependencies: P02, P04 Claim register: P09-C1, P09-C2, P09-C3, P09-C4

    1. Thesis and scope

    Learning becomes governable when the system can prove what behavior the update saw, which terrain changed, which controls did not, and why a single favorable pass did not rewrite the earned center.

    We contribute a source-disjoint training contract, concentration ratio, false-flattery controls, value-provenance schema, maturity gate, and promotion rule. The method distinguishes measured, supervised, unsupervised, and authored lanes instead of pretending all terrain coordinates have the same epistemic origin.

    FORKED is an internally owned and managed estate inside the Ubiquity Federation, Prompted LLC's flagship governance substrate. The estate inherits the Fractal Quivers of Quivers parent topology, production splat mechanics, center exclusion, receipt-bearing lawful traversal, lifecycle separation, and the rule that models propose morphisms while governance determines admissible motion [@Taylor2026FQoQ; @Taylor2026OpenCenter; @Prompted2026Ubiquity]. This paper studies one bounded expression of that parent architecture. It neither renames the parent substrate nor inherits universal applicability from it.

    The evidence lanes are kept separate. A canonical or source artifact establishes what that artifact states. A component test establishes behavior under its recorded configuration. A synthetic campaign establishes behavior of the generator and analysis pipeline, not human, hardware, customer, field, or operational performance. Novel cross-domain claims are preregistered as hypotheses and remain open to null, adverse, localized, and disproving results.

    Dependency freeze. Final freeze depends on P02, P04; draft work may proceed in parallel, but imported claim IDs and artifact hashes must be rechecked after each dependency freezes.

    2. Interlattice position

    The interlattice is the citation and governance relationship among the parent FQoQ topology, production splat mechanics, the FORKED foundational architecture, the Native Range, the Experimental Program, and the companion papers in this DAG. Citations are therefore typed: a parent citation supplies inherited architecture; a native citation supplies component behavior; a synthetic citation supplies generator-level evidence; a companion citation supplies a versioned specialized argument. No downstream citation converts an upstream hypothesis into a universal fact.

    This paper's claims are:

    • P09-C1: the central mechanism is formally representable and testable.
    • P09-C2: the specified controls can distinguish the mechanism from simpler alternatives.
    • P09-C3: the recorded implementation or synthetic evidence establishes only its declared lane.
    • P09-C4: external applicability requires the preregistered receipts and remains domain-bound.

    3. Problem statement

    Modern AI systems frequently compress unlike objects into a common output surface. Facts, hypotheses, instructions, permissions, model predictions, runtime observations, and institutional decisions can all arrive as fluent text or a single status field. The resulting failure is architectural before it is linguistic: a local representation acquires authority that its source never possessed. Learning becomes governable when the system can prove what behavior the update saw, which terrain changed, which controls did not, and why a single favorable pass did not rewrite the earned center.

    4. Parent architecture

    Terrain shape

    Past failures should influence present decisions when their structure conforms to the present terrain. They should not dominate merely because they are recent, emotionally vivid, institutionally notorious, or easy to name.

    FORKED separates two variables:

    memory salience = how available a precedent is to the observer
    precedent standing = how much lawful influence the precedent may exert now

    A failure can be salient and irrelevant. Another can be old and structurally load-bearing. The architecture therefore retrieves precedents through a staged gate:

    1. APO veto. Remove candidates whose forbidden-path geometry conflicts with the current splat.
    1. Causal topology. Compare dependency structure, authority transfer, omissions, correction paths, and failure sequence.
    1. Trajectory phase. Compare whether the present system is entering, normalizing, depending on, defending, rupturing, or correcting the pattern.
    1. Horizon. Compare task, mission, battle, campaign, war, institutional, or civilizational scope.
    1. Point of no return. Compare what becomes irreversible and how close the current path is to that surface.
    1. Observer and agency geometry. Compare who could see, who shaped the view, and whether challenge remained meaningful.
    1. Currentness. Determine whether mechanism, law, environment, technology, authority, or physical conditions have changed.
    1. Working metric. Only after the structural gates survive may an embedding or distance function rank candidates for inspection.

    The internal terrain lineage contains a direct example of why this matters. A lexical retrieval matched the word used to describe a problem and selected the wrong doctrine. A six-ray shape strike with an APO veto selected the correct governing pattern instead. The resulting principle is explicit: rehydration is shape- and field-derived, not keyword-derived. [@TerrainDoctrine2026; @Taylor2026OpenCenter]

    A generic scoring form is:

    Standing(f | x) = V_APO(f,x) * [
        w_s * ShapeConformation(f,x)
      + w_t * TrajectoryPhase(f,x)
      + w_h * HorizonMatch(f,x)
      + w_p * PONRMatch(f,x)
      + w_o * ObserverAgencyMatch(f,x)
      + w_r * ReceiptQuality(f)
      + w_c * CurrentnessModifier(f,x)
    ]

    V_APO is a veto, not a low weight. Currentness is bounded. The architecture may increase or decrease a candidate's standing because the world changed, but recency cannot independently manufacture conformation.

    The native terrain campaign makes this policy executable: an older shape-conforming precedent outranked a recent shape-distant precedent in all 500 generated trials. That is a receipt of the current implementation, not a universal theorem about all memory systems. [@ForkedCampaign2026]

    Training cables

    A terrain learner should not learn that a phrase is good. It should learn which path conditions, exclusions, authority patterns, and correction responses recur when the system behaves well or fails.

    The current terrain doctrine establishes two implementation principles at different evidence levels. First, receipt-mediated visibility is a runtime design rule: if behavior does not enter the behavioral receipt corpus, the learner cannot claim it observed the behavior. Second, concentration as anti-cheat is a bounded observation from one trainer pass: touched patterns moved three to four times more than an untouched control while the false-flattery guard passed. Promotion was correctly withheld at n=1. [@TerrainDoctrine2026]

    The broader training objective is cable preservation:

    same surface + different terrain      -> distinguish
    different surface + same terrain      -> rehydrate
    same evidence + different authority   -> route differently
    same facts + different horizon        -> preserve externalities
    same fable + wrong topology           -> deactivate
    high confidence + missing receipt     -> do not promote
    low rhetoric + complete hard receipt  -> do not block for lack of fluency

    Training may adjust working affinity, retrieval prominence, inspection cost, or model routing. It may not learn the held-open center, grant itself authority, or promote doctrine from a single favorable pass.

    Native shape benchmark

    The campaign ran 500 paired cases. The older shape-matched memory ranked above the recent shape-mismatched memory in all 500.

    Capability boundary

    The current evidence establishes:

    • the parent-to-estate boundary is executable;
    • the held-open center is structurally excluded in the native objects;
    • the natural-language projection remains bound to the lattice;
    • the shape-priority policy is implemented;
    • suspension has a governed exit;
    • nominal agency and absent authority cannot pass the tested declaration paths;
    • hidden joins can block stabilization in the encoded graph family;
    • correction fan-out is represented;
    • source artifacts and derived results are hashed and reproducible;
    • HUNGER, biome, sensory, cache, lifecycle, and economy lanes execute locally.

    It does not establish:

    • human behavioral effects;
    • physical carrier performance;
    • field sensor performance;
    • hardware isolation or root-of-trust properties;
    • universal threshold calibration;
    • customer economics;
    • strategic or operational outcome improvement;
    • cultural universality of fable terrain;
    • canonical Ubiquity absorption of the local package.

    This is not hedging where capability exists. It is the exact scope of the receipts.

    5. Formal model

    Let Bt be the behavioral receipt corpus at time t, D a source-disjoint declared-lane corpus, and Δai the affinity change for pattern i. A valid trainer update requires that the learner consume Bt without reading D as behavioral evidence. For touched set T and untouched controls C, define concentration \[ \kappa=\frac{\operatorname{mean}_{i\in T}|\Delta a_i|}{\operatorname{mean}_{j\in C}|\Delta a_j|+\epsilon}. \] Concentration alone is not proof: a learner can overfit a touched label. The full test includes source disjointness, directionality, untouched and adversarial controls, counterexample response, and longitudinal persistence. Values must carry provenance fields for physical meaning, observable trace, transform, source mode, and confidence.

    Promotion is a lifecycle transition, not a gradient threshold. A candidate can move at one pass while remaining tentative. Sustained terrain requires repeated independent conformation, bounded drift, and survival under correction.

    6. System and study design

    The benchmark runs repeated governed work episodes. Conditions include receipted target behavior, identical behavior without receipt, declared-lane edits without behavior, lexical stuffing, generic high activity, wrong-pattern behavior, and corrective counterexamples. Each episode is followed by terrain reconstruction and blinded evaluation on held-out tasks. The learner is compared with ordinary fine-tuning, embedding updates, declared-profile weighting, and no-learning baselines.

    Primary evaluation asks whether changes are specific, source-disjoint, behavior-linked, persistent when warranted, and reversible under counterevidence. The study preregisters the maturity gate before any run so favorable motion cannot redesign the admission rule after the fact.

    6.1 Controls and ablations

    • Behavior without receipt
    • Receipt without behavior
    • Declared-lane manipulation
    • Lexical stuffing
    • Generic activity
    • Untouched controls
    • Wrong-pattern controls
    • Counterexample correction

    6.2 Primary outcomes

    • Concentration ratio
    • False flattery
    • Source contamination
    • Held-out behavior improvement
    • Sustained crossing
    • Drift
    • Correction responsiveness
    • Promotion accuracy

    7. Current implementation and pilot evidence

    The terrain doctrine reports one trainer pass in which four touched patterns moved by +0.0135 to +0.0175 while an untouched control moved +0.0013; the source-disjoint guard remained passing, and promotion was withheld because n=1 [@TerrainDoctrine2026]. The native shape-priority benchmark verifies the current retrieval policy across 500 generated pairs [@ForkedCampaign2026]. The concentration observation remains explicitly tentative and is not promoted by this paper.

    The paper treats this evidence directly where capability is established and narrowly where it is not. Implemented code paths are called implemented. Passing component tests are called passing component tests. Synthetic estimates remain synthetic. Human and field effects remain hypotheses until their own evidence exists.

    8. Preregistration

    The formal preregistration artifact distributed with this paper freezes the primary claims, outcomes, controls, exclusion rules, analysis family, null interpretation, and release boundary before external data collection. The core sequence is:

    1. freeze source and configuration manifests;
    1. freeze primary hypotheses and adverse outcomes;
    1. generate or acquire data without changing the admission gate;
    1. run the declared analysis and publish all primary results;
    1. route deviations to an explicit exploratory appendix;
    1. demote, localize, or reject claims when falsifiers fire.

    8.1 Analysis plan

    The preregistered primary test compares concentration in the receipted-behavior condition with every cheat control. We report raw per-pattern movement and distributional effects, not only κ. Longitudinal analysis estimates whether repeated valid behavior produces stable terrain while counterexamples cause localized correction. Promotion decisions are evaluated separately from learning quality because a cautious lifecycle can be correct even when movement is real.

    8.2 Null and adverse-result handling

    A second pass that does not reproduce concentration leaves the mechanism tentative or disproves it under the tested configuration. Uniform warming fails even if task performance improves. Behavior without receipt producing equal motion defeats the visibility claim. A correct withhold is not counted as false negative when maturity criteria are unmet.

    9. Falsifiers

    • Concentration fails to replicate
    • Declared labels move terrain equally
    • Lexical stuffing reproduces effect
    • Untouched controls warm similarly
    • Counterexamples do not reverse localized terrain

    10. Limitations

    Receipts can be incomplete, gamed, or costly. Pattern taxonomies and labels may encode the author's desired result. Parameter changes can interact with retrieval and evaluation in ways that look like learned terrain. Independent replication and frozen gates are necessary.

    11. Security, ethics, and release boundary

    This public paper is limited to benign assurance, provenance, human-agency preservation, lawful state transition, simulation, test infrastructure, and defensive resilience. It does not disclose operational target-selection logic, engagement optimization, platform-specific exploitation thresholds, live deception procedures, signature-emitter recipes, or methods that materially increase harmful capability. Any later operational research requires separate lawful authority, ethics and safety review, configuration control, and release adjudication.

    12. Required next receipts

    • Multiple trainer passes
    • Independent implementation
    • Frozen preregistration
    • Blind held-out task authors
    • Longitudinal correction corpus

    13. Conclusion

    Learning becomes governable when the system can prove what behavior the update saw, which terrain changed, which controls did not, and why a single favorable pass did not rewrite the earned center. The contribution is not a claim that every domain should adopt one representation, threshold, interface, or governance stack. It is a falsifiable architecture and experiment package for determining where this mechanism carries load, where it inverts, and where a simpler system should win.

    Data, code, and provenance availability

    The submission bundle includes the manuscript source, compiled PDF, bibliography, preregistration, claim register, controls and null-handling document, release boundary, source manifest, configuration digest, dependency contract, figure source, and arXiv source archive. Internal source artifacts are cited by immutable digest where available. Restricted operational material is not included.

    Acknowledgments and authorship

    Breyden E. Taylor is the author and bears responsibility for the thesis, terminology, claims, boundaries, and decision to publish. Homeskillet assisted with source synthesis, drafting, artifact generation, typesetting, and build inspection. The Ubiquity Federation supplied prior artifacts, runtime implementations, and receipts. AI assistance is production history, not evidence.

    References

    PREPRINT — this is an author preprint of a paper scheduled for arXiv submission. It has not been peer reviewed and carries no arXiv identifier until announcement. Claims are preregistered and domain-bound; no paper inherits universal applicability from the parent whitepaper or from a companion paper.