Trading Bot HQModels & runsTestsResearch

D4 pre-read (draft v1) — for the 2026-10-22 quarterly review

Drafted 2026-07-25, three months early and deliberately so: the evidence is fresh, and drafting the review before the intervening quarter's results exist is itself a pre-commitment — October's session should diff reality against this document, not compose a narrative after the fact. Per Research Program §6 D4: reviews spend against §3 budgets, D1–D3 status, and whether the program document still describes reality.

1. Budget status (as of this draft)

Budget Spent Remaining Notes
Hypotheses (2,000 lifetime) 223 ledger (+ prereg-side execution hypotheses, reconciliation in progress) ~1,770 11% spent in ~2 weeks of active research; at this pace the budget is a ~6-month constraint, which is fine — D2's floor is 1,000
Validation looks (10) 1 (run 20, FAIL) 9 Look 2 decision pending (streak candidate memo)
Vault 0 contact sealed Runner-level assertions in place

2. What the quarter established (headline evidence, citations in RESEARCH-FINDINGS)

  1. The machinery is done and trustworthy. Edge Lab (bit-exact acceptance), level engine (causality-proven), paper loop (streaming==batch), MC harness (block-only), prereg template with ex-ante cost kill rule, 199-test suite, public evidence site. The project's marginal cost of a new rigorous experiment is now ~hours.
  2. The mentor's mechanizable claims failed comprehensively (0-for-5 on direction: entries, direction logic, stop-structure fit, level reactions, H4 ignition). What remains of Track M value: his event concepts seeded two real research objects (displacement, streaks), his session observation matched measured structure, and his risk discipline informed the executor design. D1 (October 31) should be assessed against this base rate: even verified mentor data is unlikely to yield a mechanized edge directly; its realistic value is calibration (original stop placements for the B4 question) and cost truth (his fills).
  3. One effect family is asset-class structure (displacement continuation: 4/5 instruments, XRP cleanest) — but its BTC execution died on the Validation holdout (regime inversion post-2022). The event class is real physics; harvesting its direction is regime-fragile.
  4. One candidate is alive but weakened (up-streak persistence: BTC execution BAR MET; instrument map patchy 2/5 decisive + XRP inversion; regime conditioning ambiguous/underpowered). Look-2 decision rests with Richard.
  5. Cost structure is destiny. Vantage-class BTC ≈ 3.1bp is the only venue/instrument combination where any tested effect clears its own drag; ETH (22.6bp) killed a registration by arithmetic alone. Any instrument-expansion ambition runs through sourced cost models first.

3. Questions D4 must actually answer in October

4. Standing risks to name at the review

5. Pre-committed for October

This draft is the baseline. The October session diffs: budgets vs §1, D1–D3 status vs §2–§3, and answers Q1–Q4 with the intervening quarter's registered evidence only. If this document's framing turns out wrong, the review says so and amends the program — it does not quietly rewrite this pre-read.