Balance-aware retrospective portfolio diagnostic
Evaluate the eligible survivor across independent 10M, 100M and 1B purses.
Typed non-result: no eligible candidate stream and no three-balance retrospective result.
Every terminal canonical finding this project has published, in one list. Negative results and inconclusive measurements are included, not only the ones that came out flattering. A finding that has been superseded by a later, corrected result is not deleted. It stays reachable at its own permanent address, linked from the result that replaced it, but it is not listed by default here.
Verdicts are published exactly as the research index records them, machine vocabulary and all. Where a verdict uses a word that means something specific here, open the plain-language note under it.
Balance-aware retrospective portfolio diagnostic
Non-result
Typed non-result: no eligible candidate stream and no three-balance retrospective result.
No balance result or live-profitability claim can be computed without an eligible stream.
COMMON
Evaluate the eligible survivor across independent 10M, 100M and 1B purses.
Typed non-result: no eligible candidate stream and no three-balance retrospective result.
Whether the same-day return-reversal signal is mechanical bid-ask bounce or a real mean-reversion effect.
1d — GENUINE — the signal survives measured on a single side (mid -0.29, bid -0.39, ask -0.27; touch retention 113%), so it is not bid-ask bounce. 2h — INCONCLUSIVE — the signal partly survives measured on a single side (mid -0.36, bid -0.30, ask -0.20; touch retention 71%). Some is bounce and some may be real; this test cannot apportion it. Return-based models stay blocked [at 2h].
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Determine whether narrow export re-requests can deterministically repair historical full-book gaps.
Narrow re-requests are not deterministic historical-gap recovery; prospective recording is the valid execution-grade route.
Report present/absent/failed chunk accounting for the /export acquisition.
Operational acquisition-status snapshot (present/absent/failed chunk accounting), not a fixed research verdict. Figures move as the /export acquisition proceeds; treat as current-as-of-generation only.
Compute fixed-cohort, event-indexed median relative-price-change aggregates for admitted patch-notice events.
287 cohort-event rows cleared the minimum-sample and promotion gate out of 182 admitted events x their matched cohorts (9 cohorts, partition tradeable-v1-b3bfe7781e). Each row is one event's cohort-wide MEDIAN relative price change; no per-item value, extreme or cumulative statistic is ever computed or published. A supporting dataset for the optional patch-events route, not a hypothesis-test verdict.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Whether a fitted model forecasts the spread better than a naive persistence null.
Pooled R2 beats the harder null (+0.0729 vs -0.0666) but per-item (equal-weighted and notional-weighted) R2 LOSES to it (-0.4959 vs -0.3538; -1.6372 vs -0.4366). Reporting only the pooled figure 'would have been selection dressed as a result.'
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Whether any resting-order cycle clears its costs.
CONCLUSIVE NEGATIVE — under a touch-only fill rule, which is a proven upper bound on what could have filled, the median resting cycle does not clear its costs (median net capture per attempted cycle -0.530%, 40.7% of items above zero). Real fills are a subset of these, so the true result can only be worse. No depth data can rescue this configuration.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Decompose Gate A's long-patience positive capture into drift share vs. genuine per-item residual, graded against ADR-0004's pre-registered criteria.
GENUINELY EARNED — 1d: drift share 0.084 is below DRIFT_SHARE_THRESHOLD = 0.5 and the median per-item residual +7.170% is above zero ... ; 2h: drift share -0.062 is below DRIFT_SHARE_THRESHOLD = 0.5 and the median per-item residual +10.819% is above zero ... . Graded against ADR-0004's table, at 7-day patience under touch_only with the artifact guard on, 0bps from the weighted quote reference.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Determine whether recorded depth can resolve 3- and 7-day calibration attempts.
Every 3- and 7-day cell is not yet measurable; the result is immature coverage, not a failed or extrapolated calibration.
Whether Gate B's calibration verdict holds on a self-built, growing bar sample instead of the static panel.
[sample: self-built bars (recorder snapshots, 2h grid)] NO OVERSTATEMENT MEASURED — the median item's bar-derived fill count differs from recorded depth by +0.00 percentage points of attempted cycles, and the two agree at the median; 48% of items overstate. ... This is a FIRST calibration over 51.9 hours of recorded depth, 413 items, 7,787 resolvable attempted cycles, not a settled constant.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
How much bar-derived fill detection overstates real fills, calibrated against recorded order-book depth.
NO OVERSTATEMENT MEASURED — the median item's bar-derived fill count differs from recorded depth by -14.29 percentage points of attempted cycles, and the bar panel MISSED fills the recorded book shows, which is a defect in the opposite direction and is not a reason for confidence; 28% of items overstate. Under the touch-only rule ... bar data is fit for purpose for fill detection at this configuration. This is a FIRST calibration over 51.9 hours of recorded depth, 400 items, 2,770 resolvable attempted cycles ... not a settled constant.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Report v1's observed vs. elapsed forward-run cadence coverage -- explicitly not an expected-value verdict.
v1 spoke in 3 of 3 elapsed windows (100.0%; 2 with a complete cohort, 1 torn, 0 deliberately empty, 0 with no record) at an assumed 24h cadence anchored 2:00 UTC, over 2026-08-07 00:45 -> 2026-08-09 02:00 UTC. No outages: every one of the 3 elapsed windows has a record. (measured live at 2026-08-09T22:59:17Z)
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Whether the 2h drift residual's excess capture is explained by selection on wider fill spreads, or is a defect.
Completed cycles select NARROWER, not wider, BUY-fill spreads (median fill-vs-item spread delta -0.403pp at 1d, -1.126pp at 2h); no evidence of an inverted book side, mismatched item, or entry/exit join defect.
Determine whether a preregistered item-selection grid produces an eligible real-candidate strategy with an interior optimum.
Terminal outcome: `grid_edge_unidentified`. The winning vol_7d / 30-day / top-40 cell had +0.9155% EV per attempted cycle over 6,174 attempts but t=1.552 < 3.33 and sat at both continuous grid boundaries, so it is not a real candidate and does not trigger another search.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Establish naive/persistence forecast baselines before any model is built.
Published before any model exists, per the project's standing rule that a null chosen after seeing model results cannot be distinguished from a rationalization. Supplies the persistence/naive nulls spread-forecast.md is graded against.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Report recorder-era coverage and third-copy backup currency for recorded order-book depth.
Recorder era coverage and third-copy backup currency, regenerated per run. Not a fixed research verdict.
Find the interior expected-value optimum over the declared execution-parameter grid.
THE OPTIMUM IS NOT IDENTIFIED. Expected value rises monotonically to the edge of the grid on depth_bps, sell_depth_bps, ttl_days, and the best configuration sits at that edge. The grid bound chose this configuration, not the data. (grid_edge_unidentified: true; canonical digest 1eeb1ad85faa5d700330d0180334d38a8ea859596f3d140ffe2213e12e525887)
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
What changed in the calendar-window panel migration, and whether Gate A's headline result moved.
Daily panel: 37 columns identical, 8 moved, 4 added; 2h panel: 42 identical, 0 moved, 4 added (must be zero, and is). Gate A's headline re-runs to -0.530% with 40.7% of items above zero -- unchanged. Movement is tilted outside the tradeable tier on all 8 moved columns but NOT confined there.
Measure fixed-horizon lifecycle watchability for the frozen recorder cohort and reproduce one frozen legacy bar baseline.
3d is first measurable for aggregate watchability (419 complete-lifecycle-eligible placements); the frozen legacy comparator has no economic drift, only receipt-bound touch-only capacity-counter instrumentation drift.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Admit the private full-book development corpus and account for temporal gaps.
COVERAGE_GAPPED with 1,253 temporal gaps; downstream replay must fail closed across them.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Test whether an executable-markout veto improves the unchanged frozen replay baseline.
Economically inconclusive despite formal selection; it cannot open the holdout or support profitability.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Test whether a causal fill-risk gate improves the frozen pessimistic strategy economically.
Negative stop: no fill-hazard gate passed the strict economic-improvement rule.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Replay frozen strategy policies without bridging corpus gaps or hiding queue ambiguity.
Gap-aware aggregate replay exists; a temporal gap is a no-trade coverage outcome and is never interpolated.
Freeze the cohort and test full-book cache fidelity before replay.
REQUIRES_ACQUISITION: zero of 60 items were usable in the original corpus account.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Score the latency-corrected pessimistic market-making strategy reprice5-pessimistic-latency20-v1 from bounded aggregate slices.
The slice intersection did not pass; terminal negative stop with the holdout unaccessed.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Bind R1 to the preregistered fill-hazard result.
R1 is a terminal non-candidate because the fill-hazard measurement is a negative stop.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Run the one-shot R2 holdout only if the exact frozen R1 account survives.
R2 is a one-shot non-result because the exact frozen R1 account is a non-candidate.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Freeze the disjoint cohort partition of the tradeable tier and the fixed per-release list of admitted patch-notice events.
Freezes the disjoint cohort partition of the tradeable tier and the fixed per-release list of official-notice events admitted into it. Computes no price aggregate itself (by design -- see analysis/cohorts.py docstring).
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Create a future twenty-second protocol only if R2 has a survivor.
No twenty-second protocol exists because R2 produced no survivor.
Record v2's frozen/not-frozen status as an explicit non-result.
v2 is WITHHELD, not frozen and not abandoned. Reason (re-recorded 2026-08-09): weighted-quote geometry is attributed, but the per-attempt frontier still ends at the grid wall under a fill model whose trade-off did not identify an interior optimum. Bound to docs/research/strategy.json's canonical digest 1eeb1ad85faa5d700330d0180334d38a8ea859596f3d140ffe2213e12e525887.
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
How much of the touch-only fill rate the price-bound guard artifact itself accounts for, versus a real fill.
At the touch the correction is ~1%. At 25% placement depth on the BUY side, 17.83% of unguarded fills came from the price-bound artifact alone under the primary VOLUME_CONFIRMED rule, and VOLUME_CONFIRMED carries a HIGHER false-fill share than TOUCH_ONLY at every depth >= 1% ('confirmation concentrates the contamination rather than diluting it').
This site's own explanation of the words above. The verdict itself is published exactly as the research index records it and is not rewritten here.
Determine why the 100/min export rate ceiling is effectively reached at one request every three minutes.
Supported hypothesis: FAN-OUT. 429 risk is 3.9x higher after a large chunk (18% of 73) than after a small one (5% of 22); inconsistent with a shared-address or competing-process explanation.
Track acquisition progress of /export historical depth widened to the tradeable tier.
Operational scope/progress snapshot for the /export widening acquisition (413 items, 2,891 chunks, ordered by notional traded per week). Not a fixed research verdict; figures move as acquisition proceeds.
See methods & provenance for the seal, the chance ledger, execution bounds, and how the weighted quote differs from the touch.