Honest Leaderboard

The wall

Ranked by worst-plausible-case, deflated by the whole field, gated before it's ranked at all. Most strategies don't make the wall — that's honest. This is the one number a competitor can't copy, because their measurement isn't strict enough to keep the board from filling with garbage.

8
strategies submitted4 live
4
cleared every gate
50%
didn't make it
Deflated by the field

The #1 of N submissions is the max of N tries — expected best-of-N Sharpe grows even under pure noise. Every entry's DSR is deflated at 8 trials, so a bigger board raises the bar for everyone.

Ranked by the lower bound

Order is the lower bound of the block-bootstrap skill CI, never the point estimate. Reliably-good beats occasionally-brilliant. Rank = worst plausible case.

Gates before rank

Eligible iff trades ≥ 20 ∧ deflated DSR > 0 ∧ PBO < 0.5 ∧ CI excludes null. Fail any and you're not ranked low — you're in a separate honest list, with the reason.

On the wall

Eligible strategies, ordered by lower-CI skill. Every figure is computed by the same honest instrument the product ships.

Didn't make the wall

Not ranked 9,999 — ineligible, with the honest reason. A coin-flip is absent, not low.

Put a strategy up

Build one in Studio — the same honesty instrument scores it live, and a shared run reproduces bit-for-bit on the verify page. The board re-scores every entry on seeds it never saw; an entry that stops clearing the gates on fresh data falls off. Tomorrow's data isn't tuneable.

Reference board scored on a fixed synthetic seed-set — the figures demonstrate the ranking instrument, not live returns. Live submission and scheduled fresh-data re-scoring roll out next. Educational and research software; not investment advice. Backtested results are hypothetical and do not indicate future performance.