Does this scoring model actually work?
The page that is allowed to say no
Every wallet tracker on the market ranks by trailing profit, which guarantees the top of the list is whoever got luckiest recently. The only honest test is out-of-sample: score wallets on one window, then check whether that score predicts a completely separate later window the model never saw.
That is what the backtest below does. It reports Spearman's ρ between train-window and test-window performance, plus the lift of the top-decile wallets over the median. If ρ is not meaningfully positive, the model is decoration and the correct response is to stop using it — not to reshuffle the weights until the number looks better.
Dataset too short for a verdict. 24.2h of trade history is indexed; a train/test split needs at least 14 days before persistence can be measured out-of-sample. Any backtest run before then will honestly report insufficient-data, and the leaderboard should be treated as provisional.
Latest verdict
not run
no backtest recorded
Spearman ρ
—
train rank vs test rank
Top-decile lift
—
best 10% vs median wallet
Wallets evaluated
0
present in both windows
Backtest history
Every run is kept, including the bad ones
No backtest has been recorded yet
Run
pnpm backtest once the worker has accumulated enough history. Until then the leaderboard is unvalidated by construction, and this page will keep saying so.Current data coverage
Trades indexed
34,065
Distinct wallets
8,584
Tokens
46
Pools tracked
184
Wallet scores stored
118
Rings detected
0
First trade
Aug 25, 10:22 PM
Last trade
Aug 26, 10:33 PM