Method leaderboard

Every prediction method we run, scored on real draws — including a pure-random control arm. This is the table no tipster site would ever publish.

How to read this: 26 draw nights scored so far, averaging 6.0 operator boards a night. By pure chance, one pick should collect about 0.33 box hits per night — that's the baseline every method (and the random control) is measured against. Straight = exact match; box = any digit arrangement.
MethodPicksStraightBoxBox/pickvs chanceReturn
ProfileEdge720380.53+0.20153%
Cold endingsCollecting510210.41+0.0837%
Weighted ensembleCollecting570230.40+0.07124%
FollowerCollecting572190.33+0.0080%
AI (Claude)Collecting510150.29-0.0426%
RhythmChance570160.28-0.0584%
Own-cycleCollecting430120.28-0.05105%
DueChance781210.27-0.0631%
Trend-builtChance723180.25-0.0896%
In formChance511120.24-0.1090%
AnniversaryChance511120.24-0.1067%
HotChance781180.23-0.10111%
Near missChance570130.23-0.1067%
CompanionsChance722160.22-0.11102%
JiranChance630140.22-0.1129%
CoolingChance511110.22-0.1290%
Random controlcontrolChance570110.19-0.14107%
PositionalChance663120.18-0.15112%
DateCollecting11020.18-0.159%
Head pairsChance51090.18-0.1526%
Repeat watchChance720120.17-0.1618%
EndingChance53480.15-0.1869%
ML modelChance57170.12-0.2128%

The return column

What RM1 iBox Big on every pick, on every board, would actually have returned — priced with the real dividend table. Counting box hits treats a consolation and a 1st prize as the same event; this does not. A fair draw returns about 67%, and 100% is break-even.

Read this column last. A single RM1 bet has an average return of 67 sen but a standard deviation of RM5.75, because 38% of all expected return sits in the 1st prize alone — an event with a 0.24% chance. That makes return roughly 5 times noisier than the box-hit rate and in need of about 25 times more data to mean anything, which is why the verdict is decided on hits and the table is not sorted on money. For scale: the pure-random control arm is currently returning 107%.

Read this before reading the table. With only 26 nights, every gap in this table is statistical noise — a method "leading" today will regress. If any method were consistently and significantly ahead of the random control over hundreds of nights, that would be extraordinary (and we would say so loudly). Until then: this page exists to show they are all the same, honestly. The random control is currently collecting 0.19 box hits per pick — any method below it is doing worse than throwing darts.

From 20 Aug 2026 every row here is scored on the Malaysian-licensed boards only (Magnum, Toto, Da Ma Cai, Sabah 88, STC, CashSweep) — the same six the picks are built from. Fewer boards per night means fewer chances to hit, so both the methods and the random control sit lower than they did before, and the whole history is recounted on this basis.

How we scoreThe statistics behind each verdict, and the rules for retiring a method›

How the verdict is decided (and why it isn't decided yet)

Re-checking a leaderboard every night and calling the best row a winner is how false discoveries are made — repeated peeking inflates the error rate. The verdict column instead runs Wald's sequential probability ratio test per method: chance-level box rate vs a 1.5× edge, with 5% error caps that hold no matter how often we look. A method is only called an Edge or confirmed as Chance when the accumulated evidence crosses the threshold; until then it honestly says Collecting.

For scale: a method performing exactly at chance needs roughly 95 picks before the test can confirm it — verdicts here are earned slowly by design.

Method retirement policy

Pre-committed

A method is retired from the nightly picks when its SPRT verdict is "chance" AND it has at least 1,000 scored picks. Retired methods keep their full record on this leaderboard, marked retired. Any change to the line-up is announced here with its date. The random control is never retired — a control removed for behaving like a control is not a control.

Status right now: no method meets this rule. That is the whole point of publishing it today — a rule written after a method fails is not a rule, it is an excuse.

Why a pick floor as well as a verdict: the SPRT can confirm "chance" fairly early on a method that simply had a quiet run. The floor makes sure nothing is benched on a short record, and keeping the retired rows visible means nothing quietly disappears from the scoreboard once it stops being flattering.

Relegation & promotion

The softer stage before retirement: a MAIN method (hot, overdue, own-cycle, ending) whose verdict is "chance" with at least 150 scored picks loses its main-pick slot and continues in the deep set — still locked, still scored, still on this board. Its slot goes to the deep method with the highest ensemble trust weight and at least 150 picks of its own. The random control and the ensemble itself are never eligible. Swaps happen only when this rule fires, and are announced here with their date.

Status right now: no main method meets the relegation rule. Published today, before any does.

Are “hot numbers” real? The shrinkage test

Empirical Bayes on the last 365 days of top-3 hits, fitted across all 10,000 numbers.

6.7% of the spread between numbers exceeds what Poisson noise explains — worth watching, but see the fairness audit before reading anything into it.
NumberHits (365d)Honest estimateExcess kept
425161.047%
237550.977%
326450.977%
643750.977%
762350.977%
767550.977%
771350.977%
787450.977%
804750.977%
813450.977%

“Honest estimate” is the posterior expected hit count for the NEXT 365 days, after shrinking toward the all-number average of 0.69. “Excess kept” shows how much of a number's above-average performance the math believes; near 0% means its hot streak was luck.