Four 4D myths, tested properly.
Everyone "knows" that 8888 never comes out, that one draw day is luckier, that consolation numbers look different, and that Magnum's machine has its own habits. We ran all four beliefs against the full archive in a single batch. Here is every number we got.
1. Pattern numbers
Do 8888-style numbers come up less often than ordinary ones?
Every 4-digit number has the same 1-in-10,000 chance, so a family of numbers should take exactly the share of draws its size entitles it to. This counts every prize slot ever published by every operator — 715,271 of them — and checks each family against that share.
| Family | Of 10,000 | Expected | Observed | Obs / exp | p |
|---|---|---|---|---|---|
| QuadAAAA · 8888 | 10= 10 | 715 | 735 | 1.028 | 0.472 |
| TripleAAAB · 8887 | 360= 10 × 9 × 4 | 25,750 | 25,565 | 0.993 | 0.242 |
| Two pairsAABB · 1212 | 270= C(10,2) × 6 | 19,312 | 19,393 | 1.004 | 0.559 |
| One pairAABC · 1123 | 4,320= 10 × C(9,2) × 12 | 308,997 | 309,103 | 1.000 | 0.801 |
| All four differentABCD · 1234 | 5,040= 10 × 9 × 8 × 7 | 360,497 | 360,475 | 1.000 | 0.960 |
| PalindromeABBA · 1221 | 90= 10 × 9 | 6,437 | 6,537 | 1.015 | 0.215 |
| Doubled pairAABB · 1122 | 90= 10 × 9 | 6,437 | 6,484 | 1.007 | 0.564 |
| AlternatingABAB · 1212 | 90= 10 × 9 | 6,437 | 6,372 | 0.990 | 0.416 |
| Ascending runA,A+1,A+2,A+3 · 1234 | 7= 7 | 501 | 499 | 0.997 | 0.958 |
| Descending runA,A-1,A-2,A-3 · 4321 | 7= 7 | 501 | 516 | 1.031 | 0.508 |
Quads, triples, two-pairs, one-pairs and all-different are mutually exclusive and add to exactly 10,000, which is why they can be tested together in one goodness-of-fit. ABBA, AABB and ABAB all sit inside the two-pair shape (90 × 3 = 270) and are listed separately only because players name them separately. Each family gets a two-sided binomial test with a continuity correction.
2. Draw days
Do Wednesday draws produce different numbers from Saturday, Sunday or special draws?
The classic operators draw on Sunday, Wednesday and Saturday, plus the extra special draws the Ministry of Finance approves — historically almost always a Tuesday. Each group's numbers are compared with the others', digit by digit and by digit shape.
| Draw day | Draw dates | Numbers |
|---|---|---|
| Sunday | 2,055 | 21,918 |
| Wednesday | 1,863 | 21,318 |
| Saturday | 2,151 | 22,563 |
| Special draw | 1,411 | 7,572 |
1,411 special-draw dates were identified in the archive by weekday. The curated 2026 special-draw list covers 6 of them, because that list only carries the current year — deriving the group from the weekday rather than from the list is what lets this test reach back to 1985.
Grand Dragon, 9 Lotto and Lucky Hari Hari draw seven nights a week and are excluded entirely; pooling them would compare draw schedules instead of numbers. One honest wrinkle: the special-draw group also contains Magnum's and Sports Toto's pre-1995 Monday and Thursday schedule, so it mixes eras as well as days.
3. Prize tiers
Do 1st-prize numbers look different from consolation numbers?
Every slot on a board comes out of the same machine on the same night, so a 1st prize's ending digit and digit shape should be indistinguishable from a consolation number's. Both are compared across all five tiers.
| Prize tier | Numbers published |
|---|---|
| 1st prize | 31,111 |
| 2nd prize | 31,111 |
| 3rd prize | 31,111 |
| Special | 310,962 |
| Consolation | 310,976 |
The tiers have very different sample sizes — three slots a night against twenty — which chi-square handles, but it does mean the two big tiers dominate the total. Numbers within one board are also drawn without replacement, a dependency of roughly 23 in 10,000 that is far too small to matter here and errs on the conservative side.
4. Operator signatures
Is Magnum's machine different from Toto's?
Each operator's top-three numbers are broken into digits and the operators are set against each other. The test runs twice: once across every operator with enough history, and once across Magnum, Sports Toto and Da Ma Cai alone — the longest and most heavily regulated archives, and the three the myth actually names.
| Operator | Digits counted | Largest gap from 10% |
|---|---|---|
| Magnum 4D | 81,828 | 0.24 ppon digit 0 |
| Da Ma Cai 1+3D | 71,880 | 0.23 ppon digit 8 |
| SportsToto 4D | 68,268 | 0.29 ppon digit 3 |
| Lucky Hari Hari 7:30PM | 35,700 | 1.79 ppon digit 4 |
| Grand Dragon | 27,276 | 2.08 ppon digit 1 |
| Special CashSweep | 18,492 | 0.31 ppon digit 3 |
| Sandakan 4D | 18,432 | 2.07 ppon digit 4 |
| Singapore 4D | 17,784 | 0.48 ppon digit 4 |
| Sabah 88 4D | 16,800 | 0.57 ppon digit 8 |
| Lucky Hari Hari 3:30PM | 16,416 | 1.99 ppon digit 4 |
| 9 Lotto | 456 | 3.42 ppon digit 1 |
This is the test most exposed to sample size. Magnum's archive starts in 1985 and Grand Dragon's in 2020, and chi-square grows with the number of observations, so a longer history produces a smaller p for exactly the same sized difference. That is why every row also carries the largest gap between an operator's digit share and a flat 10% — a plain-percentage-point measure that does not move when you add more years.
How to read these numbers
A p-value answers one question: if the belief were false and the draws were fair, how often would data look at least this odd by luck alone? A small p is surprising under fairness. It is not the probability that the myth is true.
We ran 17 tests in a single batch. About one test in twenty lands under p = 0.05 even on perfectly fair data, so with this many tests a couple of "significant" results are guaranteed noise. The flag line is therefore p < 2.9e-3 — 0.05 divided by the number of tests — and nothing above it is treated as a finding.
Every p is printed next to an effect size. Cramér's V measures how strongly two things are associated, from 0 to 1, and does not grow with sample size; the largest single gap is the biggest difference between two shares in plain percentage points. A tiny p next to a V of 0.01 means a real but microscopic difference found by a very large archive — not a big one.
The expected outcome for all four of these is nothing, and nothing is a publishable result. A test run honestly that finds no effect is the entire point of this page: it is what lets us say a popular belief is not supported, instead of merely asserting it.