The Data Desk · Accuracy Ledger

How our predictions have held up

Every projection we make — and every tracked commentator's call — is written to a public ledger before the round locks, then graded against what actually happened. Nothing is quietly deleted: when a call is revised, the earlier version is kept and marked superseded, and only the final pre-lockout call is scored.

69,971
Predictions logged
30
Predictors tracked
18
Commentators graded

Our own models have been tracked live since 04 Jul 2026. Commentator calls are graded from the back catalogue too — the earliest transcribed prediction we hold dates to 11 May 2016.

The engine's own record
47.9%
price-direction calls correct
over 1,805 graded calls · rounds 19–25 · average miss $29.4K · 2,154 withdrawn

Our price model is the first of our own predictors to earn a graded record. Across 1,805 price-change calls it read the direction of a player's price move — up, down or flat — correctly 47.9% of the time. A three-way call gives roughly 33.3% to a coin, so that is better than guessing, though not by a margin worth leaning on. Its average miss was $29.4K, which is a lot of money to be wrong by on a single trade. We show this because grading ourselves by the same rule we apply to everyone else is the entire point of this page — a better number would have been more comfortable to publish, not more honest.

Corrected 25 August 2026. Up to round 25 we published price calls for players who could not change price — fewer than three games that season, or not playing that round. There were 2,154 of them, and every one was scored against a movement of zero. They have been withdrawn from the graded record. Our hit rate over the calls that remain is 47.9% rather than 27.4%; our average miss is $29.4K rather than $26.2K. The second figure is worse because the withdrawn calls were easy to be nearly right about.

A price call counts as correct when it gets the direction right, with any move under $3,000 treated as flat. The average miss is the mean gap, in dollars, between the predicted change and the real one.

The trade engine's record
61.6%
trade-in calls that beat the field
over 138 graded calls · rounds 19–24 · produced by the nightly form-value-fixture blend

Each buy the engine recommends is judged on what the player actually scored over his next three played rounds, against the median of every player the engine could have named that week. Across 138 graded calls, 61.6% beat that median. Every one of those calls came from the nightly form-value-fixture blend — when the method changes, this figure starts again under the new method's name rather than quietly absorbing the old record.

A caution on the sample: 138 calls across rounds 19–24 of a single season, all under this year's Stats Perform adjudication. Six rounds is a window, not a settled record, and we publish it because the ledger is graded, not because the sample is large enough to lean on.

What this figure covers: the buy calls this engine publishes each night across every squad it reads, chosen by the nightly form-value-fixture blend. It does not cover the trade of the round on your front page, which is picked for your squad alone by a different method — that one keeps its own record, and it will appear here under its own name once enough of its calls have graded.

The engine's trade-in calls — regret

For each buy the engine recommends, regret is the SC points the best affordable same-position alternative averaged over the next three completed rounds, less the pick's — priced as at the time of the call. Zero means the pick was the best affordable buy. A call only grades once all three of those rounds have been played and settled, so a short or part-settled window waits rather than counting.

27
Avg points foregone
8%
Was the best buy
112
Graded buy calls

What this figure covers: the buy calls the engine publishes each night across every squad it reads, chosen by the nightly form-value-fixture blend — the same calls the record above describes. It does not cover the two-trade planner's combinations, or the trade of the round on your front page, both of which are different methods over different squads. The comparator differs too: regret is measured against the single best affordable alternative, where the record above is measured against the median of the field.

The two-trade planner's record

The planner's graded calls have not yet reached the 20 the commentators on this page are held to — 8 have graded so far, so the rate is withheld until the record clears the same bar.

What this figure covers: the buy calls the two-trade planner served to readers who opened the trade board, chosen by the two-trade optimiser. Each buy is graded on its own merits rather than as part of the combination it was shown in, so a two-trade plan counts here as its separate buys and not as one call.

The commentators we grade

The YouTube channels and podcasts we track, ranked by their trade-in record. A trade-in call counts as correct when the player bought beat the median of his position over his next three played rounds. Captain and trade-out calls are tracked too; the fuller per-call breakdown lives on the commentators page.

# Commentator Trade-in hit rate Calls graded Seasons
1 Aman Talks NRL SuperCoach 25.9% 3,545 6
2 Ballr NRL SuperCoach 26.9% 2,547 2
3 Rugby League Guru 28.9% 2,235 5
4 SC Playbook 24.7% 1,730 4
5 CODE Sports 25.2% 1,695 5
6 NRL Fantasy Analysis 25.5% 938 1
7 NRL Physio 31.4% 870 4
8 NRL Fantasy with The Casual Athlete 28.6% 821 1
9 Talking League - NRL Fantasy Podcast 28.2% 628 1
10 Seven Tackle Set | NRL Supercoach Podcast 28.1% 552 1
11 SuperCoach Insight 26.8% 512 1
12 NRL Fantasy Amateurs 33.2% 455 1
13 The Weekly Rub Down | NRL SuperCoach 23.6% 406 1
14 The NRL Supercoach Therapy Podcast 25.4% 394 1
15 The SuperCoach NRL Podcast 25.2% 274 1
16 The NRL Supercoach BDE Podcast 34.6% 257 1
17 Jones and Stones Fantasy Podcast 20.6% 170 1
18 Head to Head NRL Fantasy Podcast 21.8% 133 1

Hit rates cluster in the low-to-high twenties. Trade-in calls are hard: beating the positional median over three rounds is a genuine bar, and no tracked commentator clears it much more than three times in ten.

Technical · Confidence calibration

When a tipster flags a call as high confidence, it ought to land more often than their run-of-the-mill medium call. Across the tracked field it mostly does not: only 5 of 16 commentators with a large enough sample show their high-confidence trade-ins beating their medium ones. Shown only where both bands clear 100 graded calls; smaller samples say nothing reliable and are left out.

Commentator High conf. Medium conf. Higher confidence…
Rugby League Guru 28.7% (1,596) 29.6% (639) did no better
Ballr NRL SuperCoach 26.4% (1,480) 27.5% (1,067) did no better
Aman Talks NRL SuperCoach 25.3% (1,412) 26.3% (2,133) did no better
CODE Sports 25.6% (1,095) 24.5% (600) did better
SC Playbook 22.9% (1,065) 27.5% (665) did no better
NRL Physio 30.3% (545) 33.2% (325) did no better
NRL Fantasy Analysis 21.7% (360) 27.9% (578) did no better
SuperCoach Insight 27.2% (335) 26.0% (177) did better
NRL Fantasy with The Casual Athlete 28.9% (332) 28.4% (489) did better
Seven Tackle Set | NRL Supercoach Podcast 26.2% (290) 30.2% (262) did no better
The Weekly Rub Down | NRL SuperCoach 23.9% (285) 23.1% (121) did better
Talking League - NRL Fantasy Podcast 25.5% (271) 30.3% (357) did no better
The NRL Supercoach Therapy Podcast 19.7% (183) 30.3% (211) did no better
NRL Fantasy Amateurs 30.2% (179) 35.1% (276) did no better
The SuperCoach NRL Podcast 26.5% (136) 23.9% (138) did better
The NRL Supercoach BDE Podcast 30.0% (120) 38.7% (137) did no better

Methodology: predictions are snapshotted with the exact inputs used at the time and only entries recorded before a round's first kickoff are eligible for grading. Revised predictions supersede their earlier version — only the final pre-lockout call is scored. Round scores are graded on the official SuperCoach scale against games the player actually played; price calls against the realised price change; trade-in calls against the player's next three played rounds versus his positional peers. The same substrate grades our models and the tracked commentators, so the comparison is honest.