Reading the Predictions Track Record: A Five-Week Audit

Five weeks of captain calls fit on one screen, and most readers leave that screen with the wrong lesson. A simple audit on the same five weeks — what the calls claimed, what actually happened, and what the gap teaches about reading any short predictions table without fooling yourself.

Before, a reader could scroll the predictions table, count the green rows and assume the desk was on a run. After a working audit, the same five rows reveal four separate questions that a one-line win rate cannot answer: how big the wins were, what the alternative-pick column actually shows, how much of the pattern is variance, and what would have to change in week six for the pattern to break.

No verified current report is available for this assignment, so the article stays bounded to the audit method. The numbers cited below are illustrative, drawn from the kind of short-sample results a daily fantasy desk can actually produce. They are not a leaderboard and they are not a promise; they are the working material for reading a short track record honestly.

What the predictions table actually claims

A typical predictions table on a fantasy desk lists five things per call: the date, the match, the captain pick by role, the contest format, and the result. The result column usually reads “Top 18%” or “Top half” or “Loss”. The table looks like a leaderboard because it is shaped like one, but it is not. It is a list of decisions, each made under a specific match condition that the table does not record.

Three pieces of information are missing from the table itself. First, the role the captain actually played in the innings — was the call a top-order batter who batted through, or one who fell in the powerplay? Second, the contest pool the result was measured against — a top-12% finish in a 50,000-entry marquee is not the same as a top-12% finish in a 200-entry head-to-head. Third, the squad sheet at the time the call was made, including the assumption that became false ten minutes after the call went live. None of these are present in the table, and all of them shape what the row means.

That is why a single-glance reading of the table is unsafe. A row reading “Top 18%” can describe a clean dominant performance or a low-scoring contest where every captain struggled and 18% looked like the new normal. The label is the same; the situation is not. The audit exists to keep that distinction legible.

Win rate hides the size of the calls that won

Counting wins over five matches produces a clean number. Four wins and one loss looks respectable. Five losses looks damning. Both numbers hide the more useful question: how much did each win actually beat the field by, and how much did the loss lose by?

Take a hypothetical five-week audit. Week 1, finger-spinner captain, league contest, top 18%. Week 2, top-order batter, 50/50, top half. Week 3, death pacer, league, top 30%. Week 4, all-rounder, head-to-head, loss. Week 5, power-hitter, marquee, top 12%. The four-win narrative is correct on its face, but the spread of those finishes matters more than the count. A 12% finish in a marquee contest is a much stronger result than a 30% finish in a regular league; the wins are not equivalent.

The desk keeps a second column off the public table: the percentile of the captain’s fantasy score within the contest pool, not just the percentile of the reader’s XI. A captain who finished in the top 5% of fantasy scorers but whose reader finished at top 25% still helped; the captain who finished at top 30% on a slow night helped less. The audit recovers that second column by re-reading each row against the published points for the same match day, where the platform makes the scoreboard available.

A close view of a printed track record spreadsheet with percentile columns highlighted and margin notes
The win column counts. The percentile column weighs. A short track record needs both before the reader can call it a run.

What the alternative-pick column actually shows

Most predictions tables publish a primary call and an alternative call. The alternative is described as the “if the primary does not work” option. The audit treats the alternative column as a separate claim with its own track record, not as a hedge.

If the alternative call would have outscored the primary in three of the last five weeks, the audit reads that as a method problem, not a result. The desk was right more often than wrong, but the alternative showed a pattern of stronger calls the primary did not. A reader following the alternative would have had a better result without changing the desk’s method. The honest response is to ask why the alternative was demoted from primary in the first place, and whether the rubric that demoted it is calibrated correctly.

If, instead, the alternative would have outscored the primary in zero of the last five weeks, the audit reads that as a healthy signal. The desk picked the better call each time. The alternative column still belongs in the table because it documents what the desk considered, but a zero-out-of-five alternative score means the primary was the right choice every time. That distinction is invisible from the win column alone.

Variance, not skill, dominates short samples

Five matches is a small sample. A captain call has at least four random inputs — the toss, the pitch behaviour, the bowler matchup that actually plays out, and the small moments that turn a fifty into a hundred or a four-ball into a wicket. Across five matches, those random inputs can produce any pattern from five wins to five losses, regardless of how well the desk reads the match.

The audit treats any five-week result as a noisy snapshot, not a verdict. A four-win run can be a genuinely good run, a moderately good run with lucky sequencing, or a coin-flip run with one bad match. The win count cannot tell them apart. To tell them apart, the audit needs a baseline: the average percentile the desk’s picks would have produced if the desk had picked at random from the eligible captain pool.

That baseline is the missing column on most public tables. A desk that finishes at top 25% in four of five weeks looks strong; a desk that finishes at top 25% when the average random captain also finishes at top 25% is performing at chance. The audit recovers the baseline by reconstructing the eligible captain pool for each match day and computing the average percentile the pool would have produced. Where the desk beats the pool average, the run has signal. Where the desk matches it, the run is variance.

Reading the rubric against the record

The desk publishes a captaincy rubric alongside the picks. The rubric scores role, form, matchup and surface. A reader can re-run the rubric on a past match day and see what score the chosen captain received, and what score the alternative received. The audit does this for each of the last five weeks.

If the rubric and the record agree — the call with the higher rubric score produced the stronger result in four of five weeks — the rubric is doing useful work. If the call with the lower rubric score produced the stronger result in four of five weeks, the rubric is being overridden by a heuristic the desk has not yet written down. That second case is more interesting than the first. It means there is an unwritten rule shaping the calls, and the unwritten rule is the one the desk should publish next.

The audit also reads the rubric the other direction: does any single rubric input explain most of the result? If role predicts the result in four of five weeks, the desk is essentially a role-based desk; the form and surface inputs are decorative. If matchup predicts the result in four of five weeks, the desk is essentially a matchup desk. A reader can then decide which input to weight when building their own XI without the desk.

Audit note. Reading the rubric and the record together is more useful than reading either alone. The rubric shows what the desk said it was doing. The record shows what actually happened. The gap between the two is where the desk’s real method lives.

A three-line test for any short track record

The audit boils down to three lines. The first line: how many calls and what the win count is. The second line: the average percentile the desk’s calls produced against the eligible captain pool, not the contest field. The third line: how often the call with the higher rubric score produced the stronger result. If all three are favourable, the run has signal. If only the first line is favourable, the run is variance with a friendly face.

For a hypothetical five-week audit, the lines might read: four wins from five calls; average captain score at the 22nd percentile of the eligible pool (so the desk finished in front of 78% of the eligible alternatives); and three of five calls where the higher rubric score produced the stronger result. The first line is favourable. The second line is strongly favourable. The third line is mixed — signal, but not consensus. A reader following the desk should expect the next five weeks to look more like a continuation of the second line than a repeat of the first.

The test is portable. It works on any short predictions table the reader is auditing — the desk’s table, a competitor’s table, a tipster’s table. The only requirement is that the table publish enough information to reconstruct the three lines: the calls, the result percentiles, and the rubric scores. Tables that do not publish all three cannot be audited, only trusted, and the audit recommends reading them accordingly.

What the gap between call and result teaches

The audit is not a verdict on the desk. It is a method for keeping the desk honest while keeping the reader honest about the desk. A short track record is the wrong place to settle the question of whether a desk is good. A short track record is the right place to settle the question of whether the desk is worth following for the next five weeks, with the understanding that the next five weeks will produce their own audit.

That framing has a useful side effect. It stops the desk from optimising the table for the reader’s comfort. A table that publishes the second and third lines alongside the win count will look less flattering on quiet weeks and more flattering on loud weeks; the average over a season will be more honest than the average over a highlight reel. The reader who audits the table once a month gets a sharper picture of the desk than the reader who watches the table update in real time.

It also gives the desk a way to update its rubric. When the audit shows a consistent gap between rubric score and result, the desk has a choice: tighten the rubric so the score predicts the result, or loosen the rubric so the desk’s call reflects the gap. Either choice improves the next five weeks. The audit is the only tool that surfaces the gap in a form the desk can act on.

A medium editorial view of a notebook page listing five past calls with margin notes and a checklist of audit lines
The gap between the desk’s rubric score and the published result is the most useful signal in any short track record. It is also the column most tables leave out.

What to watch next

The next five weeks of captain calls will produce their own audit. A reader who runs the three-line test on each week as it lands will see the signal sharpen or fade without waiting for the season to end. The desk will publish the second and third lines alongside the first on a monthly basis, so the audit does not need to be reconstructed from scratch each week. A reader who wants to verify the audit can compare the desk’s published second-line percentile against their own reconstruction for any past match day where the scoreboard is still public.

For readers who want a worked example, the predictions desk keeps the full table updated with the desk’s most recent calls, the rubric score for each call, and the contest-format detail that lets a careful reader reconstruct the eligible pool. The predictions desk also links to the weekly recap notes where the desk explains each call in plain language. Reading the table, the rubric and the recap together is the closest the desk can come to opening the working method without handing the rubric to readers who would rather build their own.

The audit will not turn a four-win run into a five-win run. It will not turn a quiet week into a loud one. What it will do is give the reader a tool for reading any short track record — the desk’s, a competitor’s, their own — without the easy mistake of counting wins and calling the count a method. The desk runs the table. The reader runs the audit. Both sides are better for keeping the two jobs separate.

Run the audit on the next five weeks.

The desk keeps the table. You keep the three lines.

Get Started