NFL

How accurate are NFL prediction models?

A plain answer with the numbers behind it, from 2010–2025 of games, plus our own live record for comparison. · predictions are published to show the model works. nothing here is betting advice.Last updated: predictions Oct 1, 11:02 AM ET · results graded Sep 28, 11:22 PM ET

The short answer

The best public forecast of an NFL game, the closing betting price, picks the winner about two games in three (66.5% over 4,142 regular-season games, 2010–2025). Good statistical models land between 63% and 67%. Anyone claiming to sustain 70% or more over full seasons is, so far as public records show, either measuring something else, counting selectively, or getting lucky over a short stretch. A model is not "good" because it beats a coin flip; picking the home team every game gets 55%. It is good when it states probabilities that come true at the stated rate, over enough games to rule out luck.
66.5%
Market pick at kickoff, 2010–2025
65.0%
Our model, same games, walk-forward backtest
55%
Just pick the home team
62.4–71.4%
The market's worst and best single seasons (2021, 2013)
29–17
Our live 2026 record so far (63.0%); market pick at kickoff 67.4% on the same games

What the numbers mean

Why two in three is the ceiling for now

NFL games are close and short. Sixteen or seventeen games a season, one possession deciding a third of them, and injuries that change a team from week to week. The betting market absorbs a week of public and private information, price movement, sharp money, injury news, and still only reaches about 66–67%. That number is the honest benchmark for everything else: a model that cannot beat it is normal, and one that claims to beat it by a lot over many seasons deserves a very hard look.

One season proves almost nothing

Over a 272-game regular season, a picker whose true accuracy is 66.5% will land anywhere between about 61% and 72% on luck alone (one standard deviation is 2.9 points). The market itself ranged from 62.4% to 71.4% across seasons with no change in how good it was. So a model that went 70% last year has told you very little; a model that went 65% for ten years has told you a lot.

Accuracy is the wrong headline anyway

Two models can both go 65% while one says "55%" on every game and the other says "80%" on every game. The second is lying about what it knows. The better measure is calibration: when a model says 70%, do those teams win about 70% of the time? And the Brier score, which penalizes confident misses more than cautious ones (lower is better; a coin flip scores 0.250, the market's final price about 0.211, a good model near 0.215–0.220). A model that publishes probabilities and lets you check them is worth more than one that publishes a record.

Against the market margin is a different question

Picking winners and beating the point market margin are not the same skill. The market margin is set so that either side is about a coin flip, and a bettor has to win 52.4% of market margin picks just to cover the bookmaker's margin. Public models that beat the final market margin consistently are, in our experience of the public record, essentially nonexistent, and ours does not either. Any site showing a big against-the-market margin record without showing every pick before kickoff is asking you to take it on faith.

How to judge any model's claim

Five questions, in order:

1. Were the picks published before kickoff, and can I see all of them? A record built after the fact, or shown selectively, is not a record.

2. Over how many games? Anything under a few hundred is mostly noise. Sixteen seasons of backtest is about 4,000 games; that is where the noise gets small enough to mean something.

3. Was the backtest honest? "Walk-forward" means the model only ever used games that had already been played. A model tested on the same seasons it was tuned on will look better than it is.

4. What is the benchmark? Compare against the market's final price and against "pick the home team," on the same games. "65%" means nothing until you know the market got 66.5% on those games.

5. Are the probabilities calibrated? Ask for the table: stated probability by bucket versus how often those picks won. If the site cannot produce it, it has not checked.

Our own answers: picks stored before kickoff and never edited (every one on the weekly pages and Record); 4,142 backtest games walk-forward with a leak audit (Backtest); the market and the home-team baseline shown beside us; calibration published live and below. We are a normal model, 65.0% against the market's 66.5%, and we say so.

Season by season, 2010–2025

Regular season, ties excluded, games with a market's final price. Our numbers are from the walk-forward backtest (the model retrained each week on what had been played by then), not from live picks.
SeasonGamesMarket pick at kickoffOur modelHome team
2010240 66.7% 61.7% 54.6%
2011256 66.8% 68.4% 56.6%
2012255 64.3% 63.1% 57.3%
2013255 71.4% 63.5% 60.0%
2014255 66.7% 69.4% 56.9%
2015256 62.5% 64.1% 53.9%
2016253 64.4% 64.8% 57.7%
2017253 70.8% 66.0% 56.5%
2018254 66.1% 68.9% 60.2%
2019255 64.3% 64.3% 51.8%
2020255 67.5% 67.1% 49.8%
2021271 62.4% 61.3% 51.7%
2022269 66.2% 62.5% 56.1%
2023272 68.0% 63.6% 55.5%
2024272 71.3% 70.2% 53.3%
2025271 65.3% 60.9% 53.9%
All4,14266.5%65.0%55.3%

How often pick actually win

By how confident the market's final price was. This is what calibration looks like: an 80% pick should win about 80% of the time, and does. It also shows where the games are hard: a third of NFL games are 50–60% propositions for everyone.
Market saidGamesPick wonOur pick won
50–55%55652.9%51.1%
55–60%73757.4%53.5%
60–65%77559.6%58.3%
65–70%61568.5%66.3%
70–75%59674.7%74.3%
75–80%44778.7%78.5%
80%+41586.3%86.3%
We agree with the market on the pick 87% of the time. On the 529 games where we did not, we were right 43.9% of the time and the market 56.1%. We publish that because it is the number a reader actually wants when they ask "are you better than the market?", and the answer is no, and nobody public is by much.

Where these numbers come from

Game results, market's final price and play-by-play from the open-source nflverse project; our backtest replays 2010–2025 week by week with the model retrained only on games already played, checked by an automated leak audit. Our live picks are stored before kickoff and graded after the final, with the market pick at kickoff graded beside them on the same games. Every term here, win probability, calibration, Brier score, market's final price, walk-forward, is defined in the glossary. Method, in the level of detail we publish, is on How this works. This page is updated automatically whenever the backtest or the live record changes.