Welcome to Regression Alert, your weekly guide to using regression to predict the future with uncanny accuracy.

For those who are new to the feature, here's the deal: every week, I dive into the topic of regression to the mean. Sometimes, I'll explain what it really is, why you hear so much about it, and how you can harness its power for yourself. Sometimes, I'll give some practical examples of regression at work.

In weeks where I'm giving practical examples, I will select a metric to focus on. I'll rank all players in the league according to that metric and separate the top players into Group A and the bottom players into Group B. I will verify that the players in Group A have outscored the players in Group B to that point in the season. And then I will predict that, by the magic of regression, Group B will outscore Group A going forward.

Crucially, I don't get to pick my samples (other than choosing which metric to focus on). If I'm looking at receivers and Justin Jefferson is one of the top performers in my sample, then Justin Jefferson goes into Group A, and may the fantasy gods show mercy on my predictions.

Most importantly, because predictions mean nothing without accountability, I report on all my results in real time and end each season with a summary. Here's a recap from last year detailing every prediction I made in 2022, along with all results from this column's six-year history (my predictions have gone 36-10, a 78% success rate). And here are similar roundups from 2021, 2020, 2019, 2018, and 2017.

The Scorecard

In Week 2, I broke down what regression to the mean really is, what causes it, how we can benefit from it, and what the guiding philosophy of this column would be. No specific prediction was made.

In Week 3, I dove into the reasons why yards per carry is almost entirely noise, shared some research to that effect, and predicted that the sample of backs with lots of carries but a poor per-carry average would outrush the sample with fewer carries but more yards per carry.

In Week 4, I explained that touchdowns follow yards, but yards don't follow touchdowns, and predicted that high-yardage, low-touchdown receivers were going to start scoring a lot more going forward.

In Week 5, we revisited one of my favorite findings. We know that early-season overperformers and early-season underperformers tend to regress, but every year, I test the data and confirm that preseason ADP is still as predictive as early-season results even through four weeks of the season. I sliced the sample in several new ways to see if we could find some split where early-season performance was more predictive than ADP, but I failed in all instances.

In Week 6, I talked about how when we're confronted with an unfamiliar statistic, checking the leaderboard can be a quick and easy way to guess how prone that statistic will be to regression.

In Week 7, I discussed how just because something is an outlier doesn't mean it's destined to regress and predicted that this season's passing yardage per game total would remain significantly below recent levels.

In Week 8, I wrote about why statistics for quarterbacks don't tend to regress as much as statistics for receivers or running backs and why interception rate was the one big exception. I predicted that low-interception teams would start throwing more picks than high-interception teams going forward.

In Week 9, I explained the critical difference between regression to the mean (the tendency for players whose performance had deviated from their underlying average to return to that average) and the gambler's fallacy (the belief that players who deviate in one direction are "due" to deviate in the opposite direction to offset).

STATISTIC FOR REGRESSION	PERFORMANCE BEFORE PREDICTION	PERFORMANCE SINCE PREDICTION	WEEKS REMAINING
Yards per Carry	Group A had 42% more rushing yards per game	Group A has 10% more rushing yards per game	None (Loss)
Yard-to-TD Ratio	Group A had 7% more points per game	Group B has 38% more points per game	None (Win)
Passing Yards	Teams averaged 218.4 yards per game	Teams average 220.1 yards per game	8
Interceptions Thrown	Group A threw 25% fewer interceptions	Group A threw 5% fewer interceptions	2

There won't be much to say about our passing yards per game prediction until a bit later in the year. It's a good sign that the total hasn't started running away from us; as long as we're within a couple yards of where we started, we'll be well-positioned for the Autumn Wind to push team averages down to the lowest levels we've seen in over a decade.

Group A continues to throw fewer interceptions than Group B, but the data so far is a bit misleading; four Group A teams were on bye last week vs. none for Group B. If they'd all played and matched the rest of Group A's average, Group B would be slightly ahead right now. That bye advantage will reverse almost immediately, with three Group B teams sitting out next week compared to just one from Group A.

Still, I think Group A's performance so far has been fairly surprising. They averaged 0.58 interceptions per game at the time of the prediction and averaged 0.60 interceptions per game in the two weeks since. I expected them to regress more by this point.

I don't think it's anything that demands an explanation. Sometimes, random data behaves randomly, especially over small samples. I still expect matters to regress just as much as I did when I made the prediction. But I think when something is surprising, it's useful to be able to note that surprise. Many are tempted to search for explanations as to why something really shouldn't have been surprising in hindsight, but I think that those explanations tend to be overfitting and this instinct leads to worse predictions in the long run. It's good to just be surprised sometimes.

Let's Compare This Column To A Medical Condition...

Let's say you go to the doctor for your routine physical. Your doctor draws some blood to run some tests and then calls you the next day with some bad news. It seems you've tested positive for some condition or other. The doctor tells you it's pretty rare -- only one out of every 10,000 or so people have it-- but the test is 95% accurate. What are the odds you actually have the condition?

Most people will say 95% here, and most people will be way off. (Don't feel bad; many doctors get this wrong, too.) Your actual odds are about 0.2%. Now, 0.2% is not nothing. It's about a 1-in-500 shot. But I'd be a lot happier with those odds than the 95% I might have reflexively expected.

If the test is so accurate, why are the odds so low? It all comes down to a concept called "base rates".

Imagine giving this 95% accurate test to a million people. Since 1-in-10,000 people have this condition, we'd expect about 100 people to actually have the condition in question. Of those 100, 95 would (correctly) test positive and 5 would test negative (a 5% error rate). We'd also expect 999,900 people to not have the condition, and of that group, 949,905 would (correctly) test negative and 49,995 would test positive (a 5% error rate).

After testing those million people, we'd have 50,090 positive results, of which 95 would be "true positives", or around 0.2% of all positive tests. There are far more "false positive" test results simply because there are so many more people who don't have the condition than who do.

Does this still seem wild to you? Let's use my favorite logical technique to illustrate it better: reductio ad absurdum. Let's say that there's a condition called "Regressionitis" that absolutely, positively does not exist. It is a fake condition that I made up and as such no one in the known universe actually has it, because it's not real.

Now let's say I develop a test for Regressionitis that is 99.9999999% accurate. Let's say I administer this test to you and it comes back positive. What are the odds that you have Regressionitis? They're zero still, Regressionitis isn't real, nobody has it, for my "99.9999999% accurate test" I just generated a random number between one and a billion and if the number was 43 I told you that you had Regressionitis.

The odds you have a condition depend not just on the test's accuracy but on the condition's rarity. The rarer the condition is, the less likely you have it, even with a positive test.

This method of starting with your baseline odds and updating as new information comes in is called Bayesian inference, and it's one of the most powerful tools in our toolkit for estimating the means that everything is expected to regress to.

Already a subscriber?

Continue reading this content with a PRO subscription.

Join Now

All About That Base (Rate)

Bayesian inference has its own lingo, but the two big terms are the prior probabilitiy (or simply "prior") and the posterior probability (usually just "posterior"). The prior is what you believed before new evidence came to light, the posterior is what you believed after. Bayesian inference is the process of taking that prior and adjusting it in the face of new evidence based on the strength of the prior and the strength of the evidence.

(In our example above, the prior is "there's a 1-in-10,000 chance you have this condition", the new evidence is "you tested positive on a test with a 5% error rate", and the posterior is "there's about a 1-in-500 chance you have this condition".)

Bayesian inference comes with its own math. Bayes' Theorem is: P(H|E) = (P(E|H)*P(H))/P(E) where H = the hypothesis, E is the new evidence, P() means "the probability of" and | means "given". You don't have to know any of this, most of the time we can just gesture in the general direction of the math-- it's the process that's important.

Let's look at an example from the NFL.

In my measured opinion, when a QB plays terribly to start his career, this is Not Good(tm).

But it’s worth noting that of the QBs who started as poorly as Bryce Young and then turned it around, they were all very high picks (four went #1, one #2, and one #7). pic.twitter.com/guSwSTu7G1
— Adam Harstad (@AdamHarstad) November 6, 2023

It may surprise you to learn that Bryce Young has not played especially well to begin his NFL career. (It probably won't surprise you, but I suppose it might.) In fact, among rookies with at least 200 pass attempts, Young ranks 11th-worst of the last 60 years in era-adjusted efficiency (measured by adjusted net yards per attempt).

It likewise may (but probably won't) surprise you to learn that most of the rookies who wound up on that list went on to become bad professional quarterbacks. Of the 25 worst qualifiers, I'd say only six became quality long-term starters, a 24% hit rate.

Does this mean things are especially bleak for Young? Not so fast; his poor play to date is not our prior, it is the evidence that we need to use to update our prior. Given that Young was the #1 pick in the draft, our prior should have been "Young is rather likely to become a good quarterback". And while poor play to start his career reduces the odds of that happening, the final posterior odds depend a lot on how high they were to start.

Indeed, of the six "exceptions" I mentioned above, four were likewise #1 overall picks (Jared Goff, Terry Bradshaw, Troy Aikman, and Matthew Stafford), while the other two were drafted 2nd (Donovan McNabb) and 7th (Josh Allen).

Nearly every #1 overall pick who found himself on that list eventually turned things around. As did the #1 overall picks who just missed the list, including Eli Manning (only 197 pass attempts as a rookie, three short of the qualifying cutoff) and John Elway and Trevor Lawrence (around 30th-worst in rookie efficiency). (The big exception was David Carr.)

If we just looked at the hit rate of Bottom 25 rookie quarterbacks, we'd be putting too much weight on the new evidence and not enough on our prior, and as a result, we'd underrate Bryce Young's chances going forward. Despite the poor play to start his career, I think the "mean expectation" for Young is still relatively favorable.

A Quick Hack To Estimate Means

When I want to estimate a specific mean, my favorite thing is to just look at a player's actual career average. How many yards is Davante Adams going to average for every touchdown he scores over the remainder of the season? Well, for his career, he averages a touchdown every 113 yards, so... I think somewhere around 113 seems like a pretty reasonable guess.

This is great for someone like Adams, who has played nine seasons, made six consecutive pro bowls, and been named first-team All Pro three times. Our prior for him is quite strong at this point. It's not as useful when looking at someone like Jordan Addison, who, for his career, averages one touchdown for every 76 yards. What is that average going to be going forward? I have no idea other than that it definitely isn't going to be 76.

Let's say I want to compare Addison, Zay Flowers (1 touchdown for every 476 yards), and Drake London (1 for every 217 yards). All three players' career averages are likely outliers since most players' "true" performance level is somewhere between 120-180. But also, all three players' performance to date tells us new information: Addison is slightly more likely to be a touchdown-dominant receiver, while Flowers and London are slightly more likely to be yardage-dominant.

Looking at the two yardage-dominant players, Flowers' average is more extreme, but London's is on a larger sample. Which of these players is more likely to have a "true" mean of 200 yards per touchdown?

These are hard questions and there are good ways to answer them, but those good ways are pretty difficult and involve a lot of math. I have an "okay" way that's not perfect, but I love it because it takes about fifteen seconds. It's called a Bayesian average and works on the same principles as Bayesian inference. You take a player's career performance, you add a chunk of league-average data to it, and you use the resulting average as your posterior.

Jordan Addison has 534 yards and 7 touchdowns. Let's add 1200 yards at league average touchdown rate (160 yards per touchdown), which would be 7.5 touchdowns; this gives Jordan Addison 1734 yards and 14.5 touchdowns, or about 120 yards per touchdown as a posterior.

Flowers has 472 yards and 1 touchdown; adding 1200 yards and 7.5 touchdowns gives us 1672 yards and 8.5 touchdowns, good for 197 yards per touchdown. London has 1304 yards and 6 touchdowns. After adding 1200 more yards and 7.5 more touchdowns, his expected scoring rate is 185 yards per touchdown.

Here's a quick conceptual rundown of what's going on here: we're setting our prior as "all players are probably about league average" and we're updating that prior based on how each young receiver produces. But this works out in a neat way where the more data we have on a receiver, the more the posterior will resemble their actual career production; Flowers' expected yard-to-touchdown ratio is nearly 200 yards lower than his actual ratio to date, while London's only falls by 20 yards. (If we did the same to Davante Adams, it would only move his expected ratio by 3 yards.)

How did we decide to add 1200 yards worth of league-average production? Why not 600 or 2000? No particular reason; using more league average performance is equivalent to having a stronger prior and will result in a posterior estimate that's closer to league average. I picked 1200 yards because it's somewhere between 1-2 years worth of expected production, but there's no magic to it. (There is a way to find the "optimal" amount of data to add, but it's math-intensive and largely defeats the purpose of a quick hack.)

You can use this trick for all kinds of data. I project punt and kickoff returns for Footballguys and use Bayesian averages to estimate players' "true" return averages. It's especially useful when comparing players with different sample sizes.

Develop and Trust Your Intuitions

At the end of the day, these tools are just that-- tools. Bryce Young is not Troy Aikman, and he's not David Carr; his chances of turning things around wouldn't be better today if Carr had developed into a franchise quarterback for the Texans, nor would they be worse if Aikman had gotten shell-shocked going 0-11 for the 1989 Cowboys. Base rates are only useful because they're useful, not because they're right.

Similarly, Jordan Addison's "true" touchdown rate isn't one for every 76 yards, but it's probably not one for every 120 yards, either. His true touchdown rate certainly doesn't depend on how much league-average performance I decide is appropriate to add.

We can never know a player's true underlying performance level. All we can do is make our best guess, and if our best guess is better than our leaguemates', we'll profit over time.

Thinking in terms of Bayesian inference-- taking a prior and updating for new evidence based on the strength of the prior and the strength of the evidence-- is merely a useful mindset for generating better guesses.

Photos provided by Imagn Images

Fantasy Draft

Season-Long Projections

Fantasy In-Season

Weekly Projections

Fantasy Draft

Season-Long Projections

Fantasy In-Season

Weekly Projections

Dak's Ceiling, Purdy's Payday, Henry's Reward, Achane's Ambition, and Schedule Fallout: The Fantasy Notebook

All 32 NFL Schedule Release Videos Ranked (from worst to best)

More Content

Dak's Ceiling, Purdy's Payday, Henry's Reward, Achane's Ambition, and Schedule Fallout: The Fantasy Notebook

All 32 NFL Schedule Release Videos Ranked (from worst to best)

More Content

Fantasy Draft

Fantasy In-Season

Daily Fantasy (DFS)

Statistics

Fantasy Draft

Fantasy In-Season

Daily Fantasy (DFS)

Statistics

Community

Contests

Useful Links

Community

Contests

Useful Links

Regression Alert: Week 10

The Scorecard

Let's Compare This Column To A Medical Condition...

Continue reading this content with a PRO subscription.

All About That Base (Rate)

A Quick Hack To Estimate Means

Develop and Trust Your Intuitions

Share This Article

Featured Articles

More by Adam Harstad

Latest Features