Game quality

What each result was worth beyond the score. Every play is graded by EPA (expected points added), then adjusted for who it came against: a defence's rating is what it allowed to everyone else that season, so a big day against a bad defence counts for less. Season views include playoff games.

Tim Boyle — 2023

Adj EPA / dropback
-0.389
Raw EPA / dropback
-0.360
Dropbacks
86
Turnover EPA
-16.9
WPA
-0.56
WeekOpp DropbacksAdj EPA/dbRaw EPA/dbScheduleTotal adj EPA INT / FLTurnover EPASacksSack EPAWPACPOE
11BUF 15 -0.483 -0.510 +0.026 -7.2 1 / 0 -2.0 1 -1.8 -0.01 -19.0
12MIA 45 -0.451 -0.415 -0.036 -20.3 2 / 0 -11.6 7 -11.4 -0.44 +0.7
13ATL 26 -0.226 -0.180 -0.046 -5.9 1 / 0 -3.3 1 -1.7 -0.11 -7.3
Opponent adjustment. A unit's rating for a game is its average over its other games that season, blended with 16 games' worth of prior (league average plus half of last season's gap from it; 24 games for pass defence, which is noisier). Early in a season that leans on last year; by midseason it is mostly this year. Chosen by how well it predicts a held-out game: it cut that error by about a quarter against no shrinkage, and adjusted numbers agree between odd and even weeks more than raw ones do (team EPA margin r 0.52 → 0.55, defence 0.32 → 0.36, QB EPA/dropback 0.47 → 0.49, 2016–2025).
Win expectancy is a logistic fit of wins on EPA margin over 6,036 completed games (it names the winner 86.5% of the time). These describe what happened; one game is not a forecast. Play-by-play credits the quarterback with every sack and dropback, so a bad line shows up in his numbers. Seasons 2016–2026, built 2026-09-14 01:25 UTC, from nflverse. Write-up: backtests/nfl_game_quality/FINDINGS.md.