The Predictability
Paradox
Every NFL offense telegraphs its run–pass tendencies to some degree, and some are far more readable than others. Conventional wisdom says the predictable ones pay for it. They don’t — and the number that actually separates good offenses from bad is a different one entirely.
How predictable is measured
Take every early-down snap in a competitive game and ask what the situation alone — down, distance, field position, score, clock — says a team should do. A legible offense does exactly that: its pass rate climbs steeply with the situation. An unpredictable one stays flat — running when you’d pass, passing when you’d run.
Kliff Kingsbury’s 2022 Cardinals are the most legible offense of the last decade — a near-perfect diagonal. Kyle Shanahan’s 2016 Falcons, the best offense in football that year, are nearly flat. Same league, opposite legibility.
Legibility is the steepness of this line, scored as the AUC of the league tendency model against a team’s calls. It is not the same as how often a team passes: ARI 2022 and ATL 2016 threw at similar rates overall — they differ in when.
| Situation | ≤44% | 44-48% | 48-52% | 52-65% | >65% |
|---|---|---|---|---|---|
| League — “what the situation says” | 34% | 45% | 49% | 56% | 75% |
| ATL 2016 — unpredictable (a top offense) | 55% | 44% | 57% | 58% | 76% |
| ARI 2022 — most legible | 22% | 36% | 51% | 75% | 87% |
The paradox: legible ≠ beaten
Here is every team-season on two axes of identity: how legible it is (left–right) and how aggressive it is — pass rate over expected, or PROE (bottom–top). Each dot is colored by how good the offense was. If predictability were the tax everyone assumes, the red would pool on the left. It doesn’t — what faint order the color carries runs top-to-bottom, with aggressiveness, not left-to-right.
Legibility and success are statistically independent (r = +.06, its 95% interval straddles zero). PROE and success are not (r = +.25). And the two axes don’t move together (r = +.04) — they are genuinely separate dimensions of who an offense is. Single dots carry ~±.03 of sampling noise; read the field, not the point.
| Team-season | Legibility | PROE | Offensive EPA |
|---|---|---|---|
| ARI 2022 | 0.78 | +1.7 | −.067 |
| TB 2019 | 0.77 | +0.6 | −.091 |
| WAS 2024 | 0.76 | +2.3 | +.011 |
| TB 2020 | 0.76 | +1.9 | +.087 |
| TB 2016 | 0.76 | −2.3 | −.011 |
| ARI 2021 | 0.75 | +5.3 | +.019 |
| DET 2024 | 0.55 | −5.1 | +.118 |
| WAS 2023 | 0.55 | +10.0 | −.069 |
| ARI 2023 | 0.55 | −7.0 | +.011 |
| CHI 2025 | 0.54 | −1.0 | +.078 |
| CHI 2022 | 0.54 | −11.5 | −.032 |
| SF 2016 | 0.51 | −6.1 | −.091 |
What actually gives you away
If tendencies don’t cost you, what does a defense read? Not the things fans obsess over. Train a model on the pre-snap picture and only one family of cues moves the needle: alignment — are you in the gun, how many backs, pistol or under center. Motion and personnel grouping add essentially nothing.
Pre-snap alignment adds +.10 AUC over the situation alone, and about +.05 of that is beyond simply being in the shotgun. Motion and personnel — the classic “tells” — test as noise once alignment is known.
An empty backfield is the loudest tell in football.
The gun screams pass; under center and pistol both lean run — pistol most of all.
Every offense, decade by decade
Legibility is a coaching fingerprint — it moves with coordinators, not rosters. Pick a team to trace its predictability and aggressiveness across the last ten seasons, and see the one alignment that gives it away.
| Season | Legibility | PROE | EPA |
|---|---|---|---|
| 2016 | 0.73 | −1.5 | +.041 |
| 2017 | 0.72 | −4.1 | −.073 |
| 2018 | 0.64 | −7.6 | −.180 |
| 2019 | 0.61 | +4.3 | +.027 |
| 2020 | 0.73 | −0.0 | +.002 |
| 2021 | 0.75 | +5.3 | +.019 |
| 2022 | 0.78 | +1.7 | −.067 |
| 2023 | 0.55 | −7.0 | +.011 |
| 2024 | 0.64 | +2.4 | +.109 |
| 2025 | 0.67 | +8.1 | −.027 |
Biggest tell: 0 (empty) → 95% pass.
Pass rate by backfield count, 2022–2025.