CAPABILITIES
& LIMITATIONS
What the model does, what it doesn't, and the data behind both. This page is generated from the same backtest artifacts as the weekly picks.
← BACK TO THIS WEEK'S POLL
MODEL OVERVIEW
A logistic regression over 8 named signals, trained walk-forward on FBS-vs-FBS games since 2015 (CollegeFootballData). Predictions lock at publication. Every probability decomposes into signal contributions that sum to a points-style lean — the breakdown shown on each pick is the model's actual computation.
home_field rating_gap rest_diff opp_pass_def_trend rush_def_trend pass_off_trend lookahead_spot travel_km
HISTORICAL PERFORMANCE (WALK-FORWARD)
Each season was predicted using only data from prior seasons.
SEASONGAMESSTRAIGHT-UPVS THE LINEATS %BRIERML ROI
2016 76070.3%358–382–2048.4%.190 no ML data
2017 77670.4%390–365–2151.7%.185 no ML data
2018 77272.4%391–361–2052.0%.175 no ML data
2019 77473.1%377–384–1349.5%.176 no ML data
2020 (COVID) 53471.2%274–251–952.2%.187 no ML data
2021 77069.5%383–378–950.3%.187 -4.0%
2022 77668.7%389–374–1351.0%.199 -4.6%
2023 79272.3%395–381–1650.9%.185 -1.2%
2024 79869.2%391–388–1950.2%.193 -5.4%
2025 80872.5%396–394–1850.1%.184 +1.0%
CALIBRATION
Career held-out calibration, 2021 onward: actual win rate per predicted-probability bucket.
40–50%
47.1% n=444 50–60%
52.6% n=521 59–69%
65.0% n=549 70–80%
72.4% n=557 80–90%
85.6% n=536 90–100%
93.0% n=359
Every reliable bucket within ±7 points of its midpoint. Career Brier .190 vs .191 for the Elo-only baseline.
MONEYLINE ANALYSIS
Flat $1 stakes at posted moneylines, 2021-2025 (3,771 games). Historical books: Bovada/DraftKings/ESPN Bet (used as a proxy for FanDuel pricing).
STRATEGYBETSWINSFLAT ROI
bet every straight-up pick3,7712,619 -2.81%
any modeled edge1,537895 -0.62%
edge > 2 points1,278709 -0.35%
edge > 5 points968514 +2.91%
robot backs the favorite3,3012,415 -4.83%
robot backs an underdog470204 +11.38%
Note: the underdog subset (470 bets) is a small sample identified among six tested strategies. It is tracked as a paper-trading experiment pending live confirmation.
SIGNAL EVALUATION RECORD
Every candidate signal, its verdict, and the evidence. Acceptance requires: improves held-out Brier, stable coefficient sign, and improvement holds post-2021.
SIGNALVERDICTEVIDENCE
talent_gap AUDITIONING passes 2/3 criteria — needs more seasons of evidence
returning_production AUDITIONING passes 2/3 criteria — needs more seasons of evidence
rush_def_trend GRADUATED improves held-out Brier (mean +0.00003), stable sign, holds 2021+
pass_off_trend GRADUATED improves held-out Brier (mean +0.00003), stable sign, holds 2021+
rush_off_trend CUT no lift on held-out seasons (mean Brier delta -0.00003)
turnover_luck CUT no lift on held-out seasons (mean Brier delta -0.00002)
letdown_spot CUT no lift on held-out seasons (mean Brier delta -0.00014)
lookahead_spot GRADUATED improves held-out Brier (mean +0.00020), stable sign, holds 2021+
travel_km GRADUATED improves held-out Brier (mean +0.00020), stable sign, holds 2021+
body_clock CUT no lift on held-out seasons (mean Brier delta -0.00014)
line_movement ADVISORY evaluated but cannot graduate on backtest evidence — open→close movement is not observable at Tuesday lock time
KNOWN LIMITATIONS

No injury or roster awareness. The model uses no injury reports, practice news, or depth charts. Ratings adjust to a missing starter only after games are played, while betting lines typically adjust immediately. Mid-week line movement of 3+ points against a locked pick is flagged 'steamed' in the ledger so the cost of this limitation can be measured.

Does not beat the closing spread. Career 49.5% against the closing line; break-even at standard odds is 52.4%. This is why the betting analysis runs in paper mode.

Cover probabilities are unvalidated. Win probabilities are calibrated (see above). Cover probabilities on the bet card are derived through a margin model that has not been validated to the same standard; the paper ledger exists to test them.

One signal carries most of the predictive weight. The rating-gap (Elo) signal accounts for most of the model's accuracy. The other signals passed evaluation but contribute marginally to career Brier score.

Positive subsets rest on small samples. The positive underdog moneyline ROI comes from 470 bets across five seasons and was identified while testing six strategies. It is treated as unconfirmed until live paper results agree.

Historical odds are a proxy. Backtests price against Bovada/DraftKings/consensus lines (CFBD), not FanDuel directly. Live pricing uses FanDuel via The Odds API.

OPERATING RULES

• Predictions lock at publication and are not edited. The scorecard grades the locked record.

• The model does not use in-week news. Line movement ≥3 points against a locked bet is flagged 'steamed', graded normally, and reported separately.

• If the available line improves on a locked bet's number by ≥1 point, a paper top-up may be recorded at the better line. Bets are never added after adverse movement.

• Bets are restricted to games with |spread| ≤ 21. The margin model underestimates large favorites' margins, which can produce spurious large-underdog edges; the restriction stands until live paper results show the mapping holds in the tails.

• Live betting mode requires positive held-out backtest ROI and positive live paper closing-line value over at least half a season, followed by manual confirmation.

• 2020 is reported separately and excluded from headline metrics due to COVID-shortened schedules.

ROBOT POLL · this page regenerates from backtest data with every deploy