Research Library
Every findings document in the repo, rendered. Sourced directly from the
markdown under backtests/ and docs/ — nothing to
register, a new file shows up here automatically. Negative results are kept
deliberately: most of what's here is a hypothesis that didn't survive.
38 documents are private and not shown.
All areas
backtests
cfb_eval
docs
fantasy_eval
hr_derby
kelly_sim
ncaab_eval
nfl_context_signals
nfl_edges
nfl_prop_usage
nhl_eval
pickem
preseason_signal
qb_weather
split_miner
stacking_eval
strategies
tennis_eval
weather_ftn_value
stacking_eval · 4
Tennis consensus-stacking evaluation — REJECT (2026-07-04)
rejected
A full historical backtest was possible: [redacted]::tennis_matches (81,622 matches 2010-2026 with embedded closing odds: market average, Pinnacle, best price) + tennis_match_elo (point-in-time pre-match Elo, through 2024-11-17). Joined usable set: 69,090 …
backtests/stacking_eval/tennis/REPORT.md
NCAAB consensus-stacking evaluation — REJECT (2026-07-04)
rejected
The best-prior candidate (soft small-conference markets), tested on our largest capture. Verdict: consensus wins everything.
backtests/stacking_eval/ncaab/REPORT.md
MLB consensus-stacking pilot — INSUFFICIENT DATA, leaning REJECT (2026-07-04)
rejected
Question: does blending the pitcher-adjusted Elo v2 prob into the devigged closing consensus (NHL v3 recipe) beat consensus-only?
backtests/stacking_eval/mlb/REPORT.md
NFL consensus-stacking evaluation — REJECT (2026-07-04)
rejected
Question: does the NHL v3 "top-down" recipe (tiny logistic on devigged closing consensus + small orthogonal features) beat the NFL closing moneyline?
backtests/stacking_eval/nfl/REPORT.md