Playground
Demos a reviewer can poke at.
These read from committed project outputs and a Supabase-backed sample of real shots. No demo here renders an invented number: every value traces to a committed result or a real model output. Each one sits on its project page, next to the write-up that explains it.
Model scorecard
Every model across the portfolio, one card each. Filter by domain. The benchmark is the point: each card sits its value beside the incumbent, and a metric with nothing to beat is greyed.
Opponent-Adjusted Football Metrics
LiveCxG diagnostic (sklearn logistic/Ridge/GBM)
Contextual expected goals
Behind the benchmark.
Opponent-Adjusted Football Metrics
LiveCxG diagnostic
Calibration
No incumbent to beat
Opponent-Adjusted Football Metrics
LiveCxA baseline
Contextual expected assists
No incumbent to beat
Opponent-Adjusted Football Metrics
LiveCxA diagnostic v1 (sklearn GradientBoostingClassifier, sigmoid-calibrated)
Contextual expected assists
Ahead of the benchmark.
Opponent-Adjusted Football Metrics
LiveCxA diagnostic v1
Precision among the model’s most confident actions
Ahead of the benchmark.
Contextual Football Metrics
LiveContextual GLM
Contextual expected goals
Ahead of the benchmark.
Contextual Football Metrics
LiveContextual GLM
Cross-validation
No incumbent to beat
Frame2Threat
LiveXGBoost + GRU ensemble
Possession-danger prediction
No incumbent to beat
Frame2Threat
LivePossessionGRU
Possession-danger prediction
No incumbent to beat
Frame2Threat
LiveXGBoost (event+360)
Pass-level line-breaking
No incumbent to beat
Retail Growth Intelligence
LiveTwo-stage X-learner (LightGBM stage 2)
Uplift / campaign targeting
No incumbent to beat
Retail Growth Intelligence
LiveTwo-stage X-learner
Uplift / campaign targeting
Ahead of the benchmark.
Retail Growth Intelligence
LiveTwo-stage X-learner
Uplift ranking quality
No incumbent to beat
Retail Growth Intelligence
LiveChurn classifier
Churn classification
Ahead of the benchmark.
HealthBeauty360 Retail Analytics
LiveChurn classifier
Churn classification
Ahead of the benchmark.
Multi-Horizon Demand Forecasting
ProfessionalDemand forecaster
Quarterly SKU demand (~7,000 SKUs)
vs Previous approach
Residual-Load Forecasting
ProfessionalResidual-load framework
Forward price-curve accuracy (far seasons)
vs Previous methodology
Residual-Load Forecasting
ProfessionalResidual-load framework
Christmas-period demand RMSE
vs Pre-fix
Network-Charge Forecasting
ProfessionalDUoS/TNUoS forecasting engine
Network-charge forecast accuracy
vs Prior internal approach
Attendance Forecasting
ProfessionalGridSearchCV Random Forest
Cardio attendance forecast (out-of-sample)
Ahead of the benchmark (lower is better).
Attendance Forecasting
ProfessionalGridSearchCV Random Forest
Holistic attendance forecast (out-of-sample)
Ahead of the benchmark (lower is better).
Every value traces to a committed result. A card with no incumbent is greyed: a metric with nothing to beat is not a win. MAE is lower-better and reported as a count, not a percentage.
Data provenance. Hand-typed from committed results (content/metrics.ts). Football and retail are open or synthetic data; energy and consulting are professional results.