Posts
All the articles I've posted.
-
Fraud Detection at Sub-200ms: The Latency Budget Nobody Talks About
Real-time fraud scoring that lives inside the payment authorization path has a tiny latency budget, and a slightly worse model that fits it beats a better one that does not.
-
A Forecasting Ensemble That Actually Ships
A demand-forecasting ensemble (a classical statistical model, a sequence model, and gradient boosting) that took accuracy far enough to cut inventory hard, plus the boring data problems that mattered more than the model.
-
Llama 2 Is Here. Should You Self-Host?
The week Llama 2 dropped, half my inbox asked whether to pull inference in-house. The break-even math, the GPU scarcity, and the on-call tax nobody puts in the spreadsheet.
-
Platform Engineering: Paving Roads vs Building Cages
An internal ML platform that dozens of teams actually used, and the one test that told me whether I was paving a road or building a cage around it.