Reproducibility 3 min read
Cold-start evaluation: identities before architectures
Why identity-disjoint splits, paired deltas, and an explicitly incomplete evidence grid matter more than an impressive model diagram.
Read noteShort notes on methods, evaluation, reproducibility, and engineering. Dates and references are explicit; claims tied to a paper return to its source record.
10 of 10 notes
10 research notes
Subscribe via RSSReproducibility 3 min read
Why identity-disjoint splits, paired deltas, and an explicitly incomplete evidence grid matter more than an impressive model diagram.
Read noteReproducibility 9 min read
Patient leakage, imbalance, uncertainty, and other decisions that matter before selecting a model.
Read noteInterpretable ML 7 min read
A practical distinction between explaining a fitted black box and designing a model whose structure is inspectable.
Read noteMethods 8 min read
A compact explanation of kernels, inductive bias, and why small-data settings still reward careful similarity design.
Read noteRisk-aware learning 7 min read
A methods note on threshold risk, expected tail loss, and why evaluation assumptions matter.
Read noteMethods 8 min read
Architecture differences matter less than leakage control, baselines, horizons, and repeated evaluation.
Read noteInterpretable ML 3 min read
How structured constraints can keep a clinical-risk model inspectable, plus the validation still needed before deployment.
Read noteBiomedical AI 3 min read
A research note on similarity geometry, patient-wise evaluation, and what the QADK study does—and does not—show.
Read noteRisk-aware learning 3 min read
Why sequential decision systems need explicit downside objectives and skeptical backtesting.
Read noteResearch engineering 8 min read
A current, framework-neutral setup checklist built around isolated environments, locked dependencies, smoke tests, and official install selectors.
Read noteTry a broader search or a different topic.