Practical write-ups on the failure modes we see most often in scientific ML codebases, and how we address them.