Small, working tools and step-through explainers for reliability engineering. The tools cover reliability targets, telling real outages from blips, and the failures that keep coming back. The explainers cover how the systems actually work, from Kubernetes and Linux to infrastructure as code, delivery, observability, security and incident response, and what breaks once you run them at scale.