Small, working tools and step-through explainers for reliability engineering. The tools cover reliability targets, telling real outages from blips, and the failures that keep coming back. The explainers cover how the systems actually work, from Kubernetes and Linux to delivery, observability, security and incident response, and what breaks once you run them at scale.