Reliability engineering for a data-sensitive platform
SLI/SLO practice and hardening that held 99.99% uptime on a platform handling sensitive personal records.
Who it was for: A healthcare SaaS platform handling sensitive personal data.
The problem. A platform holding health records can't treat availability and confidentiality as separate problems — an outage and a leak are both incidents, and both are reportable.
What I did. Defined SLIs and SLOs so reliability had a number attached rather than a feeling. Hardened networking and access, and enforced encryption in transit and at rest across the platform. Ran vulnerability scanning continuously rather than at release checkpoints.
The result. 99.99% uptime sustained, with protective controls around personal data documented and testable.
Like this project
Posted Aug 6, 2026
Reliability engineering for a data-sensitive platform
SLI/SLO practice and hardening that held 99.99% uptime on a platform handling sensitive personal record...